Recent Articles - Page 16

Latest News

Anthropic's $1.5B Book Piracy Settlement Wins Approval

Anthropic's $1.5B Book Piracy Settlement Wins Approval

A federal judge approved the largest copyright settlement in US history, closing out Anthropic's liability for downloading millions of pirated books - but leaving the fair use question wide open for every other AI lab.

View All News →

Guides

View All →

Reviews

View All →
Kimi K3 Review: Best at Code, Worse at Honesty

Kimi K3 Review: Best at Code, Worse at Honesty

Moonshot's Kimi K3 tops LMArena's Frontend Code Arena and undercuts Opus 4.8 on cost per task, but a tripled price tag, a rising hallucination rate, and an unresolved distillation question complicate the win.

Leaderboards

View All →

Models

View All →
Qwen3-VL-235B-A22B

Qwen3-VL-235B-A22B

Alibaba's flagship open-weight vision-language MoE beats every proprietary model on DocVQA at 96.5% and MathVista at 85.8%, but trails GPT-5.4 and Gemini 3.1 Pro on broad MMMU-Pro reasoning.

DeepSeek-VL2

DeepSeek-VL2

DeepSeek-VL2 is DeepSeek's open-weight Mixture-of-Experts vision-language model, activating just 4.5B of its 27B parameters to hit 93.3% on DocVQA and beat GPT-4o on OCRBench.

Qwen2.5-VL-72B-Instruct

Qwen2.5-VL-72B-Instruct

Alibaba's dense 72B vision-language model tops the open-weight DocVQA leaderboard at 96.4% and remains the default self-hosted choice for document and chart understanding.

Recent

Embedding Models Pricing - June 2026

Embedding Models Pricing - June 2026

Embedding API cost comparison: voyage-4-lite, OpenAI 3-small, Jina v3, and Amazon Titan V2 tie at $0.02/MTok. Gemini Embedding 2 now GA, Cohere Embed 4 dimensions corrected to 1,536 default.

Meta Restricts Claude Code Over Training Data Leakage

Meta Restricts Claude Code Over Training Data Leakage

Meta has restricted engineers from using Claude Code and Codex, citing training data distillation risk. The policy change exposes a structural mismatch between how AI coding tools work and what enterprise AI labs can tolerate.

Ford Rehires 350 Engineers After AI Quality Fails

Ford Rehires 350 Engineers After AI Quality Fails

Ford climbed from No. 15 to No. 1 in JD Power's 2026 quality study - not by deploying more AI, but by admitting it had over-relied on automation and bringing back 350 veteran engineers to fix what the machines got wrong.

Apple's Vision Pro Chief Joins OpenAI Hardware Unit

Apple's Vision Pro Chief Joins OpenAI Hardware Unit

Paul Meade, Apple's VP for the Vision Pro and smart glasses, is joining OpenAI's io hardware division - the most senior hardware engineering defection yet in a two-year talent war that is reshaping who builds the next computing platform.

Gemini 3.5 Pro

Gemini 3.5 Pro

Google DeepMind's upcoming flagship model with a 2M-token context window and Deep Think reasoning, announced at Google I/O 2026 and expected in July.

Claude Mythos 5

Claude Mythos 5

Claude Mythos 5 is the full release of Anthropic's restricted Mythos family - same weights as Fable 5 but without safety classifiers for cybersecurity and biology, at $10/M input and $50/M output tokens.