Recent Articles - Page 37

Latest News

View All News →

Guides

View All →

Reviews

View All →
Kimi K3 Review: Best at Code, Worse at Honesty

Kimi K3 Review: Best at Code, Worse at Honesty

Moonshot's Kimi K3 tops LMArena's Frontend Code Arena and undercuts Opus 4.8 on cost per task, but a tripled price tag, a rising hallucination rate, and an unresolved distillation question complicate the win.

Leaderboards

View All →

Models

View All →
Qwen3-VL-235B-A22B

Qwen3-VL-235B-A22B

Alibaba's flagship open-weight vision-language MoE beats every proprietary model on DocVQA at 96.5% and MathVista at 85.8%, but trails GPT-5.4 and Gemini 3.1 Pro on broad MMMU-Pro reasoning.

DeepSeek-VL2

DeepSeek-VL2

DeepSeek-VL2 is DeepSeek's open-weight Mixture-of-Experts vision-language model, activating just 4.5B of its 27B parameters to hit 93.3% on DocVQA and beat GPT-4o on OCRBench.

Qwen2.5-VL-72B-Instruct

Qwen2.5-VL-72B-Instruct

Alibaba's dense 72B vision-language model tops the open-weight DocVQA leaderboard at 96.4% and remains the default self-hosted choice for document and chart understanding.

Recent

Best AI Coding Tools with Real Free Tiers in 2026

Best AI Coding Tools with Real Free Tiers in 2026

Which AI coding assistants offer genuinely usable free tiers in 2026 - comparing Windsurf, GitHub Copilot, Cursor, Antigravity CLI, and Continue.dev on limits, features, and where each one runs out.

Qualcomm Hits Record High as AI Device Bet Pays Off

Qualcomm Hits Record High as AI Device Bet Pays Off

Qualcomm stock surged 12% to an all-time high on May 22 and is now up 75% in a month, driven by an OpenAI smartphone deal, a Stellantis car contract, and a data center chip push - forcing Wall Street to rethink a company it had written off.