
ZML Ships a Free LLM Server That Runs on Any Chip
ZML's LLMD inference server runs Llama, Qwen, and Mistral on Nvidia, AMD, TPU, Apple Metal, and Intel Arc from a single binary - for free.
They summarize our coverage. We write it.
Newsletters like this one rebroadcast our headlines - often without the full review, the source reading, or the analysis underneath. Our weekly briefing sends the work they paraphrase, straight from the desk, before they get to it.
Free, weekly, no spam. One email every Tuesday. Unsubscribe anytime.

ZML's LLMD inference server runs Llama, Qwen, and Mistral on Nvidia, AMD, TPU, Apple Metal, and Intel Arc from a single binary - for free.

Z.ai's free agentic IDE ships Goal Mode, multi-agent coordination, and pricing up to 82% cheaper than Claude Code - with a concrete China data law risk that teams need to weigh.

Sysdig documents the first AI-agent ransomware operation: an LLM exploited CVE-2025-3248 in Langflow, moved laterally, and encrypted 1,342 production database records with no human directing each step.

Vercel shipped Eve, an open-source TypeScript agent framework, at Ship London in June. The filesystem-first design and isolated Sandbox runtime signal a serious play for the production agent layer.

Meituan's 1.6T open-source coding model secretly topped OpenRouter for two months before revealing itself - and the price-to-performance math is hard to argue with.

Meituan's 1.6T-parameter open-source MoE coding model, trained end-to-end on 50,000 domestic Chinese ASICs, with native 1M token context and a 59.5 SWE-bench Pro score.

Meituan open-sources LongCat-2.0, a 1.6T MoE model trained on 50,000 Chinese ASICs that secretly topped OpenRouter under the alias Owl Alpha.

H Company's open-weight sparse MoE vision-language model purpose-built for desktop computer use, scoring 82.6% on OSWorld-Verified with only 3B active parameters.

Compare five leading AI developer SDKs - Vercel AI SDK 7, LangChain, LlamaIndex, Mastra, and PydanticAI - and find the right framework for your next AI-powered app.

ClinePass gives developers access to ten curated open-weight coding models for $9.99 a month, betting the agent harness matters more than the model.

Z.ai's GLM-5.2 delivers frontier coding performance with open weights and MIT license at roughly one-sixth the cost of GPT-5.5 - but can it replace Claude Opus 4.8?

Cohere's first developer-focused model - 30B sparse MoE with 3B active parameters, free Apache 2.0 license, 256K context window, and 33.4 on the AA Coding Index.