
A 30B Model Outscored a 550B One Inside Nvidia's Own Factory Planning
Nvidia and Palantir published an 86.7% versus 55.5% accuracy gap favoring a 30B fine-tuned model over a 550B general one on Nvidia's own materials allocation.

Nvidia and Palantir published an 86.7% versus 55.5% accuracy gap favoring a 30B fine-tuned model over a 550B general one on Nvidia's own materials allocation.

Moonshot AI told investors it is targeting $2 billion in annualized revenue by year-end after Kimi K3 pushed its run rate past $1 billion in August.

Cognition says SWE-2 scores 50.0% on FrontierCode 1.1 Main, a point behind Fable 5.1, at 64% lower cost — built on an open Chinese base model.

A pinned endpoint identifier is about to serve different weights with no opt-out, and engineers say that breaks the change-management contract they rely on.

IBM's Granite 4.2 ships 3B, 8B and 30B dense reasoning models under Apache 2.0 with a 512K context window and an agentic RL stage for the larger two.

Artificial Analysis scored Z.ai's GLM-5.3 at 60 on its Intelligence Index, well above the 35 median, at $4.40 per million output tokens. The catch is verbosity.

Alibaba's Qwen3.8 open weights pair a 2.4T mixture-of-experts model with a 27B dense multimodal model sized for a single 24GB consumer GPU.

NVIDIA released Nemotron 3.5 Lightning, a 30B open MoE model for agentic workloads, alongside NeMo Switchyard, an open source model-routing library.

Alibaba's Qwen team released Qwen3.8 as open weights, pairing a 2.4-trillion-parameter MoE flagship with a compact 27B vision-language model.

Meta's Muse Glimmer is a 30B open-weight agent model under Apache 2.0 that runs on one consumer GPU — its first open release in over a year.