Pulse (last 3d) · Research 59 · Products 52 · Agents 52 · Industry 22 · Open Source 20 · Models 16 · Enterprise 14 · Releases 11 · Inference 10 · Media 10

Trending · Anthropic 12 · Claude 10 · Claude Code 10 · OpenAI 8 · Ox Alpha 6 · Cursor 5 · Gemini 5 · Simon Willison 5 · Codex 4 · TechCrunch 4 · ChatGPT 3 · GLM-5.2 3


Prime Intellect NanoGPT Speedrun Frontier: 153 autonomous runs, 18 models, research taste as the measured variable. Prime Intellect ran 153 fully autonomous agent runs on the nanoGPT optimizer speedrun (8xH200 nodes, up to 8 days per run, 18 frontier models), the first public autonomous-research benchmark at this scale, with a leaderboard and 41 open traces. Source

EnvHarness & EnvRigger. Google Cloud AI ships composable wrapper components that reshape frozen agent environments at the standard reset/step/obs interface, plus EnvRigger, which rigs them automatically from the agent's own failures. Source

AI-designed molecules cleared wet-lab validation three ways this month. Three independent design-to-wet-lab validations landed within ~3 weeks, including a Claude agentic binder campaign hitting 14/15 targets, design is no longer the bottleneck. Source

Llama-Mobile: Efficient 2.7-Bit Quantization of VLMs. A framework quantizing VLMs to 2.7 bits shrinks Llama 3.2 11B Vision to 3.7GB for mobile use while preserving VQA performance. Source

Dual Nature of Generalization in On-Policy Distillation. Paper analyzes where generalization helps and hurts in on-policy distillation of LLMs. Source

0mcp. An all-in-one platform for launching and operating production-ready MCP servers from one dashboard, making your SaaS agent-ready without months of MCP development. Source

Open Artifact. Gives your AI agent's reports, docs, and dashboards a shareable URL with line-by-line commenting and agent revision support; self-hostable. Source

Mustel. A local-first static analysis layer for AI coding agents that compresses Ruff, Bandit, and pip-audit output into a single sub-200-char prompt field agents can act on directly. Source

opencodex. Open-source universal provider proxy letting Codex CLI and Claude Code run against any LLM backend. Source

Anthropic's best AI model struggles to attract users as cheaper tools thrive. Simon Willison's take on FT reporting that Claude underperforms in user adoption despite model quality, as cheaper rivals win. Source

Who's behind the new 'stealth model' Ox Alpha?. TechCrunch investigates the mysterious 'Ox Alpha' model that appeared free on OpenRouter with strong coding benchmarks, with speculation pointing to a stealth Zhipu GLM test. Source

Nvidia Customers Notified About AI-Related Price Hikes Above 15%. Nvidia has begun notifying customers of AI-related price increases exceeding 15%. Source

MCP August roadmap: Tasks head to core, agents get a first-class identity stack. MCP Core Maintainers published a roadmap fixing five priority areas for the next spec cycle, including agentic messaging primitives (Tasks SEP-2663) and agent identity via DPoP. Source

I trained a 1.57B-parameter Dreamer 4 World Model from scratch for under $150. Independent researcher trains a Dreamer 4 world model from scratch on a shoestring budget. Source

I developed my own quantized LLM from scratch, trained on 30B tokens, deploys in 60 MB. Hobbyist builds a quantized LLM with a 30B-token training run and a 60 MB deployment footprint. Source

ConvRot Quant method now in llama-cpp-turboquant. ConvRot quantization lands in llama-cpp-turboquant, with Q6 approaching Q8 accuracy. Source

UK publishers lobby to keep ChatGPT off Google's new search choice screen. UK publishers argue answer engines like ChatGPT and Perplexity send zero referral traffic, unlike link-based rivals, so should be excluded from the DMCCA choice screen. Source