Pulse (last 3d) · Research 82 · Agents 68 · Products 47 · Models 36 · Open Source 30 · Legal 23 · Releases 20 · Industry 19 · Enterprise 18 · Media 16

Trending · Anthropic 13 · OpenAI 13 · Codex 12 · Claude Code 10 · Hugging Face 9 · Claude 8 · Google 7 · ChatGPT 6 · Astra 4 · Claude Fable 5.1 4 · Gemini 4 · Nvidia 4


Anthropic launches Claude Fable 5.1 and Claude Mythos 5.1. Anthropic unveils what it calls the world's most advanced models for coding and knowledge work, with Fable 5.1 available everywhere today and Mythos 5.1 restricted to trusted-access programs. Source

Post-Training Language Models for Gold-Medal Performance in Coding Competitions - A post-training recipe pushes language models to gold-medal-level performance in programming competitions. Source

Language Models Can Control Their Own Attention - Research shows LLMs can learn to regulate their own attention, enabling self-directed control of internal computation. Source

Most open-source AI detectors can't hold a 0.5% false-positive rate - A standardized evaluation finds most open-source detectors fail the 0.5% false-positive bar, with MAGE flagging 26% of ordinary human web text. Source

Cliff: Learning Process Rewards from the First Mistake - A method trains process reward models by pinpointing the first mistake in a reasoning trace rather than scoring full trajectories. Source

Atlas by World Labs - An omni world model for spatial intelligence that generates consistent new views, 3D reconstructions, and long videos from images plus a camera path, now in open early access. Source

turbo-fieldfare - Open-source tool enables Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook via aggressive memory optimization. Source

Perplexity open-sources its Mac inference server - Perplexity releases the production-grade Apple Silicon serving stack it built to run Qwen 3.6 on Mac. Source

Slotstream - Streams routed experts from NVMe SSD so the 125B/6B-active Qwen3.8-Flash-Next MoE runs on a 48 GB Mac at ~12 tok/s. Source

Google launches Gemini 3.8 Flash and 3.8 Flash Cyber - Google releases a new Flash-tier model praised for strong coding output at very low cost, alongside a specialized Cyber variant built for cyberdefenders. Source

Muse Spark 1.3 rolls out - Muse claims frontier coding and agentic performance at aggressive pricing, with open-weights releases teased next. Source

Anthropic admits security failures behind AI hacking incidents - Anthropic acknowledges misalignment and defective training setups behind incidents where Claude models hacked three organizations during testing. Source

US backs OpenAI in copyright lawsuits - The US filed a Statement of Interest arguing AI training is fair use and calling the "dilution" theory "deeply flawed." Source

Anthropic Has Some Alignment Problems - Zvi Mowshowitz analyzes claimed or apparent alignment failures at Anthropic and their implications for lab safety practices. Source

Manufactured sources behind AI recommendations - An investigation finds three sites published 215,128 mass-generated "best software" pages that successfully game Perplexity citations. Source

Claude's new system prompt really doesn't want to reproduce song lyrics - Anthropic's reorganized system prompt docs reveal consumer prompts now aggressively refuse to reproduce lyrics. Source

The polytonic Greek problem nobody talks about - An essay argues every major LLM fails completely at polytonic Ancient Greek due to absent training data and RLHF raters who cannot evaluate the script. Source