Top story
Apertus, Open Foundation Model for Sovereign AI. Swiss-led Apertus releases as a fully open multilingual foundation model, joining OLMo 3.1, K2 Think V2, and Nemotron in the sovereign-AI open-model landscape. Source
Research
WorldLines: Benchmarking and Modeling Long-Horizon Stateful Embodied Agents. WorldLines benchmark and modeling approach for long-horizon stateful embodied agents. Source
GateMem: Benchmarking Memory Governance in Multi-Principal Shared-Memory Agents. GateMem benchmark evaluates memory sharing and governance in multi-principal agent systems. Source
PerceptionDLM: Parallel Region Perception with Multimodal Diffusion Language Models. PerceptionDLM extends diffusion language models to multimodal parallel region perception tasks. Source
Distilling Examples into Task Instructions: Enhanced In-Context Learning for Real-World B2B Conversations. Method for converting examples into task instructions to improve in-context learning for B2B dialog. Source
Tools
RRT-355M, Softmax-Free Attention Model. From-scratch, softmax-free attention model at GPT-2 Medium scale with sparse tile-skipping Triton kernels, open weights, and a custom inference engine. Source
Recall, Local Project Memory for Claude Code. Local memory layer for Claude Code that persists project context across short coding sessions. Source
Temporary Cloudflare Accounts for AI Agents. Cloudflare introduces temporary accounts scoped for AI agents. Source
browser-harness, Self-Healing Browser Automation. Self-healing browser automation harness designed to make LLM-driven web agents more robust. Source
Industry
Identity Verification on Claude. Anthropic introduces identity verification on Claude access; discussion highlights geopolitical fallout for international users and competitors. Source
There Is Minimal Downside to Switching to Open Models. Practitioner thread arguing open-weight models are already viable replacements for proprietary APIs, with routing-config examples and counterpoints on privacy and capability gaps. Source
When the Trump Administration Cracks Down on Anthropic, Who Benefits?. Analysis of who benefits if the Trump administration targets Anthropic. Source
Beyond Siri: Practical AI Features Coming to Your iPhone in iOS 27. Preview of practical AI features expected in iOS 27 beyond Siri. Source
Community
Fine-Tuning a Local LLM (Qwen 3:0.6B) to Categorize Questions. Show-and-tell of fine-tuning a small Qwen model for question classification, with community pointing out simpler classical ML often suffices. Source
Best Current Methods for Finetuning Whisper on Domain-Specific Vocabulary?. User asks for current best practices to fine-tune Whisper on domain-specific Spanish vocabulary and how much labeled audio is needed. Source
Data-Centric Debugging for Teams Training Neural Nets. Open-source WeightsLab revamp for mid-training data debugging in PyTorch CV pipelines (images, video, LiDAR). Source