The AI Wire

The signal, filed daily. Every dispatch since the wire opened, newest first.

The curated tier is coming. My picks with the why, the decode & dossier library, the full archive.

6 dispatches
  1. Researchers proposed using LLMs as semantic regularizers for enumerative feature synthesis in decision trees, aiming to filter out nonsensical features while preserving interpretability. Alibaba's Qwen team indicated its next major model will prioritize quality over release speed, signaling a shift in Chinese frontier model development. Community discussion highlighted a claim that ChatGPT 5.2 Pro solved Erdős Problem 281, intensifying debate over AI mathematical reasoning.

  2. ML hiring burnout is escalating as candidates from top labs report months of rejections and mental exhaustion, signaling structural problems in research recruiting. ClickHouse closed a $400 million Series D, acquired Langfuse, and launched a Postgres integration. DeepSeek released Engram, a static memory unit design that avoids recomputing fixed knowledge through transformer layers.

  3. PACEvolve, a new framework for integrating large language models into evolutionary search, was published today, offering structured scaffolding to fix reliability problems in AI-driven optimization. Zhipu AI and Huawei released GLM-Image, China's first state-of-the-art multimodal model trained entirely on domestic Ascend 910 chips, while Black Forest Labs shipped FLUX.2 [klein], a 9B image generator that runs in under a second on a 4090.

  4. Zhipu AI trained its first major model entirely on Huawei's hardware stack, breaking US chip reliance. Mistral published the Ministral 3 series paper detailing 3B, 8B, and 14B parameter models, while NVIDIA released Orchestrator-8B for routing tasks to specialized tools and models.

  5. Researchers released APEX-SWE, a new benchmark measuring whether frontier AI models can execute economically valuable software engineering work rather than just narrow coding tasks. ZAI open-sourced GLM-Image, a hybrid autoregressive and diffusion model matching mainstream image generators while excelling at text rendering. Google also pushed into clinical AI with MedGemma 1.5, adding improved medical image interpretation and speech-to-text capabilities for diagnostic workflows.

  6. Baichuan AI released Baichuan-M3-235B, a 235-billion-parameter medical-enhanced LLM that scales up its specialized healthcare capabilities from the previous 32B model. DeepSeek open-sourced Engram, a conditional memory framework introducing a new sparsity axis for LLM efficiency. Microsoft also released FrogBoss 32B and FrogMini 14B, coding agents fine-tuned on Claude Sonnet 4 debugging trajectories.

Newer10 / 10Older