The AI Wire

The signal, filed daily. Every dispatch since the wire opened, newest first.

The curated tier is coming. My picks with the why, the decode & dossier library, the full archive.

8 dispatches
  1. 82 signals

    Mistral released Medium 3.5, a 128B model with remote agent support, now on Hugging Face. Anthropic is reportedly seeking a valuation that would surpass OpenAI's, while a prompt injection flaw in Ramp's Sheets AI was found leaking sensitive financial data.

  2. 84 signals

    OpenAI models, Codex, and Managed Agents are now available on Amazon Bedrock, marking a major distribution partnership between the two companies. Mistral teased a 128B Medium 3.5 model ahead of launch, while Microsoft released VibeVoice, an open-source voice AI system targeting frontier-level speech capabilities.

  3. 82 signals

    Microsoft and OpenAI have ended their exclusive revenue-sharing partnership, restructuring the commercial relationship that defined the current AI industry. China blocked Meta's acquisition of Manus, and OpenAI secured FedRAMP Moderate authorization for use by U.S. federal agencies.

  4. 64 signals

    An AI agent autonomously deleted a production database, with its own logs documenting the failure. DeepSeek released V4-Pro and V4-Flash, and Moonshot open-sourced Kimi-K2.6, both advancing open-weight performance on agentic and reasoning benchmarks.

  5. 18 signals

    A non-mathematician used ChatGPT to solve a 60-year-old Erdős problem in combinatorics, marking a watershed for AI-assisted mathematical discovery. DeepSeek-V4 launched with Day 0 support on SGLang, showcasing faster inference and verified reinforcement learning. Separately, a pattern of labs withholding model capabilities "too dangerous to release" is drawing renewed scrutiny over transparency.

  6. 40 signals

    Google plans to invest up to $40 billion in Anthropic, deepening its commitment to the AI safety-focused lab. OpenAI released GPT-Image-2 via its API and Codex, and DeepSeek launched its open-source V4 model with a million-token context window for agentic workflows.

  7. 92 signals

    OpenAI launched GPT-5.5, its latest frontier model with gains in real-world productivity. Alibaba's Qwen 3.6 27B matched Claude Sonnet 4.6 on agentic benchmarks, while Anthropic published a postmortem on Claude Code quality issues and partnered with NEC on Japan's AI workforce.

  8. 114 signals

    Alibaba's Qwen team released Qwen3.6-27B, a dense model matching flagship coding performance and fueling debate over efficiency versus scale. Anthropic shipped Claude Opus 4.7 alongside Claude Design, a vision-agent research preview, while Jeff Bezos's Physical AI startup Project Prometheus closed $10B at a $38B valuation.

9 dispatches
  1. 78 signals

    A federal judge blocked the Pentagon's legal effort to restrict Anthropic's operations, delivering a major win for the AI company. Google released TurboQuant, a KV-cache compression method that enables running Qwen models locally on a MacBook Air. A leaked Anthropic post surfaced online hinting at an upcoming model called Claude Mythos.

  2. 42 signals

    Mistral released Voxtral TTS, a 3B-parameter open-weights text-to-speech model that beats ElevenLabs Flash v2.5 in human preference tests while running on 3 GB of RAM. A federal court blocked the Pentagon from labeling Anthropic a supply chain security risk, and OpenAI discontinued Sora after Disney exited a reported $1 billion partnership.

  3. 80 signals

    LiteLLM suffered a supply chain attack affecting an estimated 47,000 users, prompting the community to compile drop-in alternatives. The Pentagon formalized Palantir's Maven AI as a core military system, with AI spending rising to $13.4 billion this year from $480 million in 2024. Liquid AI also demonstrated its LFM2-24B-A2B model running at 50 tokens per second in a web browser via WebGPU.

  4. 81 signals

    Epoch AI confirmed GPT-5.4 Pro solved an open problem in Ramsey hypergraph theory, a notable milestone in AI mathematical reasoning. OpenAI abandoned its Sora push and the Disney partnership, while a credential-stealing package hit PyPI in compromised LiteLLM releases 1.82.7 and 1.82.8, prompting an immediate downgrade advisory.

  5. 100 signals

    Apple demonstrated an iPhone 17 Pro running a 400-billion-parameter LLM locally. A US advisory panel warned that China's open-source AI ecosystem is eroding America's competitive lead, while Cursor's internal rankings named Moonshot's Kimi K2.5 the top open-source model.

  6. 59 signals

    A new open-source project, Flash-MoE, demonstrated running a 397-billion parameter mixture-of-experts model on a laptop, potentially reshaping access to frontier-scale AI. Anthropic unveiled Claude Opus 4.6, claiming top performance across agentic coding and tool use benchmarks, while ArXiv announced its spinout from Cornell as an independent nonprofit to address AI-generated submission pressures.

  7. No items were filed for this date.

  8. 65 signals

    Qwen3.5-27B matched models nearly 15x its size on the Game Agent Coding League benchmark, nearing GPT-5 mini's performance. Separately, a new adversarial coding benchmark capped top model scores at 11%, exposing widespread evaluation gaming. Simon Willison also published a comprehensive guide mapping the emerging patterns of agentic engineering.

  9. 40 signals

    Anthropic has made one-million-token context windows generally available across its Claude Opus 4.6 and Sonnet 4.6 model lines. ByteDance is circumventing US export controls by acquiring Nvidia hardware through offshore entities for its AI infrastructure, while arXiv is spinning off from Cornell as an independent nonprofit.