A quiet-ish weekend in AI, but the big one is real: DeepSeek V4-Flash 0731 is out as open weights, with V4-Pro said to follow soon.

Models & research

  • DeepSeek V4-Flash 0731 released open-weights — official org repo is live (Hugging Face). Community threads report dirt-cheap API pricing (~18x cheaper input pricing vs Claude per one r/Anthropic post) — pricing and “matches Opus 4.8” claims are community-reported, not independently verified.
  • Running Kimi K3 on MI355X at better performance-per-dollar than B300 — vendor benchmark (wafer.ai) claiming AMD’s MI355X beats the B300 on inference $/token for Kimi K3. Take the numbers as vendor claims, but it’s the kind of data that decides self-host GPU buys.
  • SKILL-KD: skill distillation for frozen LLM agents (arXiv 2607.28048) — surfaced but unverified this run: claims a framework that turns teacher-student discrepancies into reusable skills for frozen agents. Check the abs page before citing.

Agent frameworks & tooling

  • Diagrid Catalyst 2.0 adds durable recovery to LangGraph — crash-safe durable recovery and signed execution histories for LangGraph via Dapr workflows (Bluesky). Relevant if you need restart-safe, auditable agent runs in production.

Industry

  • China pushes open-source AI at the UN summit — a large Chinese delegation at the UN AI for Good summit argued Chinese open models are the future for most of the world (Semafor).
  • Fields Medal winner Jacob Tsimerman to join OpenAI for AI-safety work — the Toronto mathematician takes a leave to work on AI safety (WSJ).
  • Apple caps bug-report submissions citing a deluge of AI-assisted reports — a 30-day cool-off with quota exceptions for researchers (FT). A real-world signal that AI-generated issue volume is forcing policy changes.

Compiled from the morning digest — X, HN, Reddit, Techmeme, Bluesky, arXiv. Hype cut, links kept.

All gathered items — what was cut and why (9)