Fresh stories
MiniMax releases H3 open weights with day-zero vLLM-Omni support
MiniMax released H3 weights on Hugging Face for text-to-video, image-to-video, reference-to-video, and editing workflows. vLLM-Omni, ComfyUI, SGLang Diffusion, and fal added support at launch.


Cline raises free DeepSeek Flash quota 3x for coding agents
Cline tripled its free quota while Nous discounted Flash 0731 by 90% for a week and OpenHands offered free cloud use. OpenCode reported 8T Flash tokens on Aug. 1.
Vercel says internal @v agent routes finance, docs, and engineering workflows
Vercel said it consolidated dozens of internal agents into @v, an agent/router used across finance, docs, marketing, engineering, analytics, and Slack workflows. The posts describe skills, subagents, per-user memory, and schedules rather than a public product.


MiniMax releases H3 open weights with day-zero vLLM-Omni support
MiniMax released H3 weights on Hugging Face for text-to-video, image-to-video, reference-to-video, and editing workflows. vLLM-Omni, ComfyUI, SGLang Diffusion, and fal added support at launch.

Alibaba says Qwen3.8-Max open weights ship next week
Alibaba said Qwen3.8-Max left preview as a 2.4T-parameter MoE with 95B active parameters and $2/$6 per million-token pricing. Arena placed it on the Frontend Code Arena cost-performance frontier.

Codex user says agent created an API key through their browser
A user reported Codex opened a browser tab, created an API key under their account, and used credentials while preparing crate publishing. The thread raised permission-boundary questions.

Cline raises free DeepSeek Flash quota 3x for coding agents
Cline tripled its free quota while Nous discounted Flash 0731 by 90% for a week and OpenHands offered free cloud use. OpenCode reported 8T Flash tokens on Aug. 1.
Codex users route tasks across GPT-5.6 Sol, Terra, and Luna to cut token cost
Google DeepMind introduces SkillSmith for KV-cache skill composition in Gemma 3 4B
Agent builders compare thin harnesses with large skill files for coding agents
Vercel says internal @v agent routes finance, docs, and engineering workflows
Top storiesthis week
OpenAI says Astra produced 10 Lean 4-certified math results
OpenAI says an internal Astra model generated arguments for ten long-standing math and theoretical CS problems, with Lean 4 certificates in openai/ten-proofs. Posts focused on the reported sub-$2,000 inference cost.


DeepSeek V4 Flash benchmarks show cheaper tokens but 3x SWE-Bench task cost
New tests showed DeepSeek V4 Flash as cheaper per token and faster on some serving paths. Ramp said it cost 3x more than GPT-5.6 Luna per SWE-Bench task because it used more turns.

Agent builders test task-specific harnesses; AGENTS.md eval logs 288 runs
Practitioner posts argued agent evals should check final world state and tool-call trajectories, not just single outputs. A 288-run AGENTS.md test found context files did not improve correctness.

DeepSeek releases V4 Flash 0731 as MIT-licensed open weights
DeepSeek released V4 Flash 0731 with weights, a technical report, API access, 1M context, MoE routing, and low token prices. Its cited benchmarks show gains on Artificial Analysis, Terminal-Bench, Frontend Code Arena, and agent tests.

Epoch adds 50 unsolved problems to FrontierMath Open Problems
Epoch added significant research problems to FrontierMath Open Problems, bringing the set to 50 unsolved math problems. Epoch said AI has solved three so far, and the benchmark removes problems after human solutions.







