Fresh stories
Modal launches globally available RDMA clusters
Modal made multi-node, RDMA-connected clusters globally available. The runtime also adds full-kernel VM sandboxes, sandbox sidecars, and sticky sessions that reconnect WebSockets to the same container.

Rivet reports 0.82 MB per session for its Pi integration
Rivet's Pi integration keeps the agent loop, sessions, and credentials in the backend while sandboxes provide tools. Rivet reports durable execution, shared control, and SQLite persistence alongside the 0.82 MB per-session figure.


Modal launches globally available RDMA clusters
Modal made multi-node, RDMA-connected clusters globally available. The runtime also adds full-kernel VM sandboxes, sandbox sidecars, and sticky sessions that reconnect WebSockets to the same container.

Anthropic adds Claude Code Mods for JavaScript and TypeScript plugins
Claude Code Mods let JavaScript or TypeScript plugins replace agent behavior, inspect session events, spawn subagents, and change the UI. Practitioners have shared a mod-building skill and hot-reload workflows.

Black Forest Labs launches FLUX 3 Image with native 4K output
FLUX 3 Image supports bounding-box layouts, targeted edits, native 4K output, and ten reference images. Commercial weights are available, open weights are planned, and API use is half-price through October 8.

Rivet reports 0.82 MB per session for its Pi integration
Rivet's Pi integration keeps the agent loop, sessions, and credentials in the backend while sandboxes provide tools. Rivet reports durable execution, shared control, and SQLite persistence alongside the 0.82 MB per-session figure.
Amp investigates ChatGPT subscription connection errors
Amp says it is working with OpenAI on ChatGPT connection errors and directs affected users to its legacy connection. Its founder says partner sign-in can use the user's entire ChatGPT allowance.
Pi releases version 1.0 with durable sessions backed by SQLite
Pi 1.0 introduces Pi Durable, with SQLite storage and concurrent sessions shared across multiple clients. Published examples describe replay-safe tasks and background subagents.
Kev releases 1.0 open-weight decision models with a 64k document window
Kev 1.0 introduces Kev-27B and updates Kev-9B with a 64k document window and TypeSafe SDK compatibility. The release includes weights, source code, and fine-tuning and deployment skills.
Top storiesthis week
GPT-6.1 Sol scores 86.3% on MathArena BrokenArXiv
GPT-6.1 Sol scored 86.3% on MathArena BrokenArXiv, 78.6% on Braintrust problem solving, and third in Code Arena WebDev. A user also reports that it outperformed its mini-SWE result in Codex.


Perplexity open-sources pplx-embed-v2-context-9b-preview
Perplexity released pplx-embed-v2-context-9b-preview, which encodes chunks using whole-document context. Perplexity reports leading results on ConTEB and Turbopuffer's context benchmark.

Reports rank Gemini 4 Argon highly on four engineering benchmarks
Reports place Gemini 4 Argon at 77.9% on DeepSWE, 57.6% on Terminal-Bench, 77.5% on AutomationBench-AA, and 68% on CWE-Bench. The reports also cite lower cost or token use than rivals.

Google rolls out Gemini 4 Argon to cyber defenders
Google is rolling out Gemini 4 Argon through Fairwind to government users, vetted cyber defenders, and trusted testers. The model supports up to 1M output tokens and costs $2/M input and $10/M output.

OpenAI adds capacity after GPT-6.1 Sol overloads ChatGPT
OpenAI added capacity after GPT-6.1 Sol became its most demanded model and caused heavy load in ChatGPT and Codex. OpenAI says the added capacity should nearly double speeds from the prior day.







