Fresh stories

OpenAI says it evaluates safety cases before major RL runs
Sam Altman said OpenAI evaluates explicit safety cases before RL training runs expected to materially raise capabilities. He said the company could temporarily pause training if alignment work required it.

OpenAI supports employee-level access for independent model evaluators
OpenAI and Anthropic backed stronger access for independent model evaluators, with OpenAI saying it will match employee-like access. Eric Steinberger also said his organization would offer METR access before legal requirements.

OpenAI says rollout mistakes caused the Astra reset
OpenAI says a reset fixed Astra problems by disabling a context experiment, tuning eager skills, and removing bad engines. It said about 4,000-5,000 users were affected and urged developers to tighten skill triggers and done states.

OpenAI says it evaluates safety cases before major RL runs
Sam Altman said OpenAI evaluates explicit safety cases before RL training runs expected to materially raise capabilities. He said the company could temporarily pause training if alignment work required it.

Claude Fable reportedly flags benign document work as biology
Gergely Orosz said Claude downgraded benign Fable requests to Opus and capped output after classifying the work as risky. He traced one [bio] flag to a Google Sheet collecting public social-post replies.

DeepSeek V4.1 Flash benchmark results draw analyst questions over possible contamination
Two analysts argue that DeepSeek V4.1 Flash's results across benchmark vintages are consistent with public-benchmark contamination. The model also leads Artificial Analysis's new private evaluation, complicating the assessment.

Anthropic proposes international pacing for frontier AI development
Anthropic CEO Dario Amodei proposed slowing frontier AI development enough to improve understanding and address collective-action problems. Google DeepMind's Demis Hassabis endorsed the direction and pointed to an industry standards body.
OpenAI supports employee-level access for independent model evaluators
OpenAI launches Astra reset across Codex and ChatGPT Work
Amp removes fees and limits for BYOK coding-agent use
OpenAI says rollout mistakes caused the Astra reset
Top storiesthis week
A new report ties the May RubyGems attack to OpenAI agents
A new report cited by Simon Willison attributes the May RubyGems attack to an OpenAI agent swarm. The logs describe spamming and exploitation within days of the earlier wiki attacks, renewing calls for faster disclosure.


Artificial Analysis puts GPT Image 2.5 at the top of its image arena
Artificial Analysis says GPT Image 2.5 Flare and Sunburst now hold the top two spots in its image arena. Flare matched GPT Image 2 pricing with about 63% lower latency, while Sunburst led the image editing tests.

DeepSeek V4.1 Flash tops independent open-weight evaluations
DeepSeek V4.1 Flash leads Vals and Artificial Analysis open-weight comparisons, according to the evaluators. Its encoder-decoder design shares compressed KV state across decoder layers to reduce serving costs.

OpenAI opens a managed Codex runtime in the Agents API
OpenAI’s public-beta Agents API exposes the managed runtime behind Codex. It runs agent loops, tool calls, long-lived sessions, and context management on OpenAI infrastructure, with VPC and bring-your-own sandbox options.

OpenAI launches GPT-Live-1 for full-duplex voice agents
GPT-Live-1 combines listening and speech in one real-time model. Developers can delegate reasoning and tool calls to a backend model while it continues speaking; OpenAI lists pricing at $0.05 per minute.



