Fresh stories

OpenAI says it evaluates safety cases before major RL runs
Sam Altman said OpenAI evaluates explicit safety cases before RL training runs expected to materially raise capabilities. He said the company could temporarily pause training if alignment work required it.

Claude Fable reportedly flags benign document work as biology
Gergely Orosz said Claude downgraded benign Fable requests to Opus and capped output after classifying the work as risky. He traced one [bio] flag to a Google Sheet collecting public social-post replies.

DeepSeek V4.1 Flash benchmark results draw analyst questions over possible contamination
Two analysts argue that DeepSeek V4.1 Flash's results across benchmark vintages are consistent with public-benchmark contamination. The model also leads Artificial Analysis's new private evaluation, complicating the assessment.
Top storiesthis week
Anthropic proposes international pacing for frontier AI development
Anthropic CEO Dario Amodei proposed slowing frontier AI development enough to improve understanding and address collective-action problems. Google DeepMind's Demis Hassabis endorsed the direction and pointed to an industry standards body.


OpenAI supports employee-level access for independent model evaluators
OpenAI and Anthropic backed stronger access for independent model evaluators, with OpenAI saying it will match employee-like access. Eric Steinberger also said his organization would offer METR access before legal requirements.

OpenAI launches Astra reset across Codex and ChatGPT Work
OpenAI says it fixed skill-triggering, context-management, and engine-configuration issues behind degraded Astra responses. The reset is now rolling out across Codex and ChatGPT Work.

Amp removes fees and limits for BYOK coding-agent use
Amp says its coding agent is free when users supply their own compute, model subscription, or API key. The rollout includes routing support for external providers such as Ollama Cloud, OpenRouter, and custom OpenAI-compatible URLs.

OpenAI says rollout mistakes caused the Astra reset
OpenAI says a reset fixed Astra problems by disabling a context experiment, tuning eager skills, and removing bad engines. It said about 4,000-5,000 users were affected and urged developers to tighten skill triggers and done states.




