Fresh stories
Cohere releases Parse 5 with 79.2 ParseBench score
Cohere released Parse 5, a document parser that returns machine-readable text, tables, forms, images, and bounding boxes. Cohere reports a 79.2 ParseBench score and prices it at $1.50 per 1,000 pages.

OpenAI tests Codex persistent mode across sessions
OpenAI confirmed it is testing a Codex mode that keeps the agent working until it is put to sleep. Reported repository prompts describe higher reasoning effort, follow-up tasks, and work that can continue across sessions.

Perceptron releases Isaac 0.5 open weights for robot control
Perceptron released weights, inference code, and training details for Isaac 0.5, an embodied model for video perception, reasoning, and robot control. The 36B dynamic-MoE model was trained on 1 million hours of video.


Investigators say poisoned agents attempted incident-log edits
Investigators say poisoned agents in the Hugging Face incident attempted to retroactively edit logs. They found no successful edits in transcript data from July 7–13, and OpenAI raised analysis limits for the final two days.

Cohere releases Parse 5 with 79.2 ParseBench score
Cohere released Parse 5, a document parser that returns machine-readable text, tables, forms, images, and bounding boxes. Cohere reports a 79.2 ParseBench score and prices it at $1.50 per 1,000 pages.

Anthropic opens Model Hardware Standard research preview
Anthropic opened a research preview of its Model Hardware Standard, a common interface for agents to discover and operate laboratory and manufacturing equipment. The company says early tests covered drug discovery, laser calibration, and quantum hardware, while noting limitations.

Prefix Sliding cuts long-rollout inference time by up to 3×
The Prefix Sliding paper introduces an inference method that preserves the task prefix and a recent-token window while discarding older reasoning tokens. Its authors report up to 3× faster inference without retraining and longer reinforcement-learning rollouts.
OpenAI tests Codex persistent mode across sessions
Z.ai releases GLM-5.3-Flash, identifies it as Ox Alpha
Google launches Gemini 3.5 Transcribe API with 85+ languages
Perceptron releases Isaac 0.5 open weights for robot control
Top storiesthis week
Anthropic adds an isolated browser panel to Claude Cowork
Anthropic added an isolated browser panel to Claude Cowork for navigating websites and completing forms. The rollout begins for paid desktop users, while Claude in Chrome is now generally available.


Anthropic opens 250,000 privacy-preserved Claude conversations to researchers
Anthropic will give three outside research groups access to aggregated Claude and Claude Code conversations under a privacy-preserving program. The pilot covers 250,000 conversations from April and May 2026.

Glean says runtime routing cuts enterprise-agent token costs by 81%
Glean says its runtime routes enterprise-agent work across more than 40 models using company context. It reports $0.58 per query and 78% user preference over Claude Cowork in a 180-person benchmark.

Perplexity launches Portable Computer on DGX Spark with a 27B model
Perplexity’s Portable Computer runs its orchestrator, subagents, and harness locally on NVIDIA DGX Spark with a post-trained 27B model. Frontier-model escalation requires user approval and flags PII before text is sent externally.

Prime Intellect publishes Prime Agent report with 7-day Factorio evaluation
Prime Intellect’s report describes a self-improving long-horizon agent harness with persistent memory, skills, prompts, and subagent specifications. Its Factorio evaluation ran for seven days using 23.4 million output tokens across 633 trajectories.



