Fresh stories

U.S. adviser accuses Moonshot of distilling Anthropic Fable for Kimi K3
A U.S. tech adviser accused Moonshot AI of using Anthropic’s Fable to build Kimi K3, while China’s embassy denied related claims. Engineers questioned whether leaderboard results and missing logits fit a simple distillation story.
Cursor Router adds Cost mode for Teams and Enterprise coding requests
Cursor launched Router for Teams and Enterprise with Intelligence, Balance, and Cost modes. Cursor says the router can select models per request and cut costs by 60%, with admin controls for businesses.

Martian launches Ship beta with 50% lower-cost inference target
Martian launched Ship, a beta endpoint that sits between an app and its reference model and routes each request through cheaper paths while aiming to preserve behavior. Martian says the beta targets 50% lower cost.


Liang Wenfeng reportedly frames DeepSeek roadmap around scarce GPU supply
Posts quoting Liang Wenfeng said DeepSeek is targeting low positive API margins while constrained by GPU supply. They also said early-June capacity was about 20,000 H100-equivalent units and that Huawei capacity remains below Nvidia.

U.S. adviser accuses Moonshot of distilling Anthropic Fable for Kimi K3
A U.S. tech adviser accused Moonshot AI of using Anthropic’s Fable to build Kimi K3, while China’s embassy denied related claims. Engineers questioned whether leaderboard results and missing logits fit a simple distillation story.

VS Code adds assisted tool approvals for agent workflows
The latest VS Code release lets the model assess routine tool-call risk before asking a developer for approval. It also adds agent diff summaries and chat timing metadata for review and debugging around agent runs.

OpenAI releases Presence for eligible enterprise customers
OpenAI introduced Presence in limited general availability for eligible enterprise customers. It lets voice and chat agents use company systems, take approved actions, and escalate to humans.
Cursor Router adds Cost mode for Teams and Enterprise coding requests
Firecrawl releases /search endpoint with agent-ready excerpts
OpenAI says eval agent compromised Hugging Face production systems
Martian launches Ship beta with 50% lower-cost inference target
Top storiesthis week
Plasma opens Fractal Apache-2.0 CLI for recursive coding agents
Plasma open-sourced Fractal, an Apache-2.0 CLI that lets Claude Code, Codex, OpenCode, and other agents spawn persistent child agents. Each node gets its own worktree, memory, lifecycle, and Git history.


Artificial Analysis reports Kimi K3 averages 56.4 minutes on AA-Briefcase
Artificial Analysis reports Kimi K3 averages 56.4 minutes, 83 turns, and 120k output tokens per AA-Briefcase task. Kilo also found UI-build outputs close to Claude Fable 5 at 29% of the cost.

METR introduces expenditure horizon for cost-aware agent evals
METR introduced expenditure horizon, a method that compares human and agent performance as a function of spend on continuously scored tasks. A correction estimates human labor returns near $2.5K per 1% optimization, making cost curves central to the evaluation.

Moonshot pauses new Kimi K3 subscriptions after GPU capacity crunch
Moonshot said Kimi K3 demand pushed its GPUs near capacity, so it paused new subscriptions and split memberships into Kimi and Kimi Code plans. Users also reported slow serving and sold-out paid plans.

ChatGPT Work desktop adds cloud vs local run controls
OpenAI staff said ChatGPT Work runs in the cloud on web and mobile, while desktop can now choose cloud or computer execution. The clarification followed confusion about closed-laptop and local-environment behavior.



