Skip to content
AI Primer

Explore what's new in AI

Where people deep in AI come to stay current.

Filters

Category

Tags

Breaking

Qwen 3.8 Max 2.4T open weights ship as text-only, users say

LocalLLaMA users said Qwen 3.8 Max 2.4T open weights are text-only while the API keeps vision support. A linked Qwen3.8-27B ModelScope page reportedly returned 404 before release.

Qwen 3.8 Max 2.4T open weights ship as text-only, users say
New
Qwen·12th August·4 min read
Composio and Ante benchmark coding agent harnesses with 47%–67% success range

Composio and Ante benchmark coding agent harnesses with 47%–67% success range

Composio and Ante tests reported that the same models behaved very differently by harness. DeepSeek V4 Flash ranged from 47% to 67% task success and $0.019 to $0.104 per task across harnesses.

💳Harness Engineering10th August·6 min read
Breaking

OpenRouter updates Auto Router with 30 task types and 7-day spend-based routing

OpenRouter upgraded Auto Router to classify prompts into about 30 task types, then route by anonymized 7-day spend share and cost tier. OpenRouter says the max tier beat the old router across five benchmark domains.

OpenRouter updates Auto Router with 30 task types and 7-day spend-based routing
New
Model Routing·10th August·5 min read
See all stories →
⌨️Agentic Engineering(2)
🧠Models, Serving & APIs(5)
🛡️Trust, Evaluation & Reliability(1)

Top storiesthis week

Microsoft Copilot traces report 87% of LLM calls came from agents

A Microsoft Copilot trace analysis said 87% of LLM calls came from the agent, not direct user turns. Related posts warned token use and web requests can scale far faster than human prompt counts.

Microsoft Copilot traces report 87% of LLM calls came from agents
Coding Agents·9th August·7 min read
See all stories →
AI PrimerAI Primer

Your daily guide to AI tools, workflows, and creative inspiration.

© 2026 AI Primer. All rights reserved.