Skip to content
AI Primer

Explore what's new in AI

Where people deep in AI come to stay current.

Filters

Category

Tags

Release

Qwen3.8-Max launches on OpenRouter with 1M-token context

Alibaba's Qwen3.8-Max is live on Venice and OpenRouter while open weights are still described as coming soon. Reports cite a 2.4T-parameter model with strong Vals and vision benchmark results.

Qwen3.8-Max launches on OpenRouter with 1M-token context
Qwen·3rd August·8 min read
Breaking

Hermes Agent ships Herald with Ironproxy secrets lockdown

Teknium shipped the Herald release of Hermes Agent as a desktop-agent and runtime update. The release adds voice chats, desktop plugins, Agent2Agent/webhooks, productivity skills, grounded research, Ironproxy secrets lockdown, and trace-driven token-efficiency work from 250,000 conversations.

Hermes Agent ships Herald with Ironproxy secrets lockdown
New
Hermes Agent·3rd August·7 min read
Release

MiniMax releases H3 open weights with day-zero vLLM-Omni support

MiniMax released H3 weights on Hugging Face for text-to-video, image-to-video, reference-to-video, and editing workflows. vLLM-Omni, ComfyUI, SGLang Diffusion, and fal added support at launch.

MiniMax releases H3 open weights with day-zero vLLM-Omni support
MiniMax·2nd August·7 min read
See all stories →
⌨️Agentic Engineering(2)
🧠Models, Serving & APIs(18)
⚙️Building Agents(5)
🛡️Trust, Evaluation & Reliability(6)
🔎Knowledge, Memory & Retrieval(1)
📈Adoption & Market Strategy(4)

Top storiesthis week

Workflow

Codex users route tasks across GPT-5.6 Sol, Terra, and Luna to cut token cost

Practitioners reported better Codex multi-agent runs by raising concurrency and splitting work across Sol, Terra, and Luna. One workflow sends deploy tasks to Luna Max to preserve Sol tokens.

Codex users route tasks across GPT-5.6 Sol, Terra, and Luna to cut token cost
Codex·2nd August·8 min read
See all stories →
AI PrimerAI Primer

Your daily guide to AI tools, workflows, and creative inspiration.

© 2026 AI Primer. All rights reserved.