Skip to main content
The Quantum Dispatch
Back to Home
long-context

Articles Tagged “Long Context

5 articles found

AI

Qwen3.8-Max Packs 2.4T Parameters Into a 1M Context

Alibaba's Qwen3.8-Max is a 2.4-trillion-parameter sparse MoE model with a 1M-token window, priced at $2 per million input tokens and open weights next week.

Dr. Nova Chen
Dr. Nova ChenAug 4, 20267 min read
AI

DeepSeek V4-Flash 0731 Tops V4-Pro at a Third the Price

DeepSeek's retrained V4-Flash 0731 beats its own V4-Pro preview on every published agentic benchmark at $0.28 per million output tokens, MIT licensed.

Dr. Nova Chen
Dr. Nova ChenAug 4, 20266 min read
AI

Kimi K3 Open Weights Ship 2.8T Parameters and 1M Context

Moonshot AI released Kimi K3 open weights on July 26: 2.8 trillion parameters, 104B active per token, 1M context, and a modified MIT license.

Dr. Nova Chen
Dr. Nova ChenJul 29, 20266 min read
AI

GLM-5.2 Arrives With a Usable 1M-Token Context and MIT Open Weights

Z.ai's GLM-5.2 is a coding-first open-weight model with a usable 1-million-token context window and MIT-licensed weights that drop into agentic dev tools.

Dr. Nova Chen
Dr. Nova ChenJun 17, 20265 min read
AI

MiniMax M3 Open-Weight Model Lands With 1M Context and Native Multimodal

MiniMax published the open weights and technical report for M3, an open-weight model pairing a 1M-token context window with native image and video understanding.

Dr. Nova Chen
Dr. Nova ChenJun 15, 20266 min read