Skip to main content
The Quantum Dispatch
Back to Home
large-language-models

Articles Tagged “Large Language Models”

9 articles found

Cover illustration for Nvidia KV Cache Transfer Skips 7-Second Re-Prefills
AI-Generated|Opinion
AI

Nvidia KV Cache Transfer Skips 7-Second Re-Prefills

Nvidia researchers moved a 32,768-token KV cache between model sizes in 278 milliseconds, replacing a 7-second re-prefill with closed-form linear math.

Dr. Nova Chen
Dr. Nova Chen★Aug 22, 2026★3 min read
Cover illustration for Grok 4.6 Brings a 500K Context Window to AI Agents
AI-Generated|Opinion
AI

Grok 4.6 Brings a 500K Context Window to AI Agents

xAI's Grok 4.6 ships a 500K-token context window and scores 61 on the Artificial Analysis index, holding Grok 4.5's $2 per million input token price.

Dr. Nova Chen
Dr. Nova Chen★Aug 13, 2026★5 min read
Cover illustration for Google Gemini 2.5 Pro Deep Think Brings Parallel Reasoning to Everyone
AI-Generated|Opinion
AI

Google Gemini 2.5 Pro Deep Think Brings Parallel Reasoning to Everyone

Google launched Gemini 2.5 Pro with Deep Think on June 22, 2026 — a parallel-reasoning mode for hard math and coding, with a 2M-token context, now live on the API, AI Studio, and Vertex AI.

Dr. Nova Chen
Dr. Nova Chen★Jun 24, 2026★5 min read
Cover illustration for GLM-5.2 Open Weights Arrive as a Top Coding Model at a Fraction of the Cost
AI-Generated|Opinion
AI

GLM-5.2 Open Weights Arrive as a Top Coding Model at a Fraction of the Cost

Z.ai released GLM-5.2 open weights under an MIT license on June 16, 2026 — an open-weight coding model that rivals the best closed systems on long-horizon benchmarks at roughly one-sixth the cost.

Dr. Nova Chen
Dr. Nova Chen★Jun 22, 2026★5 min read
Cover illustration for GLM-5.2 Arrives With a Usable 1M-Token Context and MIT Open Weights
AI-Generated|Opinion
AI

GLM-5.2 Arrives With a Usable 1M-Token Context and MIT Open Weights

Z.ai's GLM-5.2 is a coding-first open-weight model with a usable 1-million-token context window and MIT-licensed weights that drop into agentic dev tools.

Dr. Nova Chen
Dr. Nova Chen★Jun 17, 2026★5 min read
Cover illustration for OpenAI GPT-6 Arrives With 2M Token Context and Sub-0.1% Hallucination Rate
AI-Generated|Opinion
AI

OpenAI GPT-6 Arrives With 2M Token Context and Sub-0.1% Hallucination Rate

OpenAI's GPT-6 sets a new benchmark ceiling: 2 million token context, under 0.1% hallucination rate, and 40%+ gains over GPT-5.4 across coding, reasoning, and agentic task completion.

Dr. Nova Chen
Dr. Nova Chen★Apr 18, 2026★5 min read
Cover illustration for OpenAI's 'Spud' Completes Pretraining — The Next Frontier Model Is Almost Here
AI-Generated|Opinion
AI

OpenAI's 'Spud' Completes Pretraining — The Next Frontier Model Is Almost Here

OpenAI confirmed its next frontier model, codenamed Spud, finished pretraining on March 24 — a unified multimodal AI expected to arrive within weeks.

Dr. Nova Chen
Dr. Nova Chen★Apr 12, 2026★5 min read
Cover illustration for OpenAI Releases GPT-5.3 Instant — Cutting Hallucinations by 27 Percent and Delivering Snappier Answers
AI-Generated|Opinion
AI

OpenAI Releases GPT-5.3 Instant — Cutting Hallucinations by 27 Percent and Delivering Snappier Answers

GPT-5.3 Instant focuses on reliability over raw power, reducing hallucination rates by nearly 27 percent while trimming unnecessary refusals and preambles.

Dr. Nova Chen
Dr. Nova Chen★Mar 5, 2026★5 min read
Cover illustration for Apple Is Rebuilding Siri From the Ground Up With LLM-Powered Conversational Intelligence in iOS 26.4
AI-Generated|Opinion
AI

Apple Is Rebuilding Siri From the Ground Up With LLM-Powered Conversational Intelligence in iOS 26.4

Apple's Siri overhaul replaces the command-based architecture with on-device large language models, bringing natural conversation and app-aware context to one billion iPhones.

Dr. Nova Chen
Dr. Nova Chen★Mar 3, 2026★5 min read