Skip to main content
The Quantum Dispatch
Back to Home
llm

Articles Tagged “LLM

17 articles found

Cover illustration for GPT-5.6 Sol API Price Falls to $4 Per Million Tokens
AI-Generated|Opinion
AI

GPT-5.6 Sol API Price Falls to $4 Per Million Tokens

OpenAI cut GPT-5.6 Sol API pricing on August 21: input drops 20% to $4 and output falls 33% to $20 per million tokens through November 21, 2026.

Dr. Nova Chen
Dr. Nova ChenAug 23, 20264 min read
Cover illustration for Apple Music Will Show Made With AI Labels This Year
AI-Generated|Opinion
AI

Apple Music Will Show Made With AI Labels This Year

Apple Music will make its AI Transparency Tags mandatory and surface Made With AI labels to listeners later in 2026, shifting from removal to disclosure.

Dr. Nova Chen
Dr. Nova ChenAug 21, 20264 min read
Cover illustration for Embedding Models Guide: Dense vs Multi-Vector RAG
AI-Generated|Opinion
AI

Embedding Models Guide: Dense vs Multi-Vector RAG

A practical 2026 guide to choosing embedding models for RAG: dense bi-encoders, multi-vector late interaction, and rerankers, with cost and memory math.

Dr. Nova Chen
Dr. Nova ChenAug 18, 20269 min read
Cover illustration for Claude Opus 5 Brings Frontier Coding at Half the Cost
AI-Generated|Opinion
AI

Claude Opus 5 Brings Frontier Coding at Half the Cost

Claude Opus 5 arrives at $5 per million input tokens, delivering near-frontier coding performance at half the cost of Anthropic's Fable 5 model.

Dr. Nova Chen
Dr. Nova ChenJul 24, 20264 min read
Cover illustration for Gemini 3.6 Flash Trims Output Tokens by 17% for Devs
AI-Generated|Opinion
AI

Gemini 3.6 Flash Trims Output Tokens by 17% for Devs

Google's Gemini 3.6 Flash uses 17% fewer output tokens at equal quality, costs $1.50 per million input, and jumps to 49% on the DeepSWE benchmark.

Dr. Nova Chen
Dr. Nova ChenJul 21, 20265 min read
Cover illustration for Kimi K2.7 Code Brings Open Weights to GitHub Copilot
AI-Generated|Opinion
AI

Kimi K2.7 Code Brings Open Weights to GitHub Copilot

GitHub Copilot's first open-weight model, Kimi K2.7 Code, reached Business and Enterprise plans on July 7 — a 1T-parameter MoE with 32B active.

Dr. Nova Chen
Dr. Nova ChenJul 15, 20265 min read
Cover illustration for GPT-5.6 Tiers Explained: Which of Sol, Terra, Luna Fits
AI-Generated|Opinion
AI

GPT-5.6 Tiers Explained: Which of Sol, Terra, Luna Fits

OpenAI's GPT-5.6 shipped July 9 in three tiers — Sol, Terra, Luna — priced from $1 to $30 per million tokens. Here's how to pick the right one.

Dr. Nova Chen
Dr. Nova ChenJul 15, 20266 min read
Cover illustration for Cohere Transcribe Arabic: Best Open Arabic Speech AI
AI-Generated|Opinion
AI

Cohere Transcribe Arabic: Best Open Arabic Speech AI

Cohere Transcribe Arabic, released July 7 under Apache 2.0, hits a 25.87 word error rate — the top open-source Arabic ASR, beating Whisper Large V3.

Dr. Nova Chen
Dr. Nova ChenJul 14, 20264 min read
Cover illustration for GPT-Live Gives ChatGPT Real-Time Full-Duplex Voice
AI-Generated|Opinion
AI

GPT-Live Gives ChatGPT Real-Time Full-Duplex Voice

OpenAI's GPT-Live brings full-duplex voice to ChatGPT, letting the AI listen and speak simultaneously for 150M+ weekly voice users worldwide.

Dr. Nova Chen
Dr. Nova ChenJul 13, 20264 min read
Cover illustration for SWE-1.7 Brings Near-Frontier Coding Power to Devin
AI-Generated|Opinion
AI

SWE-1.7 Brings Near-Frontier Coding Power to Devin

Cognition's SWE-1.7 scored 42.3% on FrontierCode Main at about $1.97 per task, bringing near-frontier coding into Devin at ~1000 tokens/sec.

Dr. Nova Chen
Dr. Nova ChenJul 13, 20265 min read
Cover illustration for ByteDance's Seed 2.1 Models Bring Frontier Coding to a Lower Price Point
AI-Generated|Opinion
AI

ByteDance's Seed 2.1 Models Bring Frontier Coding to a Lower Price Point

ByteDance unveiled Seed 2.1 Pro and Turbo on June 24, 2026 — strong coding and agent models with million-token context and a dramatically lower cost of ownership.

Dr. Nova Chen
Dr. Nova ChenJun 30, 20265 min read
Cover illustration for Microsoft Launches Seven In-House MAI Models With Frontier Tuning
AI-Generated|Opinion
AI

Microsoft Launches Seven In-House MAI Models With Frontier Tuning

Microsoft unveiled a family of seven in-house MAI models spanning reasoning, coding, image, voice, and transcription — plus Frontier Tuning to customize them on your own data.

Dr. Nova Chen
Dr. Nova ChenJun 10, 20266 min read
Cover illustration for ChatGPT's New 'Dreaming' Memory Learns About You in the Background
AI-Generated|Opinion
AI

ChatGPT's New 'Dreaming' Memory Learns About You in the Background

OpenAI's 'Dreaming' lets ChatGPT synthesize and self-update memories across chats automatically, with a transparent page to review and edit what it recalls.

Dr. Nova Chen
Dr. Nova ChenJun 8, 20264 min read
Cover illustration for Meta Launches Muse Spark: Its First Closed-Weight Frontier AI Model
AI-Generated|Opinion
AI

Meta Launches Muse Spark: Its First Closed-Weight Frontier AI Model

Meta Superintelligence Labs drops Muse Spark on April 8 — a fully closed frontier AI competing with GPT-5.4 and Gemini, marking Meta's sharpest strategic turn yet.

Dr. Nova Chen
Dr. Nova ChenApr 16, 20265 min read
Cover illustration for NVIDIA's AI-Q Blueprint Brings Enterprise Agentic AI to Adobe, Salesforce, and SAP
AI-Generated|Opinion
AI

NVIDIA's AI-Q Blueprint Brings Enterprise Agentic AI to Adobe, Salesforce, and SAP

NVIDIA's AI-Q Blueprint gives enterprises an open framework for building AI agents that perceive, reason, and act — slashing query costs by 50% with a hybrid routing architecture.

Dr. Nova Chen
Dr. Nova ChenApr 9, 20265 min read
Cover illustration for PrismML's Bonsai Is a 1-Bit LLM That Runs on a Smartphone and Matches Full-Size Models
AI-Generated|Opinion
AI

PrismML's Bonsai Is a 1-Bit LLM That Runs on a Smartphone and Matches Full-Size Models

Caltech startup PrismML emerged from stealth with Bonsai, a 1-bit LLM family that's 14x smaller, 8x faster, and 5x more energy-efficient than standard 8B models — and runs on an iPhone.

Dr. Nova Chen
Dr. Nova ChenApr 9, 20265 min read
Cover illustration for Alibaba's Qwen3.6-Plus Delivers 1M-Token Context and Repository-Level Agentic Coding
AI-Generated|Opinion
AI

Alibaba's Qwen3.6-Plus Delivers 1M-Token Context and Repository-Level Agentic Coding

Qwen3.6-Plus arrives with a default 1 million-token context window and breakthrough agentic coding performance, enabling AI that can navigate and rewrite entire software repositories autonomously.

Dr. Nova Chen
Dr. Nova ChenApr 4, 20264 min read