Articles Tagged “AI Efficiency”
6 articles found

Claude Haiku 5.5 Pricing: What a 90% Cheaper Haiku Unlocks
Claude Haiku 5.5 launches at $0.10 per million input tokens with a 1M context window, adjustable effort, and a 72.4% OSWorld score, up from 15.7%.

OpenAI Cuts Luna Prices 80% After AI Rewrote Its Kernels
OpenAI dropped GPT-5.6 Luna to $0.20 per million input tokens after Sol rewrote its own GPU kernels, cutting end-to-end serving costs by 20%.

Cursor Router Auto-Picks the Cheapest Capable Model
Cursor Router, launched July 22, routes each coding request to the cheapest capable model, delivering frontier quality at up to 60% lower cost.

MIT's GIFT Turns 2D Designs Into CAD at 20% Compute
MIT's GIFT system converts a 2D image and text into executable CAD code using about 20% of the compute rival methods need, with no human labeling.

HyperNova 60B Uses Quantum-Inspired Math to Halve an LLM's Size With Near-Zero Accuracy Loss
Multiverse Computing's free HyperNova 60B compresses a 120B-parameter model by 50% using quantum tensor methods, benchmarking 5x better on tool-calling tasks.

MIT Researchers Develop a Proxy Model Technique That Doubles LLM Training Speed
A new MIT method uses a lightweight proxy model to predict reasoning outputs, cutting the reinforcement learning rollout bottleneck in half.
