Articles Tagged “Small Language Models”
5 articles found

Liquid AI d1 Models: Open Decision AI That Answers in 8 ms
Liquid AI's open d1-3B and d1-omni-600M decision models skip token generation and answer in one pass, as fast as 8 ms on an RTX 4090 and 50 ms on Jetson.

Strands Decider 2B: What Amazon's Free Decision Model Does
Strands Decider 2B is a free 2B-parameter decision model that picks options in about 115ms on a single GPU. Here is how it works and where it fits.

Needle 2 Brings 14MB Function-Calling AI to Raspberry Pi 5
Needle 2 is a 14MB function-calling model that turns plain English into Raspberry Pi 5 actions in about 80ms on the CPU alone, with no AI HAT needed.

Microsoft's MAI-Code-1-Flash Brings a Tiny, Fast Coding Model to Copilot's Free Tier
Microsoft launched MAI-Code-1-Flash on June 2, 2026 — a compact 5B-parameter in-house coding model now rolling out across GitHub Copilot, including the free tier.

AT&T Slashes AI Costs by 90 Percent Using Small Language Models That Process 27 Billion Tokens Daily
AT&T's multi-agent architecture routes tasks to specialized small models, cutting AI infrastructure costs while scaling to 27 billion tokens per day.
