Articles Tagged “Hugging Face”
23 articles found

NVIDIA Nemotron Olympiad Recipe: What the Open Release Means
NVIDIA open-sourced the Nemotron 3 recipe behind a 535.4/600 IOI 2026 run and a gold-level 30/42 IMO score, with checkpoints, data and a new benchmark.

AstaBrief 8B: Ai2's Open Model for Cited Science Reports
AstaBrief 8B is Ai2's Apache 2.0 open model that writes cited research reports in 51 seconds, 3.5x faster than its Claude pipeline. Here is how it works.

Transformers Now Runs GGUF Quants on Mac at llama.cpp Speed
Hugging Face Transformers can now run GGUF quants packed on Apple Silicon. A Qwen3.5-4B Q4_K_M model shrinks from 8.42GB to 2.74GB. Here's how it works.

Ternary Bonsai 2 27B: A 5.9GB Model for 16GB Laptops
Ternary Bonsai 2 27B squeezes Qwen3.8-27B into 5.9GB with 1.76-bit weights and keeps 98.2% of its benchmark score. Here's how to run it locally.

Nvidia to Buy Hugging Face in a $12.93 Billion Deal
Nvidia agreed to acquire Hugging Face for $12.93 billion, its second-largest deal ever, and says the open model hub will stay open to everyone.

WebGPU Kernels Make Local AI in the Browser 2.57x Faster
Hugging Face published 207 WebGPU kernels as a JavaScript library, reporting a 2.57x geometric-mean speedup over ORT WebGPU on an Apple M4 GPU.

Microduck Is a $399 Open Source Robot You Can Retrain
Hugging Face and Pollen Robotics launched Microduck, a 25cm biped with 15 motors, LiDAR and an Apache 2.0 reinforcement learning stack, for $399.

Speech Recognition Benchmarks Get a Three-Test Audit
A Hume AI study of 11 open speech models introduces three diagnostics that separate genuine transcription skill from memorized benchmark patterns.

Ornith-1.5 Open Weights Score 86.1 on Terminal-Bench
Ornith-1.5 ships MIT-licensed weights from 9B to 397B, and the flagship posts 86.1 on Terminal-Bench 2.1 while a 35B MoE sibling runs far leaner.

K-EXAONE 2.0 Ships 750B Open Weights Under Apache 2.0
LG AI Research released K-EXAONE 2.0, a 750-billion-parameter mixture-of-experts model with 37B active parameters, under a permissive Apache 2.0 license.

Inkling Is a 975B Open-Weights Model Under Apache 2.0
Thinking Machines released Inkling, a 975B-parameter Apache 2.0 model with 41B active, a 1M-token context window, and native four-modality reasoning.

Real World VoiceEQ Benchmarks the Human Side of Voice AI
Hume AI and Hugging Face open a voice AI benchmark built on over one million human ratings, covering 40+ models and 60+ metrics of speech quality.

NVIDIA NeMo Automodel Fine-Tunes Any Diffusers Model
NVIDIA and Hugging Face shipped distributed fine-tuning for any Diffusers model — FLUX.1-dev LoRA hits 53.73 images per second on eight H100s.

Kimi K3 Becomes the Largest Open-Weight AI Model Yet
Moonshot AI's Kimi K3 is a 2.8-trillion-parameter open-weight model that ranks 3rd on GDPval-AA v2, with full weights arriving July 27.

Thinking Machines Inkling: A 975B Open-Weights Model
Thinking Machines Lab released Inkling, a 975B-parameter open-weights multimodal model with 41B active per token and a 1M-token context window.

PP-OCRv6 Brings Tiny, 50-Language Open-Source OCR to Hugging Face
PaddlePaddle released PP-OCRv6 on June 22, 2026 — a family of tiny open-source OCR models covering 50 languages, now available on the Hugging Face Hub.

Holo3.1 Brings Fast, Private Computer-Use AI Agents to Your Own Machine
H Company's open-weight Holo3.1 agents automate desktop and mobile tasks locally, with sizes from 0.8B to 35B and quantized builds that run on consumer hardware.

Hugging Face Hub 1.18 Adds an AI Skills Marketplace to Its CLI
The new huggingface-hub 1.18 release introduces an 'hf skills' command, surfacing a growing marketplace of reusable AI skills for the open-source community.

Hugging Face Turns the Hub Into Agent-First Infrastructure — Every Gradio Space Now Speaks Directly to AI Agents
Hugging Face shipped an /agents.md endpoint on every Gradio Space and elevated Kernels to a first-class repository type in May 2026 — making the open AI stack natively callable by AI coding agents.

Hugging Face Opens the Reachy Mini App Store — 200+ Open-Source Robot Apps for $299
Hugging Face launched an open-source app store for its $299 Reachy Mini robot on May 6, 2026, putting 200+ community-built robotics apps and an AI agent code-generator one click away.

HiDream-O1-Image Goes Open Source — An 8B Reasoning Image Model Lands on Hugging Face
HiDream-AI open-sourced HiDream-O1-Image on May 8, 2026 — an 8-billion parameter reasoning-driven image generation model with a Dev variant and prompt agent, free on Hugging Face.

HyperNova 60B Uses Quantum-Inspired Math to Halve an LLM's Size With Near-Zero Accuracy Loss
Multiverse Computing's free HyperNova 60B compresses a 120B-parameter model by 50% using quantum tensor methods, benchmarking 5x better on tool-calling tasks.

Hugging Face Acquires ggml.ai, Giving llama.cpp a Permanent Open-Source Home
Hugging Face acquires ggml.ai, bringing llama.cpp and the GGUF model format under its umbrella while keeping everything MIT-licensed and open-source for local AI inference.
