Live

Local models catch up: 30B open weights now match last year's frontier

Benchmark analysis shows the gap between open 30B models and closed frontier systems has compressed to roughly twelve months, and consumer hardware runs them. Demo content.

NFNiko Ferrar
Published 25 Jun 2026, 04:20Updated 26 Jul 2026, 00:122 min read
Local models catch up: 30B open weights now match last year's frontier

Systematic benchmark tracking shows open-weight models in the 30B class now match the closed frontier of roughly twelve months ago, and they run on hardware ordinary developers own.

The compression

  • Reasoning benchmarks: 30B open models at parity with early-2025 flagships
  • Coding: the gap is smaller still, aided by open training recipes optimized for it
  • Multimodal remains the largest remaining deficit

What runs where

A quantized 30B model runs comfortably on a 24GB consumer GPU or a unified-memory laptop, at speeds usable for interactive work.

bash■ ■ ■
# Typical local setup
ollama run qwen4:32b-q4

For workloads that fit local capability — drafting, extraction, internal search — the marginal cost of inference is now electricity.

Source: Hugging Face

Related stories

All →