Local models catch up: 30B open weights now match last year's frontier
Benchmark analysis shows the gap between open 30B models and closed frontier systems has compressed to roughly twelve months, and consumer hardware runs them. Demo content.
Systematic benchmark tracking shows open-weight models in the 30B class now match the closed frontier of roughly twelve months ago, and they run on hardware ordinary developers own.
The compression
- ▸Reasoning benchmarks: 30B open models at parity with early-2025 flagships
- ▸Coding: the gap is smaller still, aided by open training recipes optimized for it
- ▸Multimodal remains the largest remaining deficit
What runs where
A quantized 30B model runs comfortably on a 24GB consumer GPU or a unified-memory laptop, at speeds usable for interactive work.
# Typical local setup
ollama run qwen4:32b-q4For workloads that fit local capability — drafting, extraction, internal search — the marginal cost of inference is now electricity.
Source: Hugging Face


