The Chinese lab continues its cost-disruption strategy: V4 base model weights land on Hugging Face under a permissive license while hosted inference gets dramatically cheaper. Demo content.
Open-weight models now trail frontier proprietary systems by single-digit percentages on everyday tasks, while costing four to ten times less per token.