Track the latest models, compare frontier and open-source LLMs, and download open-weight models — then run any of them on Semifly with tokens, GPU servers, and AI Foundry.
Every major lab grouped by camp — US vs China, proprietary vs open — with each family’s key versions and years. The China open-weight branch is by far the busiest.
Key versions per lab, hand-curated as of 2026-06. Refreshed as new flagships ship.
Every model placed by quality (Artificial Analysis Intelligence Index) and input price — top-left is the value sweet spot. As of 2026-06-18.
LMArena Elo of the #1 proprietary model (blue) vs the #1 open-weight model (orange), 2023–2026. Open-weight nearly drew level in early 2025; since then the open #1 has been almost entirely Chinese — DeepSeek, Qwen, GLM, Kimi.
Source: LMArena (Arena) Elo via BenchLM leaderboard history. Blue tracks the frontier milestones; orange is the open-weight #1 over time.
OpenRouter routes about 25T tokens a week across 8M+ developers. By real token volume, Chinese open-weight models dominate the usage charts — yet premium US models still capture most of the dollars.
Top models · weekly tokens
By vendor · top-10 aggregate
The dollar–token split: China-origin models take 45%+ of tokens, while Anthropic holds ~12% of tokens but ~46% of dollar spend through premium pricing. Source: OpenRouter rankings (live-scraped) + market reporting. Per-language and per-use-case breakdowns are chart-only on OpenRouter and not yet scrapable.
Updated August 2026 · refreshed regularly
One paragraph of "Lord of the Rings" in, 5,500 lines of code out. Andrej Karpathy had Claude Opus 5 turn Tolkien's opening into a 3D browser scene. The article Unicorn,…
The Decoder →Two research teams independently solved the same open quantum cryptography problem using OpenAI's GPT-5.6 Sol Ultra, submitting their papers just three hours apart. "If…
The Decoder →AI agents, MCP servers, and LLM apps break the core AppSec assumption that applications do what their code says. This guide walks through a practical see-fix-protect…
MarkTechPost →Cogent AI team released Cogent VR-1, a reasoning model post-trained specifically for cybersecurity rather than picking up cyber capability as a side effect of general…
MarkTechPost →Onton, a San Francisco-based search and discovery company, has released Ontology 1, a neurosymbolic model for complex, conversational, multimodal product search. On a…
MarkTechPost →OpenAI's refutation of the Unit Distance Conjecture has sparked a wave of AI-assisted advances in mathematics. Fields Medal winner Timothy Gowers says GPT 5.6 Pro solved…
The Decoder →Ten advances in mathematics and theoretical computer science A few days ago it was Anthropic discovering cryptographic weaknesses with Claude using Mythos Preview,…
Simon Willison →smevals - a small eval suite for evaluating models, prompts, and harnesses I've been working with Jesse Vincent's Prime Radiant applied AI research lab building out this…
Simon Willison →