Semifly
Semifly · LLMs

Compare large language models

Frontier hosted models and open-weight models you can self-host, side by side. Figures as of July 2026 and refreshed regularly.

Interactive

Head-to-head comparison

Pick any two models — up to four — and see how they stack up across coding, agentic, reasoning, knowledge and context, with the key spec differences called out.

Choose modelspick 2–4
Frontier · API
Open-weight

Capability radar

BenchLM weighted scores (0–100) · June 2026

The Overall score auto-refreshes weekly from BenchLM (July 2026). The five radar axes (coding, agentic, reasoning, knowledge + a normalised context score) are a manually-curated BenchLM snapshot as of June 2026; some open-model axis values are indicative pending full per-benchmark data, and frontier parameter counts are undisclosed. Confirm specifics on each model’s official page.

Frontier

Frontier & proprietary models

Leading hosted models, accessed via API. Scores and prices move quickly; figures are indicative.

ModelLabContextNotableAccess
Claude Fable 5Anthropic1M+Tops the Artificial Analysis Intelligence Index (~65); first public Mythos-class model (Jun 9). Access intermittent under US export controls.API
Claude Opus 4.8Anthropic1M+Highest-scoring widely-available model (AA Index ~61)API
GPT-5.5OpenAI1M+Strong all-round; Pro / Instant variants (AA Index ~60)API
Gemini 3.1 ProGoogle1M+Top reasoning & data analysis (94%+ GPQA Diamond)API
Grok 4.3xAI~2.0MCheapest of the frontier four; strong agentic / tool useAPI
Gemini 3.5 FlashGoogle1MFlagship-level quality at ~4x speed (~$1.50 / 1M in)API

Sources: LLM-Stats, Morph, LM Council (July 2026).

Open source

Open-source & open-weight models

Open-weight models you can download and self-host — run on your own GPUs or on Semifly in one click.

ModelDeveloperParamsContextLicenseDownloadRun
DeepSeek V4 ProDeepSeek~1.6T (MoE)1MMITHugging Face →Run on Semifly
MiniMax M3MiniMax428B (23B active)1MApache 2.0Hugging Face →Run on Semifly
GLM-5.1Zhipu AIMoE200KMITHugging Face →Run on Semifly
Kimi K2.7-CodeMoonshot AI~1T (A32B, MoE)256KMod. MITHugging Face →Run on Semifly
Qwen3.5 (397B-A17B)Alibaba397B (17B active)256KApache 2.0Hugging Face →Run on Semifly
Llama 4 ScoutMetaMoE10MLlama 4Hugging Face →Run on Semifly
Llama 4 MaverickMetaMoE1MLlama 4Hugging Face →Run on Semifly
Mistral Small 4Mistral AI24B256KApache 2.0Hugging Face →Run on Semifly
Gemma 4 (31B)Google31B256KGemmaHugging Face →Run on Semifly

Confirm the current license on each model’s official page before deployment.

Run any of these on Semifly

Tokens & API

Access hosted models through a simple, metered token API.

Get API access →

GPU servers

Buy or lease Supermicro GPU systems to self-host open-weight models.

Browse GPU servers →

AI Foundry

Managed compute for training, fine-tuning, and inference.

Explore AI Foundry →