Skip to content
FrankX.AI

Frontier Intelligence Directory · Updated August 19, 2026

LLM Provider Hub 2026

The decision layer on top of the raw data. Every frontier provider, model, and agentic platform — categorized by capability, priced live, and paired with a verdict. Built for humans and agents.

We cite OpenRouter, Artificial Analysis, and LMArena as sources, and add what they don’t: task-first navigation, the agentic-platform comparison, curated verdicts, and a creator-stack lens.

10

Providers tracked

31

Frontier models

0

Agentic platforms

22

Live-priced

Start here: pick your constraint

The fastest path from “which model?” to an answer. One dominant constraint → a recommendation.

Interactive Simulator

Pick Your Task Constraints

Select what your agent or pipeline demands. The simulator dynamically queries the proving ground receipts to calculate the optimal route.

PROVING GROUND RESOLUTION
Dynamic Router Mode(Select standard options)

Select one or more constraints to trigger routing recommendation.

dynamic-router-integration
import { ACOS_Router } from '@acos/router'; // Initialize ACOS Dynamic Router loaded from /llm-hub.json const router = new ACOS_Router({ env: 'production', failover: true }); // Route task dynamically based on constraint requirements const task = { prompt: "Synthesize weekly content and emit output strictly as standard JSON", constraints: [] }; const result = await router.execute({ task, // Dynamic resolver picks Dynamic Router Mode based on constraints fallbackModel: 'claude-sonnet-5' }); console.log(`Routed to ${result.model} | Status: ${result.status}`);
Target endpoints synced automatically hourly/llm-hub.json

Curated Routing Table

If you need…PickRunner-upWhy
Hardest reasoning + knowledge workClaude Opus 4.8GPT-5.5Tops the intelligence index — GDPval-AA 1890 and SWE-Bench Pro 69.2% lead the field.
Agentic codingClaude Fable 5GPT-5.5New launch ceiling — 95% SWE-Bench Verified, ~80% SWE-Bench Pro vs GPT-5.5’s 58.6% (vendor-claimed).
Computer use + terminal autonomyGPT-5.5Claude Opus 4.8Best published computer-use scores (84.9% GDPval, 78.7% OSWorld, 98% Tau2 Telecom).
Long-running xAI agentsGrok 4.6Grok 4.3Current xAI flagship (AA Index 61, vendor/AA, 12 Aug 2026). Same $2/$6 list under 200k as 4.5; no SIS arena receipt yet.
Lowest cost (closed frontier)Grok 4.3Gemini 3.5 FlashStill the cheap Grok tier at $1.25/$2.50. Grok 4.6 is the flagship, not the bargain SKU.
Top open weightsKimi K2.6DeepSeek V4Highest open-weights intelligence (AA Index 54); DeepSeek V4 is the close, MIT-licensed runner-up.
Lowest cost (open weights)DeepSeek V4gpt-oss (120b / 20b)Frontier-class coding (80.6% SWE-bench Verified) at open-weight economics under MIT.
Longest contextGrok 4.3GPT-5.52M-token native window; GPT-5.5 offers 1M at GA.
Native voice + broad multimodalGPT-5.5Gemini 3.5 ProNative audio modality plus the widest general multimodal coverage.
Widest modality (incl. video)Gemini 3.5 ProGemini 3.5 FlashGoogle’s top reasoning tier across text/vision/audio/video (Pro in preview; Flash is the GA workhorse).
EU data sovereigntyMistral Large 3DeepSeek V4Apache 2.0, EU-resident endpoints, self-hostable frontier on a single 8×H200 node.
Self-host / own the weightsDeepSeek V4Llama 4 MaverickOpen-weight MoE frontier; Llama 4 for a permissive license + native multimodality.
Run on one consumer GPUGemma 4gpt-oss (120b / 20b)Gemma 4 31B runs in ~18GB at Q4 (LMArena 1452); gpt-oss-20b is the ~16GB reasoning alternative.
Laptop / edge (smallest footprint)Microsoft Phi-4 (open-weight family)Gemma 4MIT-licensed 3.8B–15B STEM specialist that runs on a laptop; Gemma 4’s E2B/E4B tiers go smaller still.
Interactive Model ROI Engine

Cost-to-Outcome Calculator

Model intelligence pricing varies up to 40x between tiers. Simulate monthly token spend across frontier models and evaluate hybrid routing savings.

Monthly Prompt Input Tokens10 Million
1M (Light)25M50M100M (Heavy)
Monthly Generated Output Tokens2 Million
0.2M5M10M20M (Heavy Agentic)

Architect Dynamic Routing Strategy

Hybrid 80/20 Architecture: $32.00/mo

Route 80% volume to Fast-Path (Gemini 3.7 Flash) + 20% to Deep-Reason (Claude Opus 5). Saves $68.00/mo (68%) vs 100% flagship.

Estimated Blended ROI
Save 68%

DeepSeek V4 Pro

DeepSeek

Lowest Cost

Cheapest frontier-class coding and agentic reasoning

Estimated Monthly:$4.90
In: $2.70Out: $2.20Speed: 140 t/s

Gemini 3.7 Flash

Google

Blazing speed leader with hybrid thinking

Estimated Monthly:$15.00
In: $7.50Out: $7.50Speed: 340+ t/s

Grok 4.6

xAI

Real-time grounded orchestration and agent swarms

Estimated Monthly:$32.00
In: $20.00Out: $12.00Speed: 110 t/s

Claude Sonnet 5

Anthropic

Balanced daily driver for coding and analysis

Estimated Monthly:$40.00
In: $20.00Out: $20.00Speed: 90 t/s

GPT-5.6 Sol

OpenAI

Unified frontier reasoning and multi-modal synthesis

Estimated Monthly:$90.00
In: $50.00Out: $40.00Speed: 85 t/s

Claude Opus 5

Anthropic

Elite situational judgment, architecture & deep code craft

Estimated Monthly:$100.00
In: $50.00Out: $50.00Speed: 75 t/s

Claude Fable 5

Anthropic

Mythos-class ceiling for long-horizon constraint precision

Estimated Monthly:$200.00
In: $100.00Out: $100.00Speed: 70 t/s

Browse by capability

Pick the job, jump to the providers that lead.

Reasoning & Analysis

Complex problem-solving, math, abstract reasoning, long-horizon planning

No tracked providers yet

Multimodal Understanding

Vision, document, chart, and cross-modal reasoning across text/image/audio

No tracked providers yet

Video Generation

Generative video models, text-to-video, image-to-video, editing

No tracked providers yet

Coding & Engineering

Agentic coding, terminal use, debugging, multi-file refactors

No tracked providers yet

Agentic Infrastructure

Tool use, function calling, agent SDKs, computer use, long-horizon execution

No tracked providers yet

Voice & Audio

Native speech in/out, real-time conversation, audio understanding

No tracked providers yet

Image Generation

Text-to-image, editing, in-painting, brand-consistent generation

No tracked providers yet

Model explorer

Sort and filter every tracked model. Live pricing via OpenRouter where available. Click a model for the full breakdown.

31 models live pricing via OpenRouter

Qwen3.8-27BAlibaba (Qwen)2026-08-14262KOpenOpen
Gemini 3.7 FlashGoogle DeepMind2026-08-131.0M$0.75$3.75
DeepSeek V4 Pro 0813DeepSeek2026-08-131.0M$0.66$1.98
Grok 4.6xAI2026-08-12500K$2.00$6.00
Claude Opus 5Anthropic2026-07-241M$5.00$25.00
Claude Sonnet 5Anthropic2026-06-301M$2.00$10.00
Claude Fable 5Anthropic2026-06-091M$10.00$50.00
MAI-Thinking-1Microsoft AI2026-06-02256K
MAI-Image-2.5Microsoft AI2026-06-02
MAI-Code-1-FlashMicrosoft AI2026-06-02
Claude Opus 4.8Anthropic2026-05-281M$5.00$25.00
Gemini 3.5 FlashGoogle DeepMind2026-05-191.0M$1.50$9.00
Qwen3.7-MaxAlibaba (Qwen)2026-05-191M$1.48$4.43
Grok 4.3xAI2026-04-301M$1.25$2.50
DeepSeek V4DeepSeek2026-04-241.0M$0.87$1.74
GPT-5.5OpenAI2026-04-231.1M$5.00$30.00
Kimi K2.6Moonshot AI2026-04-20262K$0.95$4.00
Gemma 4Google DeepMind2026-04-02256KOpenOpen
Claude Sonnet 4.6Anthropic2026-02-171M$3.00$15.00
Claude Opus 4.6Anthropic2026-02-051M$5.00$25.00
GPT-5.2 ProOpenAI2026-01-01400K$21.00$168.00
Mistral Large 3Mistral AI2025-12-02262K$0.50$1.50
Gemini 3 ProGoogle DeepMind2025-12-012M
Claude Opus 4.5Anthropic2025-11-01200K$5.00$25.00
Grok 4.1xAI2025-11-012M
Claude Haiku 4.5Anthropic2025-10-01200K$1.00$5.00
Claude Sonnet 4.5Anthropic2025-09-291M$3.00$15.00
gpt-oss (120b / 20b)OpenAI2025-08-05131K$0.04$0.17
Llama 4 MaverickMeta AI2025-04-051.0M$0.20$0.70
Microsoft Phi-4 (open-weight family)Microsoft AI2024-12-1216K$0.07$0.14
Gemini 3.5 ProGoogle DeepMind

Creator stacks

Creators don’t pick a model — they assemble a stack. Here’s what to use across each modality, and the workflow.

Agentic platforms

Where the models actually do work — IDEs, CLIs, desktops, agent platforms, managed runtimes. The layer the data sites skip.

Frequently asked

What is the best LLM in 2026?+

There is no single winner. As of 14 August 2026, Grok 4.6 is the current xAI flagship and scores 61 on the Artificial Analysis Intelligence Index, matching GPT-5.6 Sol on that composite. Other seats still depend on the task — see the decision matrix and dated model pages rather than a global crown.

How is this different from OpenRouter or Artificial Analysis?+

Those are the raw-data sources — OpenRouter for live pricing and routing, Artificial Analysis for independent benchmarks, LMArena for human preference. We cite all three. The FrankX LLM Hub adds the decision layer they don’t: task-first navigation, the agentic-platform comparison (Claude Code vs Antigravity vs Cursor vs Codex), curated verdicts, and a creator-stack lens — for humans and agents.

Which is the cheapest frontier reasoning model?+

DeepSeek V3.2 leads on pure cost ($0.27 / $1.10 per 1M tokens, MIT license). Gemini 3.5 Flash is the cheapest closed-frontier option at $0.30 / $2.50. Both deliver frontier-class reasoning for production agentic workloads.

What is the best agentic LLM in 2026?+

By category: coding agents — Gemini 3.5 Flash (76.2% Terminal-Bench 2.1) and Claude Opus 4.6; long-horizon enterprise — Gemini Spark and Claude Agent Teams; computer-use — GPT-5.2 Operator and Claude Opus 4.6 (72.7% OSWorld).

Is the pricing live?+

Where a model maps to OpenRouter, pricing is fetched live (hourly) and marked with a ⚡ icon and "via OpenRouter." Otherwise it comes from our curated registry. Always verify against the provider before relying on it for billing.

Can AI agents consume this hub?+

Yes. The full curated dataset — models, pricing, verdicts, decision matrix, comparisons — is available as clean JSON at /llm-hub.json, plus JSON-LD structured data on every page and deep links in /llms.txt.

FrankX Intelligence Pipeline · Last refreshed May 20, 2026

Source of truth: data/model-registry.json · Agent surface: /llm-hub.json · Add a model via /new-model