InferenceIndexerInferenceIndexer.ai
/
ModelsProvidersAPIEmbedFor agentsHarnessesMethodologyAboutLoginSign Up

AI inference recommendation engine

Models indexed677Providers indexed75Prices sourcedhourly
Coverage·What is verified today, and what comes next
PriceVerified

Composite and per-model prices pulled directly from provider APIs and rebuilt every hour.

75providers polled hourly
QualityVerified: intelligence

Intelligence verified against the AA index, then divided into price so rankings reward value, not cheapness.

274models quality-adjusted
In developmentCriteria published before any provider is rated against them.
Privacy

Whether a provider keeps your tokens, for how long, and under whose jurisdiction. Matched on provider statements today.

Security

What could go wrong using this inference, stated in terms that can be checked. Not yet scored.

Who attests each claim today, and who verifies it

The claimProvider attestedVerified by
Price per million tokensThe providerInference Indexer
Model quality and intelligenceThe providerInference Indexer
Model identity and quantizationThe providerIn development
Data retention and privacyThe providerIn development
Security postureThe providerIn development
Standard Inference Token price·Verified hourly, pulled directly from provider APIs
$1.35/ M tokens
SIT index↑ 0.4%today
7 day↓ 11.1%
30 day↑ 26.6%
90 day↓ 8.4%

The Standard Inference Token (SIT) tracks the cost of producing one million GPT-4-Turbo-equivalent inference tokens, the commodity unit for AI compute. Basket: cheapest qualifying model per provider, eligibility set relative to the scored-model population (top 40%), based on Artificial Analysis v4.3.

→ Read methodology
31-day spot
low $1.03 · high $1.85API keys →
v4.2 ↻v4.3 ↻
$1.88
$1.44
$1.00
2026-08-302026-09-14today
era break (basket reconstitution)green = price down/red = price up

Quality-adjusted price across 677 models

Grouped by quality tier, ranked within tier by Cost/IQ: verified price per million tokens per unit of AA Intelligence Index. Input, output, and blended prices per model.

Full rankings →
#ModelCreatorBasisInput $/MOutput $/MBlended $/MCost / IQ
🥇Meta: Muse Spark 1.3 Contributor🇺🇸 MetaPrice: verified$0.10$0.20$0.160.13🥈Meta: Muse Spark 1.3🇺🇸 MetaPrice: verified$1.25$4.25$3.052.54🥉OpenAI: GPT-6 Sol Pro🇺🇸 OpenAIPrice: verified$1.50$7.50$5.104.29Anthropic: Claude Sonnet 5.5🇺🇸 AnthropicPrice: verified$2.00$10.00$6.804.86OpenAI: GPT-6 Sol🇺🇸 OpenAIPrice: verified$2.20$11.00$7.486.30Anthropic: Claude Opus 5.5🇺🇸 AnthropicPrice: verified$4.00$20.00$13.609.44Anthropic: Claude Opus 5🇺🇸 AnthropicPrice: verified$5.00$25.00$17.0013.39OpenAI: GPT-6 Astra Pro🇺🇸 OpenAIPrice: verified$7.50$37.50$25.5019.36Anthropic: Claude Fable 5.1🇺🇸 AnthropicPrice: verified$10.00$50.00$34.0025.49OpenAI: GPT-6 Astra🇺🇸 OpenAIPrice: verified$10.00$50.00$34.0025.82

Quality here means intelligence only. Latency and uptime verification is in development; neither is scored on this page.

Subscribed to by the teams that buy inference

Create a free account for unlimited recommendations and API access. The engine on this page is the same API your agents can call.

Create a free accountView API documentation

Free for 1,000 requests/day. No credit card required.

InferenceIndexerInferenceIndexer.ai · Inference Recommendation Engine
ProvidersModel TypeMethodologyAPI DocsFor AgentsHarnessesData QualityAboutPrivacy PolicyTerms of Service@inferenceindex
677 models · 75 providers · Last updated: 2026-09-29 06:00 UTC