# Tokenomy > Tokenomy is the Economic Intelligence Layer for AI — the platform for AI Economics, the discipline of understanding how intelligence is produced, consumed, priced, allocated, optimized, governed, and monetized. Tokenomy operates the six-stage AI Economics lifecycle: Measure, Understand, Control, Optimize, Monetize, Transact. ## Category AI Economics — the discipline of running AI as an economy. Tokenomy is the platform (operate), the Academy (teach), the Library (organize), and Research (advance) for this category. ## The AI Economics lifecycle - **Measure** — Instrument every token, call, agent, and outcome. - **Understand** — See where cost, waste, and margin actually live. - **Control** — Enforce budgets, policies, and guardrails before spend happens. - **Optimize** — Route, cache, right-size, and prune every workload. - **Monetize** — Price AI products, protect margin, run chargeback. - **Transact** — Metering, settlement, and commerce for agents. ## Primary pages - [Homepage](/): Economic Intelligence Layer for AI. - [About](/about): Mission, team, and product story. - [Pricing](/pricing): Plans and pricing tiers. - [Contact](/contact): Get in touch with the team. - [Documentation](/documentation): Developer and user docs. ## AI Economics — the discipline - [AI Economics Academy](/academy): Curriculum organized by lifecycle stage — Foundations, Measure, Understand, Control, Optimize, Monetize, Transact. - [Library](/library): Curated third-party news, benchmarks, papers, and events, organized by lifecycle stage. - [Tokenomy Research](/research): Original research franchises published on a regular cadence. - [AI Economics Assessment](/waste): Free scanner that finds waste in your current AI stack. ## Tokenomy Research — original franchises - [AI Economics Index](/research/ai-economics-index): Monthly headline gauge of AI cost, margin, and efficiency. - [Model Economic Efficiency Benchmark](/research/model-economic-efficiency): $/useful-token and $/successful-task across frontier models. - [Agent Economic Efficiency Benchmark](/research/agent-economic-efficiency): Same lens applied to agent harnesses. - [State of AI Economics](/research/state-of-ai-economics): Quarterly flagship report. - [AI Pricing Observatory](/research/ai-pricing-observatory): Tracked provider price changes and elasticity. - [Agent Commerce Research](/research/agent-commerce): Metering, settlement, take-rate. ## Platform — tools by lifecycle stage - Measure: [Token Calculator](/tools/token-calculator), [Token Observability](/tools/token-observability), [Energy Estimator](/tools/energy-usage-estimator), [GPU Monitoring](/tools/gpu-monitoring), [Token Ledger](/enterprise/token-ledger), [Token Energy Ledger](/enterprise/token-energy-ledger), [Token Yield Score](/enterprise/token-yield-score), [API Keys](/app/keys). - Understand: [Cost Explorer](/app/cost-explorer), [Find Your AI Waste](/waste), [Book an AI Cost Audit](/audit), [Prompt Visualizer](/tools/prompt-visualizer), [Token Leaderboard](/tools/token-leaderboard), [Context Waste Scanner](/enterprise/context-waste-scanner), [Agent Cost Forecast](/enterprise/agent-cost-forecast), [Agent Harness Profiler](/enterprise/agent-harness-profiler), [Dev Token Coach](/enterprise/dev-token-coach), [Token Economy Benchmark](/enterprise/token-economy-benchmark). - Control: [Alerts](/app/alerts), [Policy-as-Budget Router](/enterprise/policy-budget-router-v2), [AI Budget Gate](/enterprise/ai-budget-gate), [Token Anomaly Sentinel](/enterprise/token-anomaly-sentinel). - Optimize: [Cost Optimization Suite](/tools/cost-optimization), [Alternatives Explorer](/tools/alternatives-explorer), [Model Right-Sizing Lab](/enterprise/model-right-sizing-lab), [Cache ROI Optimizer](/enterprise/cache-roi-optimizer). - Monetize: [Prompt Margin Analyzer](/enterprise/prompt-margin-analyzer). - Transact: [Agent Commerce Meter](/enterprise/agent-commerce-meter). ## Enterprise - [Enterprise Hub](/enterprise): All 15 enterprise tools grouped by lifecycle stage. - [Enterprise Observability](/enterprise-observability): Cross-team telemetry, SLOs, and policy enforcement. ## Library (curated) - [AI News Hub](/research/ai-news-hub) - [Model Benchmarks (external)](/research/model-benchmarks) - [Research Papers](/research/research-papers) - [Innovation Tracker](/research/innovation-tracker) - [Conference Calendar](/research/conference-calendar) - [Tiktoken & Tokenization Guide](/research/tiktoken-guide) ## Community - [Community](/community) ## Machine-readable formats Every page below is also served as markdown at the same path with a `.md` suffix. The whole corpus in one file: https://tokenomy.ai/llms-full.txt (updated 2026-09-02). ## Live data endpoints (no login, CORS-open) Tokenomy publishes the AI model pricing corpus as a machine-readable API. Use `X-API-Key: tkm_demo` for anonymous/agent access. - `GET https://dakfcntliydkpzbmiurt.supabase.co/functions/v1/data-api/openapi.json` — OpenAPI 3.1 description of every route. - `GET https://dakfcntliydkpzbmiurt.supabase.co/functions/v1/data-api/models` — full catalog: id, provider, context window, input/output price per 1M tokens. - `GET https://dakfcntliydkpzbmiurt.supabase.co/functions/v1/data-api/models/{id}` — one model. - `GET https://dakfcntliydkpzbmiurt.supabase.co/functions/v1/data-api/pricing/history?model={id}&days=90` — daily price snapshots. - `GET https://dakfcntliydkpzbmiurt.supabase.co/functions/v1/blog-rss` — RSS 2.0 feed of newly published analysis. MCP (Model Context Protocol) server for agents: `https://tokenomy.ai/mcp` — discovery at `https://tokenomy.ai/.well-known/mcp.json`. Tools: list-models, estimate-cost, recent-usage, check-budgets, list-api-keys. ## Full page index (101 pages, updated 2026-09-02) - [Tokenomy — FinOps for AI: Meter, Budget & Route LLM Spend](/): Tokenomy is the FinOps platform for LLMs and AI agents. See where every AI dollar goes — and get 20–40% of them back. BYOK across OpenAI, Anthropic, Google, xAI and open-weight endpoints. — markdown: /index.md - [FinOps for AI — The Category, Explained | Tokenomy](/finops-for-ai): FinOps for AI is the discipline of measuring, attributing and optimizing every LLM and agent dollar. See how Tokenomy operationalizes it with runtime rails, budgets, routing and chargeback. — markdown: /finops-for-ai.md - [Pricing — Tokenomy FinOps for AI](/pricing): Simple pricing indexed to AI spend under management. Free tier, Starter $9/mo, Pro $39/mo, Team/Business $499/mo, Enterprise from $50k/yr. BYOK — provider spend stays on your accounts. — markdown: /pricing.md - [Platform Features — Tokenomy FinOps for AI](/features): Runtime rails for the token economy: metering proxy, smart router, budget guard, unified cost graph, policy engine, chargeback, SLO monitoring, MCP server and agent commerce rails. — markdown: /features.md - [Trust & Security — Tokenomy](/trust): Security, privacy and compliance at Tokenomy. BYOK isolation, per-tenant encryption, RLS-enforced multi-tenancy, SOC 2 in progress (Q4 2026), GDPR and DPA available. — markdown: /trust.md - [Team — who is behind Tokenomy](/team): The people accountable for Tokenomy's AI pricing data, methodology and platform. Named humans, direct contact, published methodology. — markdown: /team.md - [About Tokenomy — The Economic Intelligence Layer for AI](/about): Tokenomy is the Economic Intelligence Layer for AI. BYOK metering, budgets, smart routing and chargeback for teams shipping agents in production. — markdown: /about.md - [Documentation — Tokenomy FinOps for AI](/documentation): Deploy the Tokenomy metering proxy, smart router and budget guard in a day. BYOK guides for OpenAI, Anthropic, Google, xAI, OpenRouter and Ollama. MCP server for ChatGPT, Claude and Cursor. — markdown: /documentation.md - [Contact Tokenomy](/contact): Talk to Tokenomy about FinOps for AI, the AI Economics Assessment, enterprise deployments and security review. — markdown: /contact.md - [AI Economics Index — Tokenomy Research](/research): The AI Economics Index tracks model efficiency, pricing, agent economics and the state of AI economics. Public research franchises from Tokenomy. — markdown: /research.md - [Attention Evolution Lab — GPT-2 to Kimi K3, Priced Out | Tokenomy](/tools/attention-evolution-lab): Free interactive lab: step through linear attention, DeltaNet, gated delta, Kimi Delta Attention, MLA and MoE from GPT-2 to Kimi K3 and see live KV cache, decode throughput and cost per million tokens. — markdown: /tools/attention-evolution-lab.md - [AI Token Calculator — Estimate LLM Cost per Prompt | Tokenomy](/tools/token-calculator): Free AI token calculator. Count tokens and estimate cost per prompt across GPT-5, Claude, Gemini, xAI Grok, Llama and other frontier LLMs — with July 2026 pricing. — markdown: /tools/token-calculator.md - [Token Observability — Real-Time LLM Cost & Latency | Tokenomy](/tools/token-observability): Real-time observability for every LLM call. Track cost, latency, tokens, cache hits and error rates across providers, models, workspaces, customers and agents. — markdown: /tools/token-observability.md - [Token Speed Simulator — LLM Latency & Throughput | Tokenomy](/tools/token-speed-simulator): Simulate LLM output speed. Compare tokens-per-second, time-to-first-token and total latency across GPT-5, Claude, Gemini, Grok, Llama and open-weight endpoints. — markdown: /tools/token-speed-simulator.md - [LLM Memory Calculator — VRAM & KV Cache Sizing | Tokenomy](/tools/memory-calculator): Calculate VRAM, KV cache and system memory for self-hosted LLM inference. Size hardware for Llama, Mistral, Qwen, DeepSeek and other open-weight models. — markdown: /tools/memory-calculator.md - [AI Energy Usage Estimator — Watts, kWh & CO₂ per LLM Call](/tools/energy-usage-estimator): Estimate the energy and CO₂ footprint of any LLM workload. Compare hosted APIs and self-hosted open-weight inference on H100, H200, B200 and MI300X. — markdown: /tools/energy-usage-estimator.md - [AI Content Detector — Spot LLM-Generated Text | Tokenomy](/tools/ai-content-detector): Free AI content detector. Identify whether text was likely generated by GPT-5, Claude, Gemini or other LLMs — with confidence scoring and per-passage analysis. — markdown: /tools/ai-content-detector.md - [GPU Throughput Monitor — Real-Time AI GPU Metrics | Tokenomy](/tools/gpu-monitoring): Monitor GPU utilization, VRAM, temperature and throughput for AI training and inference. Real-time dashboard for H100, H200, B200, MI300X fleets. — markdown: /tools/gpu-monitoring.md - [Token Leaderboard — Compare LLM Cost, Speed & Quality | Tokenomy](/tools/token-leaderboard): Live leaderboard ranking frontier LLMs by cost per useful token, latency, throughput and quality. Updated every 12 hours across 400+ hosted and open-weight models. — markdown: /tools/token-leaderboard.md - [Prompt Processing Visualizer — See How LLMs Read Prompts | Tokenomy](/tools/prompt-visualizer): Visualize how LLMs tokenize, attend to and process your prompt step by step. See tokenization, attention hints and estimated cost per stage. — markdown: /tools/prompt-visualizer.md - [LLM Alternatives Explorer — Cheaper, Faster, Open-Weight Swaps](/tools/alternatives-explorer): Find cheaper, faster or open-weight alternatives to any LLM. Compare hardware, model variants and inference strategies with total-cost-of-ownership modeling. — markdown: /tools/alternatives-explorer.md - [AI Cost Optimization Suite — Cut LLM Spend 20–40% | Tokenomy](/tools/cost-optimization): Multi-provider comparator, prompt library, model router simulator, batch processing optimizer and cost tracking dashboard. Everything you need to cut LLM spend 20–40%. — markdown: /tools/cost-optimization.md - [Unified Cost Graph — Tokenomy FinOps for AI](/features/unified-cost-graph): One cost graph across every provider, model, workspace, customer, agent and environment. Real attribution from the Tokenomy usage ledger. — markdown: /features/unified-cost-graph.md - [Policy & Budget Routing — Tokenomy FinOps for AI](/features/policy-budget-routing): Enforce spend policies before a request hits a paid provider. Route by cost, quality and latency; block or throttle at budget. — markdown: /features/policy-budget-routing.md - [Agent Commerce Rails — Tokenomy FinOps for AI](/features/agent-commerce-rails): Let agents transact, meter, budget and settle spend. MCP server, per-agent budgets and Stripe-metered billing built in. — markdown: /features/agent-commerce-rails.md - [Advanced Telemetry — Tokenomy FinOps for AI](/features/advanced-telemetry): OpenTelemetry-native telemetry for LLM traffic. Traces, spans and metrics per request with cost, latency and cache attribution. — markdown: /features/advanced-telemetry.md - [Token Flow Visualizer — Tokenomy FinOps for AI](/features/token-flow-visualizer): Visualize every hop a token takes — from prompt to router to provider to cache to response — with cost and latency per stage. — markdown: /features/token-flow-visualizer.md - [SLO Monitoring — Tokenomy FinOps for AI](/features/slo-monitoring): Define and track SLOs for LLM latency, error rate and cost. Alert on burn-rate breaches with automatic failover. — markdown: /features/slo-monitoring.md - [Policy Governance — Tokenomy FinOps for AI](/features/policy-governance): Security-review-ready governance for LLMs and agents. PII policies, jailbreak controls, model allow-lists and full audit logs. — markdown: /features/policy-governance.md - [Route Health — Tokenomy FinOps for AI](/features/route-health): Real-time health of every provider and model route. Automatic failover on latency spikes, error surges or provider incidents. — markdown: /features/route-health.md - [Billing & Revenue — Tokenomy FinOps for AI](/features/billing-revenue): Turn LLM usage into revenue. Stripe metered billing, chargeback exports, per-customer invoicing and margin dashboards. — markdown: /features/billing-revenue.md - [Free AI Cost & Token Tools — Tokenomy](/tools): Free calculators and estimators for AI teams: token cost, throughput, VRAM, energy, model alternatives and cost optimization. No sign-up to try. — markdown: /tools.md - [AI Economics Research — Index, Benchmarks & Reports | Tokenomy](/research): Tokenomy Research: the AI Economics Index, model and agent economic efficiency, the AI Pricing Observatory and the State of AI Economics report. — markdown: /research.md - [AI Economics Academy — Learn FinOps for AI | Tokenomy](/academy): Free curriculum on AI economics: unit economics of tokens, routing strategy, caching ROI, budget guardrails and chargeback models for AI teams. — markdown: /academy.md - [AI Economics Library — Curated Sources & Methods | Tokenomy](/library): A curated library of pricing pages, benchmarks, papers and methodology behind Tokenomy's AI economics research and cost models. — markdown: /library.md - [Tokenomy Blog — AI Economics & FinOps for AI](/blog): Essays on AI economics, token unit economics, model routing, caching ROI and building the financial control layer for AI systems. — markdown: /blog.md - [Tokenomy Community — AI Cost & FinOps Discussions](/community): Ask questions and share benchmarks with engineers and finance teams working on AI cost, routing, observability and unit economics. — markdown: /community.md - [Find Your AI Waste — Free Spend Scanner | Tokenomy](/waste): Free scanner that estimates where your LLM spend is leaking: oversized models, uncached prompts, retry storms and context bloat. — markdown: /waste.md - [AI Pricing Data API — Model Prices & Benchmarks | Tokenomy](/data-api): Programmatic access to Tokenomy's AI pricing dataset: list prices, context windows, throughput and quality benchmarks for 470+ models, refreshed twice daily. — markdown: /data-api.md - [State of AI Economics — Tokenomy Report](/reports/state-of-ai-economics): The State of AI Economics report: how AI unit costs, model pricing, throughput and spend efficiency are moving across the industry. — markdown: /reports/state-of-ai-economics.md - [The Economic Intelligence Layer for AI — Tokenomy](/blog/economic-intelligence-layer-for-ai): Why AI needs an economic intelligence layer: metering, attribution, routing and governance for every token your systems spend. — markdown: /blog/economic-intelligence-layer-for-ai.md - [AI News Hub — Tokenomy Research](/research/ai-news-hub): Curated AI and robotics news with an economics lens — model launches, price changes and capability shifts that move your cost model. — markdown: /research/ai-news-hub.md - [Model Benchmarks — Tokenomy Research](/research/model-benchmarks): Quality, throughput and price benchmarks across frontier and open models — SWE-bench Verified, MMLU-Pro, tokens/sec and $/1M tokens. — markdown: /research/model-benchmarks.md - [Research Papers — Tokenomy Research](/research/research-papers): Key papers on inference efficiency, attention architectures, quantization, caching and the economics of large-model serving. — markdown: /research/research-papers.md - [Innovation Tracker — Tokenomy Research](/research/innovation-tracker): Tracking architectural and hardware innovations that change the cost of inference — MLA, MoE, speculative decoding, new accelerators. — markdown: /research/innovation-tracker.md - [AI Conference Calendar — Tokenomy Research](/research/conference-calendar): Upcoming AI research and infrastructure conferences, deadlines and industry events relevant to AI economics and inference. — markdown: /research/conference-calendar.md - [Tiktoken Guide — Tokenomy Research](/research/tiktoken-guide): A practical guide to tokenization and tiktoken: how text becomes tokens, why counts differ per model and how it drives your bill. — markdown: /research/tiktoken-guide.md - [AI Model Pricing Directory — Every Model, Per 1M Tokens | Tokenomy](/models): Live input and output pricing per 1M tokens for every major AI model — OpenAI, Anthropic, Google, xAI, Meta, DeepSeek, Mistral, Qwen and more. Free, updated twice daily. — markdown: /models.md - [gpt-5.6-sol Pricing — $4.50/1M in, $18.00/1M out | Tokenomy](/models/gpt-5-6-sol): gpt-5.6-sol costs $4.50 per 1M input tokens and $18.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gpt-5-6-sol.md - [claude-fable-5 Pricing — $5.00/1M in, $22.00/1M out | Tokenomy](/models/claude-fable-5): claude-fable-5 costs $5.00 per 1M input tokens and $22.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/claude-fable-5.md - [claude-haiku-4.5 Pricing — $0.50/1M in, $2.50/1M out | Tokenomy](/models/claude-haiku-4-5): claude-haiku-4.5 costs $0.50 per 1M input tokens and $2.50 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/claude-haiku-4-5.md - [gemini-3.5-pro Pricing — $2.50/1M in, $12.00/1M out | Tokenomy](/models/gemini-3-5-pro): gemini-3.5-pro costs $2.50 per 1M input tokens and $12.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gemini-3-5-pro.md - [gemini-3.5-flash Pricing — $0.12/1M in, $0.50/1M out | Tokenomy](/models/gemini-3-5-flash): gemini-3.5-flash costs $0.12 per 1M input tokens and $0.50 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gemini-3-5-flash.md - [grok-4 Pricing — $3.50/1M in, $14.00/1M out | Tokenomy](/models/grok-4): grok-4 costs $3.50 per 1M input tokens and $14.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/grok-4.md - [llama-4.1-405b Pricing — $2.80/1M in, $8.00/1M out | Tokenomy](/models/llama-4-1-405b): llama-4.1-405b costs $2.80 per 1M input tokens and $8.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/llama-4-1-405b.md - [glm-5.2 Pricing — $0.35/1M in, $1.40/1M out | Tokenomy](/models/glm-5-2): glm-5.2 costs $0.35 per 1M input tokens and $1.40 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/glm-5-2.md - [gpt-5.2 Pricing — $2.00/1M in, $8.00/1M out | Tokenomy](/models/gpt-5-2): gpt-5.2 costs $2.00 per 1M input tokens and $8.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gpt-5-2.md - [gpt-5 Pricing — $3.00/1M in, $12.00/1M out | Tokenomy](/models/gpt-5): gpt-5 costs $3.00 per 1M input tokens and $12.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gpt-5.md - [gpt-5-mini Pricing — $0.80/1M in, $3.20/1M out | Tokenomy](/models/gpt-5-mini): gpt-5-mini costs $0.80 per 1M input tokens and $3.20 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gpt-5-mini.md - [gpt-5-nano Pricing — $0.30/1M in, $1.20/1M out | Tokenomy](/models/gpt-5-nano): gpt-5-nano costs $0.30 per 1M input tokens and $1.20 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gpt-5-nano.md - [gpt-4.1 Pricing — $2.00/1M in, $8.00/1M out | Tokenomy](/models/gpt-4-1): gpt-4.1 costs $2.00 per 1M input tokens and $8.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gpt-4-1.md - [gpt-4.1-mini Pricing — $0.50/1M in, $2.00/1M out | Tokenomy](/models/gpt-4-1-mini): gpt-4.1-mini costs $0.50 per 1M input tokens and $2.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gpt-4-1-mini.md - [o4 Pricing — $8.00/1M in, $32.00/1M out | Tokenomy](/models/o4): o4 costs $8.00 per 1M input tokens and $32.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/o4.md - [o4-mini Pricing — $2.00/1M in, $8.00/1M out | Tokenomy](/models/o4-mini): o4-mini costs $2.00 per 1M input tokens and $8.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/o4-mini.md - [o3 Pricing — $10.00/1M in, $40.00/1M out | Tokenomy](/models/o3): o3 costs $10.00 per 1M input tokens and $40.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/o3.md - [gpt-4o Pricing — $5.00/1M in, $15.00/1M out | Tokenomy](/models/gpt-4o): gpt-4o costs $5.00 per 1M input tokens and $15.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gpt-4o.md - [gpt-4o-mini Pricing — $0.50/1M in, $1.50/1M out | Tokenomy](/models/gpt-4o-mini): gpt-4o-mini costs $0.50 per 1M input tokens and $1.50 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gpt-4o-mini.md - [claude-opus-4 Pricing — $15.00/1M in, $75.00/1M out | Tokenomy](/models/claude-opus-4): claude-opus-4 costs $15.00 per 1M input tokens and $75.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/claude-opus-4.md - [claude-sonnet-4 Pricing — $3.00/1M in, $15.00/1M out | Tokenomy](/models/claude-sonnet-4): claude-sonnet-4 costs $3.00 per 1M input tokens and $15.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/claude-sonnet-4.md - [claude-haiku-4 Pricing — $0.20/1M in, $1.00/1M out | Tokenomy](/models/claude-haiku-4): claude-haiku-4 costs $0.20 per 1M input tokens and $1.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/claude-haiku-4.md - [claude-3.5-sonnet Pricing — $3.00/1M in, $15.00/1M out | Tokenomy](/models/claude-3-5-sonnet): claude-3.5-sonnet costs $3.00 per 1M input tokens and $15.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/claude-3-5-sonnet.md - [claude-3-opus Pricing — $15.00/1M in, $75.00/1M out | Tokenomy](/models/claude-3-opus): claude-3-opus costs $15.00 per 1M input tokens and $75.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/claude-3-opus.md - [claude-3-haiku Pricing — $0.25/1M in, $1.25/1M out | Tokenomy](/models/claude-3-haiku): claude-3-haiku costs $0.25 per 1M input tokens and $1.25 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/claude-3-haiku.md - [llama-4-behemoth Pricing — $1.00/1M in, $1.00/1M out | Tokenomy](/models/llama-4-behemoth): llama-4-behemoth costs $1.00 per 1M input tokens and $1.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/llama-4-behemoth.md - [llama-4-maverick Pricing — $0.50/1M in, $0.50/1M out | Tokenomy](/models/llama-4-maverick): llama-4-maverick costs $0.50 per 1M input tokens and $0.50 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/llama-4-maverick.md - [llama-4-scout Pricing — $0.20/1M in, $0.20/1M out | Tokenomy](/models/llama-4-scout): llama-4-scout costs $0.20 per 1M input tokens and $0.20 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/llama-4-scout.md - [llama-3.3-70b Pricing — $0.60/1M in, $0.60/1M out | Tokenomy](/models/llama-3-3-70b): llama-3.3-70b costs $0.60 per 1M input tokens and $0.60 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/llama-3-3-70b.md - [gemini-3-pro Pricing — $1.00/1M in, $8.00/1M out | Tokenomy](/models/gemini-3-pro): gemini-3-pro costs $1.00 per 1M input tokens and $8.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gemini-3-pro.md - [gemini-3-flash Pricing — $0.15/1M in, $0.40/1M out | Tokenomy](/models/gemini-3-flash): gemini-3-flash costs $0.15 per 1M input tokens and $0.40 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gemini-3-flash.md - [gemini-2.5-pro Pricing — $1.25/1M in, $10.00/1M out | Tokenomy](/models/gemini-2-5-pro): gemini-2.5-pro costs $1.25 per 1M input tokens and $10.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gemini-2-5-pro.md - [gemini-2.5-flash Pricing — $0.20/1M in, $0.60/1M out | Tokenomy](/models/gemini-2-5-flash): gemini-2.5-flash costs $0.20 per 1M input tokens and $0.60 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gemini-2-5-flash.md - [gemini-2.5-flash-lite Pricing — $0.05/1M in, $0.15/1M out | Tokenomy](/models/gemini-2-5-flash-lite): gemini-2.5-flash-lite costs $0.05 per 1M input tokens and $0.15 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/gemini-2-5-flash-lite.md - [azure-gpt-5.2 Pricing — $2.00/1M in, $8.00/1M out | Tokenomy](/models/azure-gpt-5-2): azure-gpt-5.2 costs $2.00 per 1M input tokens and $8.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/azure-gpt-5-2.md - [azure-gpt-5 Pricing — $3.00/1M in, $12.00/1M out | Tokenomy](/models/azure-gpt-5): azure-gpt-5 costs $3.00 per 1M input tokens and $12.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/azure-gpt-5.md - [phi-4 Pricing — $0.20/1M in, $0.40/1M out | Tokenomy](/models/phi-4): phi-4 costs $0.20 per 1M input tokens and $0.40 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/phi-4.md - [phi-4-mini Pricing — $0.10/1M in, $0.20/1M out | Tokenomy](/models/phi-4-mini): phi-4-mini costs $0.10 per 1M input tokens and $0.20 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/phi-4-mini.md - [amazon-nova-pro Pricing — $0.80/1M in, $3.20/1M out | Tokenomy](/models/amazon-nova-pro): amazon-nova-pro costs $0.80 per 1M input tokens and $3.20 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/amazon-nova-pro.md - [amazon-nova-lite Pricing — $0.20/1M in, $0.80/1M out | Tokenomy](/models/amazon-nova-lite): amazon-nova-lite costs $0.20 per 1M input tokens and $0.80 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/amazon-nova-lite.md - [titan-text-express Pricing — $0.20/1M in, $0.60/1M out | Tokenomy](/models/titan-text-express): titan-text-express costs $0.20 per 1M input tokens and $0.60 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/titan-text-express.md - [mistral-large-4 Pricing — $2.00/1M in, $6.00/1M out | Tokenomy](/models/mistral-large-4): mistral-large-4 costs $2.00 per 1M input tokens and $6.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/mistral-large-4.md - [mistral-medium-3 Pricing — $0.40/1M in, $2.00/1M out | Tokenomy](/models/mistral-medium-3): mistral-medium-3 costs $0.40 per 1M input tokens and $2.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/mistral-medium-3.md - [mistral-small-3 Pricing — $0.20/1M in, $0.80/1M out | Tokenomy](/models/mistral-small-3): mistral-small-3 costs $0.20 per 1M input tokens and $0.80 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/mistral-small-3.md - [grok-3 Pricing — $3.00/1M in, $15.00/1M out | Tokenomy](/models/grok-3): grok-3 costs $3.00 per 1M input tokens and $15.00 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/grok-3.md - [grok-3-mini Pricing — $0.30/1M in, $0.50/1M out | Tokenomy](/models/grok-3-mini): grok-3-mini costs $0.30 per 1M input tokens and $0.50 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/grok-3-mini.md - [deepseek-v4 Pricing — $0.11/1M in, $0.22/1M out | Tokenomy](/models/deepseek-v4): deepseek-v4 costs $0.11 per 1M input tokens and $0.22 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/deepseek-v4.md - [deepseek-r2 Pricing — $0.55/1M in, $2.20/1M out | Tokenomy](/models/deepseek-r2): deepseek-r2 costs $0.55 per 1M input tokens and $2.20 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/deepseek-r2.md - [deepseek-v3 Pricing — $0.15/1M in, $0.30/1M out | Tokenomy](/models/deepseek-v3): deepseek-v3 costs $0.15 per 1M input tokens and $0.30 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/deepseek-v3.md - [qwen-3-max Pricing — $0.20/1M in, $0.60/1M out | Tokenomy](/models/qwen-3-max): qwen-3-max costs $0.20 per 1M input tokens and $0.60 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/qwen-3-max.md - [qwen-3-plus Pricing — $0.10/1M in, $0.30/1M out | Tokenomy](/models/qwen-3-plus): qwen-3-plus costs $0.10 per 1M input tokens and $0.30 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/qwen-3-plus.md - [ernie-4.5 Pricing — $0.20/1M in, $0.60/1M out | Tokenomy](/models/ernie-4-5): ernie-4.5 costs $0.20 per 1M input tokens and $0.60 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/ernie-4-5.md - [ernie-lite Pricing — $0.10/1M in, $0.30/1M out | Tokenomy](/models/ernie-lite): ernie-lite costs $0.10 per 1M input tokens and $0.30 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login. — markdown: /models/ernie-lite.md ## Citation Cite as: Tokenomy, "", https://tokenomy.ai, accessed . Pricing figures are refreshed twice daily (00:15 and 12:15 UTC) from provider catalogs aggregated via OpenRouter, Ollama and Hugging Face, and stored as an immutable daily snapshot in `price_history`.