Search by name, filter by provider or capability and compare the final per-token price. Click any model for its full spec sheet.
367 models available · final price
Prices in credits per 1M tokens. 1 credit = $0.01 USD. Billed in MXN at the exchange rate on payment day + VAT (CFDI).
| Model | Capabilities | Context | ≈ messages / 100 credits | ||
|---|---|---|---|---|---|
| AI21: Jamba Large 1.7 ai21/jamba-large-1.7 Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context... Knowledge cutoff 2024-08-31 Max out 4k | Text / LLM Tools | 256k tokens | 400 | 1,600 | ≈ 83 |
| AionLabs: Aion-2.0 aion-labs/aion-2.0 Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging.... Max out 33k | Text / LLM Tools Reasoning | 131k tokens | 160 cache read 40 | 320 | ≈ 310 |
| AionLabs: Aion-3.0 aion-labs/aion-3.0 Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute... Max out 33k | Text / LLM Tools Reasoning | 131k tokens | 600 cache read 150 | 1,200 | ≈ 83 |
| AionLabs: Aion-3.0-Mini aion-labs/aion-3.0-mini Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each... Max out 33k | Text / LLM Tools Reasoning | 131k tokens | 140 cache read 36 | 280 | ≈ 360 |
| AionLabs: Aion-RP 1.0 (8B) aion-labs/aion-rp-llama-3.1-8b Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model... Knowledge cutoff 2023-12-31 Max out 33k | Text / LLM | 33k tokens | 160 | 320 | ≈ 310 |
| AllenAI: Olmo 3 32B Think allenai/olmo-3-32b-think Olmo 3 32B Think is a large-scale, 32-billion-parameter model purpose-built for deep reasoning, complex logic chains and advanced instruction-following scenarios. Its capacity enables strong performance on demanding evaluation tasks and... Max out 66k | Text / LLM JSON mode Reasoning | 66k tokens | 30 | 100 | ≈ 1,300 |
| Amazon: Nova 2 Lite amazon/nova-2-lite-v1 Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing... Max out 66k | Text / LLM Vision / OCR Video Tools Reasoning | 1000k tokens | 60 | 500 | ≈ 320 |
| Amazon: Nova Lite 1.0 amazon/nova-lite-v1 Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite... Knowledge cutoff 2024-10-31 Max out 5k | Text / LLM Vision / OCR Tools | 300k tokens | 12 | 48 | ≈ 2,800 |
| Amazon: Nova Micro 1.0 amazon/nova-micro-v1 Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length... Knowledge cutoff 2024-10-31 Max out 5k | Text / LLM Tools | 128k tokens | 7.0 | 28 | ≈ 4,800 |
| Amazon: Nova Premier 1.0 amazon/nova-premier-v1 Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models. Max out 32k | Text / LLM Vision / OCR Tools | 1000k tokens | 500 cache read 125 | 2,500 | ≈ 57 |
| Amazon: Nova Pro 1.0 amazon/nova-pro-v1 Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December... Knowledge cutoff 2024-10-31 Max out 5k | Text / LLM Vision / OCR Tools | 300k tokens | 160 | 640 | ≈ 210 |
| Anthropic Claude Haiku Latest ~anthropic/claude-haiku-latest This model always redirects to the latest model in the Anthropic Claude Haiku family. Max out 64k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 200k tokens | 200 cache read 20cache write 250 | 1,000 web 2,000,000 | ≈ 140 |
| Anthropic Claude Sonnet Latest ~anthropic/claude-sonnet-latest This model always redirects to the latest model in the Anthropic Claude Sonnet family. Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 400 cache read 40cache write 500 | 2,000 web 2,000,000 | ≈ 71 |
| Anthropic: Claude 3 Haiku anthropic/claude-3-haiku Claude 3 Haiku is Anthropic's fastest and most compact model for
near-instant responsiveness. Quick and accurate targeted performance.
See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku)
#multimodal Knowledge cutoff 2023-08-31 Max out 4k | Text / LLM Vision / OCR Tools | 200k tokens | 50 cache read 6.0cache write 60 | 250 web 2,000,000 | ≈ 570 |
| Anthropic: Claude Fable 5 anthropic/claude-fable-5 Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 2,000 cache read 200cache write 2,500 | 10,000 web 2,000,000 | ≈ 14 |
| Anthropic: Claude Fable 5 (batch) anthropic/claude-fable-5:batch Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 1,000 cache read 100cache write 1,250 | 5,000 web 2,000,000 | ≈ 29 |
| Anthropic: Claude Fable Latest ~anthropic/claude-fable-latest This model always redirects to the latest model in the Claude Fable family. Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 2,000 cache read 200cache write 2,500 | 10,000 web 2,000,000 | ≈ 14 |
| Anthropic: Claude Haiku 4.5 anthropic/claude-haiku-4.5 Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance... Max out 64k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 200k tokens | 200 cache read 20cache write 250 | 1,000 web 2,000,000 | ≈ 140 |
| Anthropic: Claude Haiku 4.5 (batch) anthropic/claude-haiku-4.5:batch Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance... Max out 64k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 200k tokens | 100 cache read 10cache write 125 | 500 web 2,000,000 | ≈ 290 |
| Anthropic: Claude Opus 4 anthropic/claude-opus-4 Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in... Knowledge cutoff 2025-01-31 Max out 32k | Text / LLM Vision / OCR Tools Reasoning | 200k tokens | 3,000 cache read 300cache write 3,750 | 15,000 web 2,000,000 | ≈ 10 |
| Anthropic: Claude Opus 4.1 anthropic/claude-opus-4.1 Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains... Knowledge cutoff 2025-01-31 Max out 32k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 200k tokens | 3,000 cache read 300cache write 3,750 | 15,000 web 2,000,000 | ≈ 10 |
| Anthropic: Claude Opus 4.1 (batch) anthropic/claude-opus-4.1:batch Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains... Knowledge cutoff 2025-01-31 Max out 32k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 200k tokens | 1,500 cache read 150cache write 1,875 | 7,500 web 2,000,000 | ≈ 19 |
| Anthropic: Claude Opus 4.5 anthropic/claude-opus-4.5 Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and... Max out 64k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 200k tokens | 1,000 cache read 100cache write 1,250 | 5,000 web 2,000,000 | ≈ 29 |
| Anthropic: Claude Opus 4.5 (batch) anthropic/claude-opus-4.5:batch Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and... Max out 64k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 200k tokens | 500 cache read 50cache write 625 | 2,500 web 2,000,000 | ≈ 57 |
| Anthropic: Claude Opus 4.6 anthropic/claude-opus-4.6 Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 1,000 cache read 100cache write 1,250 | 5,000 web 2,000,000 | ≈ 29 |
| Anthropic: Claude Opus 4.6 (batch) anthropic/claude-opus-4.6:batch Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 500 cache read 50cache write 625 | 2,500 web 2,000,000 | ≈ 57 |
| Anthropic: Claude Opus 4.7 anthropic/claude-opus-4.7 Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 1,000 cache read 100cache write 1,250 | 5,000 web 2,000,000 | ≈ 29 |
| Anthropic: Claude Opus 4.7 (batch) anthropic/claude-opus-4.7:batch Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 500 cache read 50cache write 625 | 2,500 web 2,000,000 | ≈ 57 |
| Anthropic: Claude Opus 4.7 (Fast) anthropic/claude-opus-4.7-fast Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing.
Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 6,000 cache read 600cache write 7,500 | 30,000 web 2,000,000 | ≈ 5 |
| Anthropic: Claude Opus 4.8 anthropic/claude-opus-4.8 Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 1,000 cache read 100cache write 1,250 | 5,000 web 2,000,000 | ≈ 29 |
| Anthropic: Claude Opus 4.8 (batch) anthropic/claude-opus-4.8:batch Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 500 cache read 50cache write 625 | 2,500 web 2,000,000 | ≈ 57 |
| Anthropic: Claude Opus 4.8 (Fast) anthropic/claude-opus-4.8-fast Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8.
Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 2,000 cache read 200cache write 2,500 | 10,000 web 2,000,000 | ≈ 14 |
| Anthropic: Claude Opus Latest ~anthropic/claude-opus-latest This model always redirects to the latest model in the Claude Opus family. Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 1,000 cache read 100cache write 1,250 | 5,000 web 2,000,000 | ≈ 29 |
| Anthropic: Claude Sonnet 4 anthropic/claude-sonnet-4 Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),... Knowledge cutoff 2025-01-31 Max out 64k | Text / LLM Vision / OCR Tools Reasoning | 1000k tokens | 600 cache read 60cache write 750 | 3,000 web 2,000,000 | ≈ 48 |
| Anthropic: Claude Sonnet 4.5 anthropic/claude-sonnet-4.5 Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with... Knowledge cutoff 2025-01-31 Max out 64k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 600 cache read 60cache write 750 | 3,000 web 2,000,000 | ≈ 48 |
| Anthropic: Claude Sonnet 4.5 (batch) anthropic/claude-sonnet-4.5:batch Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with... Knowledge cutoff 2025-01-31 Max out 64k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 300 cache read 30cache write 375 | 1,500 web 2,000,000 | ≈ 95 |
| Anthropic: Claude Sonnet 4.6 anthropic/claude-sonnet-4.6 Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 600 cache read 60cache write 750 | 3,000 web 2,000,000 | ≈ 48 |
| Anthropic: Claude Sonnet 5 anthropic/claude-sonnet-5 Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 400 cache read 40cache write 500 | 2,000 web 2,000,000 | ≈ 71 |
| Anthropic: Claude Sonnet 5 (batch) anthropic/claude-sonnet-5:batch Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,... Max out 128k | Text / LLM Vision / OCR Tools JSON mode Reasoning | 1000k tokens | 200 cache read 20cache write 250 | 1,000 web 2,000,000 | ≈ 140 |
| Arcee AI: Trinity Large Thinking arcee-ai/trinity-large-thinking Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7... Max out 262k | Text / LLM Tools JSON mode Reasoning | 262k tokens | 44 cache read 12 | 170 | ≈ 780 |
| Arcee AI: Virtuoso Large arcee-ai/virtuoso-large Virtuoso‑Large is Arcee's top‑tier general‑purpose LLM at 72 B parameters, tuned to tackle cross‑domain reasoning, creative writing and enterprise QA. Unlike many 70 B peers, it retains the 128 k... Knowledge cutoff 2025-03-31 Max out 64k | Text / LLM Tools | 131k tokens | 150 | 240 | ≈ 370 |
| Auto Router openrouter/auto Your prompt will be processed by a meta-model and routed to one of dozens of models (see below), optimizing for the best possible output. To see which model was used,... | Text / LLM Image Vision / OCR Audio Video Tools JSON mode Reasoning | 2000k tokens | 0 | 0 | — |
| Auto Router (Beta) openrouter/auto-beta Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your... | Text / LLM Image Vision / OCR Audio Video Tools JSON mode Reasoning | 2000k tokens | 0 | 0 | — |
| Baidu: ERNIE 4.5 VL 424B A47B baidu/ernie-4.5-vl-424b-a47b ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data... Knowledge cutoff 2025-03-31 Max out 16k | Text / LLM Vision / OCR Reasoning | 123k tokens | 84 | 250 | ≈ 480 |
| Body Builder (beta) openrouter/bodybuilder Transform your natural language requests into structured OpenRouter API request objects. Describe what you want to accomplish with AI models, and Body Builder will construct the appropriate API calls. Example:... | Text / LLM | 128k tokens | 0 | 0 | — |
| ByteDance Seed: Seed 1.6 bytedance-seed/seed-1.6 Seed 1.6 is a general-purpose model released by the ByteDance Seed team. It incorporates multimodal capabilities and adaptive deep thinking with a 256K context window. Max out 33k | Text / LLM Vision / OCR Video Tools JSON mode Reasoning | 262k tokens | 50 | 400 | ≈ 400 |
| ByteDance Seed: Seed 1.6 Flash bytedance-seed/seed-1.6-flash Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context window and can generate outputs of... Max out 33k | Text / LLM Vision / OCR Video Tools JSON mode Reasoning | 262k tokens | 15 | 60 | ≈ 2,200 |
| ByteDance Seed: Seed-2.0-Lite bytedance-seed/seed-2.0-lite Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across... Max out 131k | Text / LLM Vision / OCR Video Tools JSON mode Reasoning | 262k tokens | 50 | 400 | ≈ 400 |
| ByteDance Seed: Seed-2.0-Mini bytedance-seed/seed-2.0-mini Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,... Max out 131k | Text / LLM Vision / OCR Video Tools JSON mode Reasoning | 262k tokens | 20 | 80 | ≈ 1,700 |
| ByteDance: UI-TARS 7B bytedance/ui-tars-1.5-7b UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games. Built by ByteDance, it builds upon the UI-TARS framework with reinforcement... Knowledge cutoff 2025-01-31 Max out 2k | Text / LLM Vision / OCR JSON mode | 128k tokens | 20 cache read 20 | 40 | ≈ 2,500 |
Top up your balance, create an API key with a spend cap and use any model in the catalog.