Precios transparentes por token

Tabulador de precios AI Gateway

Documentos y preguntas a tu conocimiento: 1 crédito c/u. Chat con cualquier modelo: precio final por token — el tabulador de abajo ya es lo que pagas, todo incluido.

Pago por usoSin suscripciónTopes configurables

Recarga chica

$50

5,000 créditos

  • Todos los modelos, pago por token
  • Topes de gasto por key
  • Soporte por email
Más popular

Recarga mediana

$250

25,000 créditos

  • Todos los modelos, pago por token
  • Alertas de consumo
  • Soporte prioritario

Recarga grande

$1,000

100,000 créditos

  • Todos los modelos, pago por token
  • MCP + API + trazabilidad
  • Soporte dedicado

Planes fijos mensuales

Precio fijo en USD, créditos incluidos cada mes y todo lo de la plataforma. Si te pasas, recargas a 1¢ el crédito. Pago anual = 10 meses.

Dependencia

$199 USD/mes

Incluye 25,000 créditos/mes (valen $250)

Ahorras 20% vs. recargas sueltas

Topes forzosos · panel en vivo · soporte en español · activación en 2 días

Más popular

Municipio

$490 USD/mes

Incluye 60,000 créditos/mes (valen $600)

Ahorras 18% vs. recargas sueltas

Topes forzosos · panel en vivo · soporte en español · activación en 2 días

Ciudad

$990 USD/mes

Incluye 130,000 créditos/mes (valen $1300)

Ahorras 24% vs. recargas sueltas

Topes forzosos · panel en vivo · soporte en español · activación en 2 días

Chatbase: $32/mes por solo 10 MB · CustomGPT: desde $89/mes · Integrador Azure: ~$3,000/mes → Deployalo: desde $199/mes, todo incluido

Precio fijo en USD; CFDI en MXN al tipo de cambio del día del pago.

Precio final, sin cargos ocultos
Facturación CFDI en México
Precios por millón de tokens
Sin compromiso anual

Catálogo de modelos

400 modelos disponibles · precio final

Precios en créditos por 1M de tokens. 1 crédito = $0.01 USD. Se factura en MXN al tipo de cambio del día del pago + IVA (CFDI).

ModeloCapacidadesContexto≈ mensajes / 100 créditos
AI21: Jamba Large 1.7
ai21/jamba-large-1.7
Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 256K context...
Conocimiento hasta 2024-08-31
Max out 4k
Texto / LLM
Tools
256k tokens
220
880
≈ 150
AionLabs: Aion-2.0
aion-labs/aion-2.0
Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....
Max out 33k
Texto / LLM
Tools
Reasoning
131k tokens
88
caché read 22
176
≈ 570
AionLabs: Aion-3.0
aion-labs/aion-3.0
Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute...
Max out 33k
Texto / LLM
Tools
Reasoning
131k tokens
330
caché read 83
660
≈ 150
AionLabs: Aion-3.0-Mini
aion-labs/aion-3.0-mini
Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...
Max out 33k
Texto / LLM
Tools
Reasoning
131k tokens
77
caché read 20
154
≈ 650
AionLabs: Aion-RP 1.0 (8B)
aion-labs/aion-rp-llama-3.1-8b
Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...
Conocimiento hasta 2023-12-31
Max out 33k
Texto / LLM
33k tokens
88
176
≈ 570
AllenAI: Olmo 3 32B Think
allenai/olmo-3-32b-think
Olmo 3 32B Think is a large-scale, 32-billion-parameter model purpose-built for deep reasoning, complex logic chains and advanced instruction-following scenarios. Its capacity enables strong performance on demanding evaluation tasks and...
Max out 66k
Texto / LLM
JSON mode
Reasoning
66k tokens
17
55
≈ 2,300
Amazon: Nova 2 Lite
amazon/nova-2-lite-v1
Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing...
Max out 66k
Texto / LLM
Visión / OCR
Video
Tools
Reasoning
1000k tokens
33
275
≈ 590
Amazon: Nova Lite 1.0
amazon/nova-lite-v1
Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...
Conocimiento hasta 2024-10-31
Max out 5k
Texto / LLM
Visión / OCR
Tools
300k tokens
6.6
26
≈ 5,100
Amazon: Nova Micro 1.0
amazon/nova-micro-v1
Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...
Conocimiento hasta 2024-10-31
Max out 5k
Texto / LLM
Tools
128k tokens
3.9
15
≈ 8,700
Amazon: Nova Premier 1.0
amazon/nova-premier-v1
Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.
Max out 32k
Texto / LLM
Visión / OCR
Tools
1000k tokens
275
caché read 69
1,375
≈ 100
Amazon: Nova Pro 1.0
amazon/nova-pro-v1
Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...
Conocimiento hasta 2024-10-31
Max out 5k
Texto / LLM
Visión / OCR
Tools
300k tokens
88
352
≈ 380
Anthropic Claude Haiku Latest
~anthropic/claude-haiku-latest
This model always redirects to the latest model in the Anthropic Claude Haiku family.
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
110
caché read 11caché write 138
550
web 1,100,000
≈ 260
Anthropic Claude Sonnet Latest
~anthropic/claude-sonnet-latest
This model always redirects to the latest model in the Anthropic Claude Sonnet family.
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
220
caché read 22caché write 275
1,100
web 1,100,000
≈ 130
Anthropic: Claude 3 Haiku
anthropic/claude-3-haiku
Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal
Conocimiento hasta 2023-08-31
Max out 4k
Texto / LLM
Visión / OCR
Tools
200k tokens
28
caché read 3.3caché write 33
138
web 1,100,000
≈ 1,000
Anthropic: Claude Fable 5
anthropic/claude-fable-5
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
1,100
caché read 110caché write 1,375
5,500
web 1,100,000
≈ 26
Anthropic: Claude Fable 5 (batch)
anthropic/claude-fable-5:batch
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Fable Latest
~anthropic/claude-fable-latest
This model always redirects to the latest model in the Claude Fable family.
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
1,100
caché read 110caché write 1,375
5,500
web 1,100,000
≈ 26
Anthropic: Claude Haiku 4.5
anthropic/claude-haiku-4.5
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
110
caché read 11caché write 138
550
web 1,100,000
≈ 260
Anthropic: Claude Haiku 4.5 (batch)
anthropic/claude-haiku-4.5:batch
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
55
caché read 5.5caché write 69
275
web 1,100,000
≈ 520
Anthropic: Claude Opus 4
anthropic/claude-opus-4
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
Conocimiento hasta 2025-01-31
Max out 32k
Texto / LLM
Visión / OCR
Tools
Reasoning
200k tokens
1,650
caché read 165caché write 2,063
8,250
web 1,100,000
≈ 17
Anthropic: Claude Opus 4.1
anthropic/claude-opus-4.1
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Conocimiento hasta 2025-01-31
Max out 32k
Texto / LLM
Visión / OCR
Tools
Reasoning
200k tokens
1,650
caché read 165caché write 2,063
8,250
web 1,100,000
≈ 17
Anthropic: Claude Opus 4.1 (batch)
anthropic/claude-opus-4.1:batch
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Conocimiento hasta 2025-01-31
Max out 32k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
825
caché read 83caché write 1,031
4,125
web 1,100,000
≈ 35
Anthropic: Claude Opus 4.5
anthropic/claude-opus-4.5
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Opus 4.5 (batch)
anthropic/claude-opus-4.5:batch
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
275
caché read 28caché write 344
1,375
web 1,100,000
≈ 100
Anthropic: Claude Opus 4.6
anthropic/claude-opus-4.6
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Opus 4.6 (batch)
anthropic/claude-opus-4.6:batch
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
275
caché read 28caché write 344
1,375
web 1,100,000
≈ 100
Anthropic: Claude Opus 4.7
anthropic/claude-opus-4.7
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Opus 4.7 (batch)
anthropic/claude-opus-4.7:batch
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
275
caché read 28caché write 344
1,375
web 1,100,000
≈ 100
Anthropic: Claude Opus 4.7 (Fast)
anthropic/claude-opus-4.7-fast
Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
3,300
caché read 330caché write 4,125
16,500
web 1,100,000
≈ 9
Anthropic: Claude Opus 4.8
anthropic/claude-opus-4.8
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Opus 4.8 (batch)
anthropic/claude-opus-4.8:batch
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
275
caché read 28caché write 344
1,375
web 1,100,000
≈ 100
Anthropic: Claude Opus 4.8 (Fast)
anthropic/claude-opus-4.8-fast
Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
1,100
caché read 110caché write 1,375
5,500
web 1,100,000
≈ 26
Anthropic: Claude Opus Latest
~anthropic/claude-opus-latest
This model always redirects to the latest model in the Claude Opus family.
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Sonnet 4
anthropic/claude-sonnet-4
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
Conocimiento hasta 2025-01-31
Max out 64k
Texto / LLM
Visión / OCR
Tools
Reasoning
1000k tokens
330
caché read 33caché write 413
1,650
web 1,100,000
≈ 87
Anthropic: Claude Sonnet 4.5
anthropic/claude-sonnet-4.5
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Conocimiento hasta 2025-01-31
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
330
caché read 33caché write 413
1,650
web 1,100,000
≈ 87
Anthropic: Claude Sonnet 4.5 (batch)
anthropic/claude-sonnet-4.5:batch
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Conocimiento hasta 2025-01-31
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
165
caché read 17caché write 206
825
web 1,100,000
≈ 170
Anthropic: Claude Sonnet 4.6
anthropic/claude-sonnet-4.6
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
330
caché read 33caché write 413
1,650
web 1,100,000
≈ 87
Anthropic: Claude Sonnet 4.6 (batch)
anthropic/claude-sonnet-4.6:batch
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
165
caché read 17caché write 206
825
web 1,100,000
≈ 170
Anthropic: Claude Sonnet 5
anthropic/claude-sonnet-5
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
220
caché read 22caché write 275
1,100
web 1,100,000
≈ 130
Anthropic: Claude Sonnet 5 (batch)
anthropic/claude-sonnet-5:batch
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
110
caché read 11caché write 138
550
web 1,100,000
≈ 260
Arcee AI: Trinity Large Thinking
arcee-ai/trinity-large-thinking
Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...
Max out 262k
Texto / LLM
Tools
JSON mode
Reasoning
262k tokens
24
caché read 6.6
94
≈ 1,400
Arcee AI: Virtuoso Large
arcee-ai/virtuoso-large
Virtuoso‑Large is Arcee's top‑tier general‑purpose LLM at 72 B parameters, tuned to tackle cross‑domain reasoning, creative writing and enterprise QA. Unlike many 70 B peers, it retains the 128 k...
Conocimiento hasta 2025-03-31
Max out 64k
Texto / LLM
Tools
131k tokens
83
132
≈ 670
Auto Router
openrouter/auto
Your prompt will be processed by a meta-model and routed to one of dozens of models (see below), optimizing for the best possible output. To see which model was used,...
Texto / LLM
Imagen
Visión / OCR
Audio
Video
Tools
JSON mode
Reasoning
2000k tokens
0
0
Auto Router (Beta)
openrouter/auto-beta
Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...
Texto / LLM
Imagen
Visión / OCR
Audio
Video
Tools
JSON mode
Reasoning
2000k tokens
0
0
Baidu: ERNIE 4.5 VL 424B A47B
baidu/ernie-4.5-vl-424b-a47b
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
Conocimiento hasta 2025-03-31
Max out 16k
Texto / LLM
Visión / OCR
Reasoning
123k tokens
46
138
≈ 870
Body Builder (beta)
openrouter/bodybuilder
Transform your natural language requests into structured OpenRouter API request objects. Describe what you want to accomplish with AI models, and Body Builder will construct the appropriate API calls. Example:...
Texto / LLM
128k tokens
0
0
ByteDance Seed: Seed 1.6
bytedance-seed/seed-1.6
Seed 1.6 is a general-purpose model released by the ByteDance Seed team. It incorporates multimodal capabilities and adaptive deep thinking with a 256K context window.
Max out 33k
Texto / LLM
Visión / OCR
Video
Tools
JSON mode
Reasoning
262k tokens
28
220
≈ 730
ByteDance Seed: Seed 1.6 Flash
bytedance-seed/seed-1.6-flash
Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context window and can generate outputs of...
Max out 33k
Texto / LLM
Visión / OCR
Video
Tools
JSON mode
Reasoning
262k tokens
8.3
33
≈ 4,000
ByteDance Seed: Seed-2.0-Lite
bytedance-seed/seed-2.0-lite
Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...
Max out 131k
Texto / LLM
Visión / OCR
Video
Tools
JSON mode
Reasoning
262k tokens
28
220
≈ 730
ByteDance Seed: Seed-2.0-Mini
bytedance-seed/seed-2.0-mini
Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...
Max out 131k
Texto / LLM
Visión / OCR
Video
Tools
JSON mode
Reasoning
262k tokens
11
44
≈ 3,000
Mostrando 50 de 400 modelos (total 400). Créditos por millón de tokens, sujetos a cambio sin previo aviso.
Página 1 de 8

Ver el catálogo completo con buscador y filtros

¿Listo para empezar?

Crea tu primer API key, configura un tope de gasto y empieza a consumir modelos en minutos.