Precios transparentes por token

Tabulador de precios AI Gateway

Documentos y preguntas a tu conocimiento: 1 crédito c/u. Chat con cualquier modelo: precio final por token — el tabulador de abajo ya es lo que pagas, todo incluido.

Pago por usoSin suscripciónTopes configurables

Recarga chica

$50

5,000 créditos

  • Todos los modelos, pago por token
  • Topes de gasto por key
  • Soporte por email
Más popular

Recarga mediana

$250

25,000 créditos

  • Todos los modelos, pago por token
  • Alertas de consumo
  • Soporte prioritario

Recarga grande

$1,000

100,000 créditos

  • Todos los modelos, pago por token
  • MCP + API + trazabilidad
  • Soporte dedicado

Planes fijos mensuales

Precio fijo en USD, créditos incluidos cada mes y todo lo de la plataforma. Si te pasas, recargas a 1¢ el crédito. Pago anual = 10 meses.

Dependencia

$199 USD/mes

Incluye 25,000 créditos/mes (valen $250)

Ahorras 20% vs. recargas sueltas

Topes forzosos · panel en vivo · soporte en español · activación en 2 días

Más popular

Municipio

$490 USD/mes

Incluye 60,000 créditos/mes (valen $600)

Ahorras 18% vs. recargas sueltas

Topes forzosos · panel en vivo · soporte en español · activación en 2 días

Ciudad

$990 USD/mes

Incluye 130,000 créditos/mes (valen $1300)

Ahorras 24% vs. recargas sueltas

Topes forzosos · panel en vivo · soporte en español · activación en 2 días

Chatbase: $32/mes por solo 10 MB · CustomGPT: desde $89/mes · Integrador Azure: ~$3,000/mes → Deployalo: desde $199/mes, todo incluido

Precio fijo en USD; CFDI en MXN al tipo de cambio del día del pago.

Precio final, sin cargos ocultos
Facturación CFDI en México
Precios por millón de tokens
Sin compromiso anual

Catálogo de modelos

467 modelos disponibles · precio final

Precios en créditos por 1M de tokens. 1 crédito = $0.01 USD. Se factura en MXN al tipo de cambio del día del pago + IVA (CFDI).

ModeloCapacidadesContexto≈ mensajes / 100 créditos
AionLabs: Aion 3.5
aion-labs/aion-3.5
Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each...
Max out 33k
Texto / LLM
Tools
Reasoning
262k tokens
330
caché read 83
660
≈ 150
AionLabs: Aion 3.5 Mini
aion-labs/aion-3.5-mini
Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It is the smaller, lower-cost sibling of Aion 3.5 and uses...
Max out 33k
Texto / LLM
Tools
Reasoning
262k tokens
77
caché read 20
154
≈ 650
AionLabs: Aion-2.0
aion-labs/aion-2.0
Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....
Max out 33k
Texto / LLM
Tools
Reasoning
131k tokens
88
caché read 22
176
≈ 570
AionLabs: Aion-3.0
aion-labs/aion-3.0
Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute...
Max out 33k
Texto / LLM
Tools
Reasoning
131k tokens
330
caché read 83
660
≈ 150
AionLabs: Aion-3.0-Mini
aion-labs/aion-3.0-mini
Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...
Max out 33k
Texto / LLM
Tools
Reasoning
131k tokens
77
caché read 20
154
≈ 650
AionLabs: Aion-RP 1.0 (8B)
aion-labs/aion-rp-llama-3.1-8b
Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...
Conocimiento hasta 2023-12-31
Max out 29k
Texto / LLM
33k tokens
88
176
≈ 570
Amazon: Nova 2 Lite
amazon/nova-2-lite-v1
Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing...
Max out 66k
Texto / LLM
Visión / OCR
Video
Tools
Reasoning
1000k tokens
33
275
≈ 590
Amazon: Nova Lite 1.0
amazon/nova-lite-v1
Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...
Conocimiento hasta 2024-10-31
Max out 5k
Texto / LLM
Visión / OCR
Tools
300k tokens
6.6
26
≈ 5,100
Amazon: Nova Micro 1.0
amazon/nova-micro-v1
Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...
Conocimiento hasta 2024-10-31
Max out 5k
Texto / LLM
Tools
128k tokens
3.9
15
≈ 8,700
Amazon: Nova Premier 1.0
amazon/nova-premier-v1
Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.
Max out 32k
Texto / LLM
Visión / OCR
Tools
1000k tokens
275
caché read 69
1,375
≈ 100
Amazon: Nova Pro 1.0
amazon/nova-pro-v1
Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...
Conocimiento hasta 2024-10-31
Max out 5k
Texto / LLM
Visión / OCR
Tools
300k tokens
88
352
≈ 380
Anthropic: Claude Fable 5
anthropic/claude-fable-5
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
1,100
caché read 110caché write 1,375
5,500
web 1,100,000
≈ 26
Anthropic: Claude Fable 5 (batch)
anthropic/claude-fable-5:batch
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Fable 5.1
anthropic/claude-fable-5.1
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
1,100
caché read 28caché write 1,375
5,500
web 1,100,000
≈ 26
Anthropic: Claude Fable 5.1 (batch)
anthropic/claude-fable-5.1:batch
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 14caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Fable Latest
~anthropic/claude-fable-latest
This model always redirects to the latest model in the Claude Fable family.
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
1,100
caché read 28caché write 1,375
5,500
web 1,100,000
≈ 26
Anthropic: Claude Haiku 4.5
anthropic/claude-haiku-4.5
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
110
caché read 11caché write 138
550
web 1,100,000
≈ 260
Anthropic: Claude Haiku 4.5 (batch)
anthropic/claude-haiku-4.5:batch
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
55
caché read 5.5caché write 69
275
web 1,100,000
≈ 520
Anthropic: Claude Haiku 5.5
anthropic/claude-haiku-5.5
Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use. It succeeds Claude Haiku 4.5 with stronger coding, computer use, and...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
11
caché read 1.1caché write 14
55
web 1,100,000
≈ 2,600
Anthropic: Claude Haiku 5.5 (batch)
anthropic/claude-haiku-5.5:batch
Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use. It succeeds Claude Haiku 4.5 with stronger coding, computer use, and...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
5.5
caché read 0.6caché write 6.9
28
web 1,100,000
≈ 5,200
Anthropic: Claude Haiku Latest
~anthropic/claude-haiku-latest
This model always redirects to the latest model in the Claude Haiku family.
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
11
caché read 1.1caché write 14
55
web 1,100,000
≈ 2,600
Anthropic: Claude Opus 4.1
anthropic/claude-opus-4.1
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Conocimiento hasta 2025-01-31
Max out 32k
Texto / LLM
Visión / OCR
Tools
Reasoning
200k tokens
1,650
caché read 165caché write 2,063
8,250
web 1,100,000
≈ 17
Anthropic: Claude Opus 4.1 (batch)
anthropic/claude-opus-4.1:batch
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Conocimiento hasta 2025-01-31
Max out 32k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
825
caché read 83caché write 1,031
4,125
web 1,100,000
≈ 35
Anthropic: Claude Opus 4.5
anthropic/claude-opus-4.5
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Opus 4.5 (batch)
anthropic/claude-opus-4.5:batch
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
200k tokens
275
caché read 28caché write 344
1,375
web 1,100,000
≈ 100
Anthropic: Claude Opus 4.6
anthropic/claude-opus-4.6
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Opus 4.6 (batch)
anthropic/claude-opus-4.6:batch
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
275
caché read 28caché write 344
1,375
web 1,100,000
≈ 100
Anthropic: Claude Opus 4.7
anthropic/claude-opus-4.7
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Opus 4.7 (batch)
anthropic/claude-opus-4.7:batch
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
275
caché read 28caché write 344
1,375
web 1,100,000
≈ 100
Anthropic: Claude Opus 4.8
anthropic/claude-opus-4.8
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Opus 4.8 (batch)
anthropic/claude-opus-4.8:batch
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
275
caché read 28caché write 344
1,375
web 1,100,000
≈ 100
Anthropic: Claude Opus 5
anthropic/claude-opus-5
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
550
caché read 55caché write 688
2,750
web 1,100,000
≈ 52
Anthropic: Claude Opus 5 (batch)
anthropic/claude-opus-5:batch
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
275
caché read 28caché write 344
1,375
web 1,100,000
≈ 100
Anthropic: Claude Opus 5.5
anthropic/claude-opus-5.5
Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
440
caché read 22caché write 550
2,200
web 1,100,000
≈ 65
Anthropic: Claude Opus 5.5 (batch)
anthropic/claude-opus-5.5:batch
Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
220
caché read 11caché write 275
1,100
web 1,100,000
≈ 130
Anthropic: Claude Opus Latest
~anthropic/claude-opus-latest
This model always redirects to the latest model in the Claude Opus family.
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
440
caché read 22caché write 550
2,200
web 1,100,000
≈ 65
Anthropic: Claude Sonnet 4
anthropic/claude-sonnet-4
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
Conocimiento hasta 2025-01-31
Max out 64k
Texto / LLM
Visión / OCR
Tools
Reasoning
200k tokens
330
caché read 33caché write 413
1,650
web 1,100,000
≈ 87
Anthropic: Claude Sonnet 4.5
anthropic/claude-sonnet-4.5
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Conocimiento hasta 2025-01-31
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
330
caché read 33caché write 413
1,650
web 1,100,000
≈ 87
Anthropic: Claude Sonnet 4.5 (batch)
anthropic/claude-sonnet-4.5:batch
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Conocimiento hasta 2025-01-31
Max out 64k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
165
caché read 17caché write 206
825
web 1,100,000
≈ 170
Anthropic: Claude Sonnet 4.6
anthropic/claude-sonnet-4.6
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
330
caché read 33caché write 413
1,650
web 1,100,000
≈ 87
Anthropic: Claude Sonnet 4.6 (batch)
anthropic/claude-sonnet-4.6:batch
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
165
caché read 17caché write 206
825
web 1,100,000
≈ 170
Anthropic: Claude Sonnet 5
anthropic/claude-sonnet-5
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
220
caché read 22caché write 275
1,100
web 1,100,000
≈ 130
Anthropic: Claude Sonnet 5 (batch)
anthropic/claude-sonnet-5:batch
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
110
caché read 11caché write 138
550
web 1,100,000
≈ 260
Anthropic: Claude Sonnet 5.5
anthropic/claude-sonnet-5.5
Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade. It is especially strong at building features, fixing bugs, and producing...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
220
caché read 11caché write 275
1,100
web 1,100,000
≈ 130
Anthropic: Claude Sonnet 5.5 (batch)
anthropic/claude-sonnet-5.5:batch
Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, succeeding Claude Sonnet 5 as a direct upgrade. It is especially strong at building features, fixing bugs, and producing...
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
110
caché read 5.5caché write 138
550
web 1,100,000
≈ 260
Anthropic: Claude Sonnet Latest
~anthropic/claude-sonnet-latest
This model always redirects to the latest model in the Claude Sonnet family.
Max out 128k
Texto / LLM
Visión / OCR
Tools
JSON mode
Reasoning
1000k tokens
220
caché read 11caché write 275
1,100
web 1,100,000
≈ 130
Apodex: Apodex 1.1 Mini (free)
apodex/apodex-1.1-mini:free
Apodex 1.1 Mini is a reasoning-first model from Apodex, built for complex, long-horizon research and forecasting tasks. It works directly with files, data, code, and tools to produce verifiable results,...
Max out 236k
Texto / LLM
Tools
JSON mode
Reasoning
262k tokens
0
0
—
Arcee AI: Trinity Large Thinking
arcee-ai/trinity-large-thinking
Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...
Max out 80k
Texto / LLM
Tools
Reasoning
262k tokens
28
caché read 6.6
88
≈ 1,400
Auto Router
openrouter/auto
The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...
Texto / LLM
Imagen
Visión / OCR
Audio
Video
Tools
JSON mode
Reasoning
2000k tokens
0
0
—
Auto Router (Beta)
openrouter/auto-beta
The experimental version of our Auto Router where we test new improvements. Use it to get the latest and greatest version of our general purpose auto router, but expect beta...
Texto / LLM
Imagen
Visión / OCR
Audio
Video
Tools
JSON mode
Reasoning
2000k tokens
0
0
—
Mostrando 50 de 467 modelos (total 467). Créditos por millón de tokens, sujetos a cambio sin previo aviso.
Página 1 de 10

Ver el catálogo completo con buscador y filtros

¿Listo para empezar?

Crea tu primer API key, configura un tope de gasto y empieza a consumir modelos en minutos.