AI Gateway · 352

Compare models

Compare AI models, service types, context windows and availability in the ServerNet gateway. Find a model for your application and budget.

Only available models are exposed by the API. Availability requires a configured provider, valid prices and a supported execution path.
Google

gemini-2.5-flash

Available Text chat

Configured context limit: 4,096 tokens

Input84,263
Output702,188
Toman per million tokens
Google

gemini-2.5-flash-lite

Available Text chat

Configured context limit: 4,096 tokens

Input28,088
Output112,350
Toman per million tokens
google

gemini-2.5-pro

Available Text chat

Configured context limit: 1,000,000 tokens

Input351,094
Output2,808,750
Toman per million tokens
Google

gemini-2.5-pro

Available Text chat

Configured context limit: 4,096 tokens

Input702,188
Output5,617,500
Toman per million tokens
Google

gemini-3-flash-preview

Available Text chat

Configured context limit: 4,096 tokens

Input140,438
Output842,625
Toman per million tokens
Google

gemini-3-pro-preview

Available Text chat

Configured context limit: 4,096 tokens

Input561,750
Output3,370,500
Toman per million tokens
google

gemini-3.1-flash-lite

Available Text chat

Configured context limit: 1,000,000 tokens

Input70,219
Output421,313
Toman per million tokens
Google

gemini-3.1-flash-lite

Available Text chat

Configured context limit: 4,096 tokens

Input70,219
Output421,313
Toman per million tokens
Google

gemini-3.1-flash-lite-preview

Available Text chat

Configured context limit: 4,096 tokens

Input70,219
Output421,313
Toman per million tokens
google

gemini-3.1-pro

Available Text chat

Configured context limit: 1,000,000 tokens

Input561,750
Output3,370,500
Toman per million tokens
Google

gemini-3.1-pro-preview

Available Text chat

Configured context limit: 4,096 tokens

Input561,750
Output3,370,500
Toman per million tokens
google

gemini-3.5-flash

Available Text chat

Configured context limit: 1,000,000 tokens

Input421,313
Output2,527,875
Toman per million tokens
Google

gemini-3.5-flash

Available Text chat

Configured context limit: 4,096 tokens

Input421,313
Output2,527,875
Toman per million tokens
Google

gemini-3.5-flash-lite

Available Text chat

Configured context limit: 4,096 tokens

Input84,263
Output702,188
Toman per million tokens
Google

gemini-3.6-flash

Available Text chat

Configured context limit: 4,096 tokens

Input421,313
Output2,106,563
Toman per million tokens
google

gemini-3.7-flash

Available Text chat

Configured context limit: 1,000,000 tokens

Input210,657
Output1,053,282
Toman per million tokens
Google

gemini-3.7-flash

Available Text chat

Configured context limit: 4,096 tokens

Input210,657
Output1,053,282
Toman per million tokens
Google

gemini-3.8-flash

Available Text chat

Configured context limit: 4,096 tokens

Input210,657
Output1,053,282
Toman per million tokens
Google

gemini-flash-lite-latest

Available Text chat

Configured context limit: 4,096 tokens

Input70,219
Output421,313
Toman per million tokens
google

gemma-1.1-7b-it

Discontinued Text chat

Configured context limit: 8,192 tokens

Sale prices will appear after provider configuration is complete.

google

gemma-2-27b-it

Discontinued Text chat

Configured context limit: 8,192 tokens

Sale prices will appear after provider configuration is complete.

google

gemma-2-9b-it

Discontinued Text chat

Configured context limit: 8,192 tokens

Sale prices will appear after provider configuration is complete.

google

gemma-3-12b-it

Available Text chat

Configured context limit: 131,072 tokens

Input14,044
Output42,132
Toman per million tokens
google

gemma-3-27b-it

Available Text chat

Configured context limit: 131,072 tokens

Input22,470
Output44,940
Toman per million tokens