Clear request costs
Input, output and tax are calculated separately. The final charge and request ID appear in your usage report.
Choose a model, create a project and connect using the OpenAI request format. Toman credit and a record of every request help you control spending.
from openai import OpenAI import os client = OpenAI( api_key=os.environ["SERVERNET_API_KEY"], base_url="https://servernet.cloud/v1", max_retries=0, ) models = client.models.list() if not models.data: raise RuntimeError("No models available") available_model_id = models.data[0].id response = client.chat.completions.create( model=available_model_id, messages=[{"role":"user","content":"Hello"}], max_tokens=128, )
Input, output and tax are calculated separately. The final charge and request ID appear in your usage report.
Set monthly project budgets and daily key limits. Issue and revoke keys independently.
Change the base URL and API key. The compatible endpoint supports text chat, tool calls and streaming.
Context window: 1,048,576 tokens
Context window: 1,048,576 tokens
Context window: 1,048,576 tokens
Context window: 1,048,576 tokens
Context window: 1,048,576 tokens
Context window: 1,048,576 tokens
Billing uses reported model consumption and the prices frozen for your request. Usage and tax are recorded separately. When a response becomes uncertain after sending, consumption is reconciled according to the request state.
Each model shows its availability. Only available models can be purchased through the API. Preparing and discontinued models remain visible for catalogue reference.
Set a monthly project budget and a daily API key limit. Holds for concurrent requests count towards these limits.
Online — here to help