One API endpoint
Configure one base URL and choose an exact model per request. Keep model access and client configuration easy to inspect.
API compatibilityOpenAI-compatible AI API gateway
One endpoint.
More ways to build.
GPT-5.5, GPT-5.6 and GPT-6 through one familiar API.
Build with Python or Node.js. Stream responses. Stay in control of usage.
https://api.useinfergate.com/v1The current lineup
Use the exact model you select. No silent substitutions.
gpt-5.5gpt-5.6-solgpt-5.6-terragpt-6-astraYour authenticated model list determines account access. Availability and account limits apply.
From familiar SDK to first response
Your InferGate key, a model ID, and the SDK you already know. Start with a small request.
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["INFERGATE_API_KEY"],
base_url="https://api.useinfergate.com/v1",
)
response = client.responses.create(
model="gpt-5.5",
input="Reply with one sentence about API design.",
max_output_tokens=256,
)
print(response.output_text)import OpenAI from "openai";
const client = new OpenAI({
apiKey: process.env.INFERGATE_API_KEY,
baseURL: "https://api.useinfergate.com/v1",
});
const response = await client.responses.create({
model: "gpt-5.5",
input: "Reply with one sentence about API design.",
max_output_tokens: 256,
});
console.log(response.output_text);Built around your application
Configure one base URL and choose an exact model per request. Keep model access and client configuration easy to inspect.
API compatibilityReceive incremental text through streaming responses. Handle completion and cancellation explicitly in your application.
Streaming guideManage API keys, inspect usage records and read your prepaid balance in the InferGate application.
Explore the gatewayPrepaid. Transparent.
Review current model rates before you add credit.
Start small, then compare actual usage with your workload.
Choose a model. Follow a guide. Make your first request.