https://api.model.sale/v1 works with standard OpenAI-compatible clients.A simple ChatGPT-compatible API.
Use one Model.sale key with the OpenAI SDK, Codex, curl or your own backend. No monthly subscription: add balance when you need it and pay for measured usage.
GPT and Codex model prices
Rates below are the current public blended charge per one million charged tokens. Availability is checked separately: an unavailable model remains listed with its price, but cannot be called until it passes validation again. A blended rate is not the same as an input-only or output-only tariff.
| Model | Availability | Supported endpoints | Blended charge / 1M | Minimum request |
|---|---|---|---|---|
| gpt-5.5 | Live | chat/completions, responses | $0.45 / 1M | $0.01 |
| gpt-5.6-luna | Live | chat/completions, responses | $0.09 / 1M | $0.01 |
| gpt-5.6-sol | Live | chat/completions, responses | $0.45 / 1M | $0.01 |
| gpt-5.6-terra | Live | chat/completions, responses | $0.15 / 1M | $0.01 |
| gpt-6-astra | Live | chat/completions, responses | $1.125 / 1M | $0.01 |
| gpt-6-luna | Live | chat/completions, responses | $0.08 / 1M | $0.01 |
| gpt-6-sol | Live | chat/completions, responses | $0.38 / 1M | $0.01 |
| gpt-6.1-sol | Live | chat/completions, responses | $0.40 / 1M | $0.01 |
| gpt-5.4 | Temporarily unavailable | chat/completions, responses | $0.10 / 1M | $0.01 |
| gpt-5.4-20 | Temporarily unavailable | — | $0.10 / 1M | $0.01 |
| gpt-5.4-mini | Temporarily unavailable | — | $0.09 / 1M | $0.01 |
The API only accepts models marked LIVE in the live catalog. Compare all families and every current price on pricing.
How to make your first request
- Create an account and generate a Model.sale API key. The full secret is shown once.
- Add at least $5, or use the one-time Telegram test credit when eligible.
- Choose an ID marked LIVE in the catalog.
- Set the base URL and key in your server environment, then verify the request in Usage.
curl https://api.model.sale/v1/responses \
-H "Authorization: Bearer $MODEL_SALE_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"gpt-5.5","input":"Reply exactly OK"}'Set MODEL_SALE_API_KEY in your environment. Keep the key in a secret manager or local environment file. Never put it in a URL, browser bundle, shell argument, repository or analytics event.
OpenAI SDK examples
These snippets use the Responses endpoint and the currently published model gpt-5.5. The catalog is the source of truth; if this model is no longer marked LIVE, choose another live model that lists `/v1/responses`.
Python
import os
from openai import OpenAI
client = OpenAI(
base_url="https://api.model.sale/v1",
api_key=os.environ["MODEL_SALE_API_KEY"],
)
response = client.responses.create(
model="gpt-5.5",
input="Explain what an API base URL does in one sentence.",
)
print(response.output_text)TypeScript
import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://api.model.sale/v1",
apiKey: process.env.MODEL_SALE_API_KEY,
});
const response = await client.responses.create({
model: "gpt-5.5",
input: "Explain what an API base URL does in one sentence.",
});
console.log(response.output_text);Install the official SDK with `pip install openai` or `npm install openai`. Put the key in `MODEL_SALE_API_KEY` on your server; do not expose it in browser JavaScript.
/v1/responses. Streaming ends with a terminal response event./v1/chat/completions for compatible GPT, GLM and DeepSeek rows. Set stream: true for SSE.What “ChatGPT-compatible” means here
Model.sale provides a developer API, not the consumer ChatGPT product. Existing OpenAI SDKs can use standard request shapes, but each model still has its own validated capabilities, limits and endpoint list. The requested model ID is preserved; an unavailable model is never silently replaced.
| Need | Use |
|---|---|
| GPT or Codex integration | Responses where the model row lists it |
| GLM or DeepSeek | Chat Completions and SSE where listed |
| Current model IDs and prices | Full catalog and pricing |
| Client setup | Guides for Codex, SDKs and Open IDEs |
Billing and privacy
Rates are blended charges per 1M charged tokens; cached and reasoning fields remain visible when the protocol reports them. Missing terminal usage stays pending for reconciliation instead of being guessed. API prompts, source code and response bodies are not stored by default.
Read the pricing methodology, API documentation and reliability snapshot before production traffic.
FAQ
Can I use my regular OpenAI key?
No. Create a Model.sale key and set the Model.sale base URL.
Do I need a subscription?
No. Model.sale is prepaid. The minimum deposit is $5 and there is no monthly seat fee.
Where can I see the exact charge?
Every completed request appears in Usage with model, endpoint, token fields, request ID and the settled charge.