Cheap prepaid AI API · checked against the public catalog

One low-cost AI API key, without a subscription.

Model.sale is a prepaid API for developers who want to try current GPT, DeepSeek and GLM models without committing to a monthly plan. Add at least $5 by card, PayPal or crypto, use an OpenAI-compatible endpoint and pay only for measured tokens.

Short answer: when is prepaid cheaper?

Prepaid is often a good fit for prototypes, coding agents and irregular workloads where a recurring seat would sit unused. The fair comparison is your own input/output mix, not a headline “discount”: Model.sale publishes a blended charge per 1M charged tokens, while many official APIs split input, cached input and output prices.

A temporary reservation protects your balance before dispatch. Once terminal usage arrives, the exact charge is settled and the unused reservation is released. The selected model is never silently replaced with another one.

Compare live model prices before you integrate.

Current public rates

Rates below are USD per one million charged tokens. Availability is refreshed automatically; a model can keep its price visible while temporarily unavailable, but only a LIVE model accepts a request.

ModelBlended / 1MAvailabilityLast check
gpt-5.5$0.45LIVE2026-10-11T04:34:59.806Z
gpt-5.6-luna$0.09LIVE2026-10-11T04:35:12.548Z
gpt-5.6-sol$0.45LIVE2026-10-11T04:35:27.990Z
gpt-5.6-terra$0.15LIVE2026-10-11T04:34:41.777Z
gpt-6-astra$1.125LIVE2026-10-11T04:34:38.731Z
gpt-6-luna$0.08LIVE2026-10-11T04:34:51.271Z
gpt-6-sol$0.38LIVE2026-10-11T04:35:31.145Z
gpt-6.1-sol$0.40LIVE2026-10-11T01:35:11.992Z
claude-fable-5-1$1.00LIVE2026-10-11T01:35:20.016Z
claude-haiku-4-5-20251001$0.20LIVE2026-10-11T04:34:27.051Z
claude-opus-4-6$0.50LIVE2026-10-11T01:35:24.030Z
claude-opus-4-7$0.50LIVE2026-10-11T01:35:22.555Z
claude-opus-4-8$0.60LIVE2026-10-11T04:35:29.755Z
claude-opus-5$0.60LIVE2026-10-11T01:35:17.705Z
claude-opus-5-5$1.00LIVE2026-10-11T04:35:34.606Z
claude-sonnet-5-5$0.55LIVE2026-10-11T01:35:25.722Z
glm-5.2$0.15LIVE2026-10-11T04:35:16.219Z
glm-5.3$0.15LIVE2026-10-11T04:35:19.417Z
glm-5.3-flash$0.10LIVE2026-10-11T04:35:21.634Z
gpt-5.4$0.10TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
gpt-5.4-20$0.10TEMPORARILY UNAVAILABLE2026-10-02T09:53:56.463Z
gpt-5.4-mini$0.09TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
claude-haiku-4-5$0.075TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
claude-opus-4-5$0.30TEMPORARILY UNAVAILABLE2026-10-02T09:53:51.813Z
claude-sonnet-4-5$0.15TEMPORARILY UNAVAILABLE2026-10-02T09:53:53.468Z
claude-sonnet-4-6$0.25TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
deepseek-flash$0.03TEMPORARILY UNAVAILABLE2026-10-11T04:35:23.574Z
deepseek-v4-flash$0.20TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
deepseek-v4-pro$0.02TEMPORARILY UNAVAILABLE2026-10-11T04:35:25.514Z
glm-5$0.12TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
glm-5.1$0.16TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
qwen-3.8$0.25TEMPORARILY UNAVAILABLE2026-10-02T09:53:58.119Z
qwen3.5-plus$0.05TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
qwen3.6-plus$0.04TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
qwen3.7-max$0.11TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
qwen3.7-plus$0.03TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
kimi-k2.5$0.07TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
kimi-k2.6$0.11TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
kimi-k3$0.24TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
mimo-v2-omni$0.05TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
mimo-v2-pro$0.05TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
mimo-v2.5$0.01TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
mimo-v2.5-pro$0.03TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
minimax-m2.5$0.02TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
minimax-m2.7$0.04TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
minimax-m3$0.04TEMPORARILY UNAVAILABLE2026-10-11T04:34:24.507Z
gpt-4o-transcribe$0.10TEMPORARILY UNAVAILABLE2026-10-10T21:42:15.944Z
gpt-image-1.5$20.00TEMPORARILY UNAVAILABLE2026-10-10T21:41:47.458Z
gpt-image-2$0.10TEMPORARILY UNAVAILABLE2026-10-10T21:41:45.762Z
gpt-image-2.5$20.00TEMPORARILY UNAVAILABLEPending
gpt-image-2.5-flare$20.00TEMPORARILY UNAVAILABLE2026-10-10T21:41:43.803Z
gpt-image-2.5-sunburst$20.00TEMPORARILY UNAVAILABLE2026-10-10T21:41:41.863Z

The full catalog, minimum request charge and health status are on Models. There are currently 19 live priced models.

01 · Start small
$5 minimum top-upUse a small balance for a real integration test. There is no annual commitment or monthly seat charge.
02 · Use one key
OpenAI-compatible base URLSet https://api.model.sale/v1, choose a published model and keep the key in an environment variable.
03 · Verify the bill
Usage per requestThe dashboard shows status, token counts, charge and the request ID for JSON and streaming calls.

Good fits for a low-cost prepaid API

PrototypesValidate an idea with a capped wallet instead of opening a larger recurring account.
Coding toolsConnect compatible clients such as Codex, OpenCode or SDKs with a separate limited key.
Variable trafficTop up only when a project needs inference. RPM, TPM, concurrency and spend caps keep experiments predictable.

Three steps to your first request

  1. Create an account and verify your email or connected identity.
  2. Create a key, set a daily or monthly limit, then add at least $5 on the billing page after sign-in.
  3. Copy the base URL and run the minimal example in the API docs. Check Usage for terminal tokens and the exact charge.

This example uses currently published model gpt-5.5 with /v1/chat/completions. Availability can change; confirm the ID in GET /v1/models.

curl https://api.model.sale/v1/chat/completions \
  -H "Authorization: Bearer $MODEL_SALE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"gpt-5.5","messages":[{"role":"user","content":"Say hello in one sentence."}]}'

How to compare with official APIs and aggregators

Official APIs may offer native features, enterprise contracts or separate input/output rates. Aggregators may offer a wider catalog and routing. Model.sale focuses on a simple prepaid wallet, a curated public catalog and request-level usage. Use the comparison page and workload calculator with the same token mix before deciding.

Do not treat the lowest per-million number as a universal saving: cache rules, output share, retries, minimum charges and model behavior change the effective cost.

Questions developers ask

Can I cancel anytime?Yes. There is no subscription to cancel; stop using the key or revoke it from the dashboard.
Are prompts stored?API request and response bodies are not stored by default. Operational metadata is retained for billing, reliability and abuse prevention.
What if a model is down?The catalog marks the exact model unavailable and returns a controlled error. The gateway does not silently reroute your request.

Read the pricing methodology, reliability snapshot and reviewed developer guides before sending production traffic.