Every reAPI figure below is the live rate your key is billed at — the same number the model's own page shows. Failed generations are refunded in full.
up to 88%
against official API list prices
$0
failed generations are refunded automatically
None
pay per call from one prepaid balance
| Model | reAPI | You save | Official API |
|---|---|---|---|
| GPT Image 2 · 1KImage | $0.030per image | −86% | $0.219per image |
| GPT Image 2 · 4KImage | $0.080per image | −81% | $0.413per image |
| Nano Banana 2 · 1KImage | $0.028per image | −65% | $0.080per image |
| Nano Banana 2 · 4KImage | $0.063per image | −61% | $0.160per image |
| Nano Banana Pro · 1K/2KImage | $0.032per image | −79% | $0.150per image |
| Nano Banana Pro · 4KImage | $0.035per image | −88% | $0.300per image |
| Seedream 5.0 LiteImage | $0.031per image | −12% | $0.035per image |
| Seedream 5.0 Pro · 2KImage | $0.063per image | −30% | $0.090per image |
| GPT-5.6 Sol · outputLLM | $24.00per 1M tokens | −20% | $30.00per 1M tokens |
| GPT-5.6 Terra · outputLLM | $12.00per 1M tokens | −20% | $15.00per 1M tokens |
| GPT-5.6 Luna · outputLLM | $4.80per 1M tokens | −20% | $6.00per 1M tokens |
| Claude Fable 5 · outputLLM | $40.00per 1M tokens | −20% | $50.00per 1M tokens |
| Claude Opus 4.8 · outputLLM | $20.00per 1M tokens | −20% | $25.00per 1M tokens |
| Kimi K3 · outputLLM | $12.00per 1M tokens | −20% | $15.00per 1M tokens |
| GLM-5.2 · outputLLM | $3.00per 1M tokens | −32% | $4.40per 1M tokens |
Official API figures are the vendors' published list prices, checked 26 July 2026. reAPI figures track your account's live pricing. Models whose upstream bills in a different unit are not listed.
See all model pricingWhat teams ask before adopting an AI API aggregator
reAPI reads the model name in your request, looks at vendor health, current latency, and current load, and hands the call to whichever candidate is most likely to return cleanly fastest. Health checks run continuously, so the route table is always live, not refreshed on a schedule. When the chosen vendor regresses mid-flight, the next call moves automatically and your application never sees the failure.
No. Request bodies and model outputs travel through the AI API aggregator but are never persisted. We retain only billing and audit metadata: model name, token counts, latency, status code, and key identifier. That is what compliance reviews want and what we are willing to be on the hook for. The same posture is the industry baseline among serious gateways.
Set the base URL to https://reapi.ai/api/v1 and supply a key minted in the dashboard. The OpenAI, Anthropic, and Google SDKs continue to work unmodified — reAPI preserves their request and response shapes. Provider-specific paths like Veo or Suno mirror the original vendor request shape, so an existing integration moves over without a rewrite.
Every popular model is fronted by multiple vendors. When one degrades, reAPI moves traffic to the next within a sub-second window and the call still returns. Single-source models surface a structured error rather than guess. The status page publishes a rolling history.
One reAPI key covers every supported model. We handle the vendor relationships behind the scenes, so there are fewer keys to rotate, fewer dashboards to monitor, and a single audit trail across image, video, chat, music, and code workloads. Adding a new provider to the catalog reaches your account without any procurement work on your side.
Pay-as-you-go credits, priced per completed generation — per token, image, or video second depending on the model. There is no subscription and no monthly minimum, credits never expire, and failed generations are refunded automatically. Every model page carries a live pricing table, so you can see the exact rate before you call.
One key reaches 100+ models across image, video, chat, music, and code — Seedance, Seedream, GPT Image, Nano Banana, Veo, Kling, Midjourney, Claude, GPT-5.6, Kimi, Suno, and more. New models reach the catalog within days of vendor release, land behind the same endpoint shape, and show up in your dashboard without any extra setup.
The dispatch decision is made against an always-live route table — vendor health, latency, and load are tracked continuously, so picking a route costs a negligible slice next to model inference time. For generation workloads the model run dominates end to end, and because routing steers around degraded vendors, slow-provider tail latency often drops rather than grows.
Yes. Every new account starts with free credits — no credit card required — enough to make real calls against the cheaper image and video models and to see the dashboard, logs, and per-call costs with your own requests. When you top up later, the pay-as-you-go rate is the same one shown on the model pages.
Direct integrations multiply everything: keys to rotate, dashboards to watch, billing to reconcile, and outages to handle yourself. reAPI collapses that into one key, one bill, and one audit trail, adds automatic failover across vendors, and lets you switch models by changing the model name in the request. The zero-logging posture stays identical across every modality.