Leading video, image and language models behind one OpenAI-compatible endpoint, with streaming, async video jobs and a trace for every request. Built for teams that ship.
Every chat model streams. Requests enter Cloudflare's network near your users and reach our gateway in Frankfurt; generated media is served from a CDN.
Video runs as an async job with live progress and an ETA over Server-Sent Events, polling or webhooks. Your servers never hold a connection for minutes.
Each response carries an x-request-id. Look it up in the console to see the request, response, timing and cost.
Give each service its own key, with a budget and an allowed model list. Usage and cost are broken down by key and model, with CSV export.
Failed generations are not charged. Prompts and outputs are kept for 7 days for your traces, then deleted.
Use the OpenAI SDKs you already have. Switch models by changing one string, with one key and one usage report.
4 s · 720p · audioewest/veo-3.1-fast
ewest/flux-schnell
5 s · 480p · audioewest/minimax-h3-max
ewest/claude-sonnetCreate a key, point your OpenAI SDK at api.ewest.ai/v1 and send a request. Every response carries a trace ID you can look up in the console.
curl https://api.ewest.ai/v1/chat/completions \
-H "Authorization: Bearer $EWEST_KEY" \
-H "Content-Type: application/json" \
-d '{"model": "ewest/claude-haiku", "stream": true,
"messages": [{"role": "user", "content": "Hello!"}]}'Create an account, get a key and try every model in the playground.