ewest/gemini-flashFast, low-cost model with optional reasoning.

Sign in to run this model in the playground. Usage is billed to your organisation.
Sign in to try itOpenAI-compatible chat completions; works with the official OpenAI SDKs. Every response carries an x-request-id trace ID.
curl https://api.ewest.ai/v1/chat/completions \
-H "Authorization: Bearer $EWEST_KEY" \
-H "Content-Type: application/json" \
-d '{"model": "ewest/gemini-flash", "stream": true,
"messages": [{"role": "user", "content": "Hello!"}]}'1,000 requests of 1,000 tokens in and 300 out cost about $1.05. Billed on the tokens the model actually reads and writes.