Gemini 3.7 Flash
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Specifications
Modalities
Supported Parameters
Max Output Tokens
65,536Reasoning Configuration
- Default
- Thinking on (effort: medium)
- Selectable effort levels
- highmediumlow
- Turning thinking off
- Not possible (always thinks)
Data policy
- Prompt retention
- Unknown — we could not confirm
- Training
- Unknown — not declared
"Unknown" does not mean "safe". It means we could not confirm it.
Whether a provider trains on prompts is a declared value from our terms with them. Retention is determined from the upstream listing for every host this model can reach. Neither is guessed.
Measured performance
Measured through FastMetal's own gateway over the last 30 days (651 first-response samples).
Throughput is sampled once a day: the median of 26 observations over the last 30 days.
Time to first response is measured to the first streamed chunk; throughput counts all output tokens, including reasoning tokens. These are observations, not a performance guarantee.
Code Examples
curl https://api.fastmetal.ai/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "gemini-3.7-flash",
"messages": [{"role": "user", "content": "Hello!"}]
}'How Gemini 3.7 Flash actually answers
Real responses to our standard prompts, recorded on FastMetal.
Count the number of 'r's in 'strawberry'
Count the number of 'r's in 'strawberry'. Explain your reasoning step by step.
Debug This Error
I'm getting the following error in my Node.js application: TypeError: Cannot read properties of undefined (reading 'map') at UserList (/app/components/UserList.js:12:25) at renderWithHooks (/app/node_modules/rea…
Code Review
Please review the following Python function and suggest improvements for readability, performance, and best practices: def get_data(url, retries=3): import requests import time for i in range(retries):…
Frequently asked questions
- How much does the Gemini 3.7 Flash API cost?
- On FastMetal, Gemini 3.7 Flash is billed per token: $0.7913 per 1M input tokens and $3.96 per 1M output tokens on a US-dollar account, before tax. Usage is drawn from a prepaid balance; there is no subscription or monthly fee.
- Can I call Gemini 3.7 Flash with the OpenAI SDK?
- Yes. Point base_url at https://api.fastmetal.ai/v1 and pass "gemini-3.7-flash" as the model; existing OpenAI-style code works unchanged, including streaming, tool calls and structured output.
- What is the context window of Gemini 3.7 Flash?
- 1,048,576 tokens, shared between the prompt and the response.
- What do I need to try Gemini 3.7 Flash?
- Create an account and add credit; the model is then available both in the browser chat and over the API. There is no contract or minimum spend.
Try Gemini 3.7 Flash right now
Gemini 3.7 Flash is available on FastMetal through one API key. Start in the browser, or call it from the OpenAI SDK.