import requests r = requests.post( "https://api.aimlapi.com/v1/chat/completions", headers={"Authorization": "Bearer " + AIMLAPI_KEY}, json={ "model": "google/gemini-3.7-flash", "messages": [ { "role": "user", "content": "Hello!" } ] }, ) print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "google/gemini-3.7-flash", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"google/gemini-3.7-flash","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Cached input |
Google Search grounding billed separately at $0.0455 per call. Cached input priced at $0.0975/1M.
| Benchmark | Score | What it measures | Source | Retrieved |
|---|---|---|---|---|
| Intelligence | 39.6 | Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Coding | 71.5 | Composite score across standardised coding evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
Gemini 3.7 Flash This page | Agentic coding + high-efficiency reasoning | |||
| Reasoning + agents | ||||
| Reasoning + agents with long context | ||||
| Frontier reasoning + agents | ||||
| High-volume, cost-efficient agentic tasks |
Gemini 3.7 Flash has a 1,048,576 tokens context window and can return up to 65,536 tokens.
Gemini 3.7 Flash takes image, text as input and returns text.
Use google/gemini-3.7-flash as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.
Gemini 3.7 Flash became available on August 13, 2026.
Gemini 3.7 Flash is priced at input $0.975 / 1M tokens, output $4.875 / 1M tokens, cached input $0.0975 / 1M tokens.
Yes, Gemini 3.7 Flash can stream responses as they are generated.
Yes, Gemini 3.7 Flash accepts image input alongside text.
Gemini 3.7 Flash was built by Google.
Yes, it supports function calling and tool use, including parallel tool calls.