import requests r = requests.post( "https://api.aimlapi.com/v1/chat/completions", headers={"Authorization": "Bearer " + AIMLAPI_KEY}, json={ "model": "stepfun/step-3.7-flash", "messages": [ { "role": "user", "content": "Hello!" } ] }, ) print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "stepfun/step-3.7-flash", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"stepfun/step-3.7-flash","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Cached input |
| Benchmark | Score | What it measures | Source | Retrieved |
|---|---|---|---|---|
| Intelligence | 19.5 | Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Coding | 39.6 | Composite score across standardised coding evaluations, measured independently by Artificial Analysis | Source | September 12, 2026 |
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
Step 3.7 Flash This page | Coding + agents | |||
| Reasoning + agents | ||||
| Balanced coding + agents | ||||
| Long-context, multimodal & agentic workflows | ||||
| Reasoning + agents |
Step 3.7 Flash has a 256,000 tokens context window and can return up to 256,000 tokens.
Step 3.7 Flash takes image, text, video as input and returns text.
Use stepfun/step-3.7-flash as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.
Step 3.7 Flash became available on June 1, 2026.
Step 3.7 Flash is priced at input $0.27508 / 1M tokens, output $1.58171 / 1M tokens, cached input $0.055016 / 1M tokens.
Yes, Step 3.7 Flash can stream responses as they are generated.
Yes, Step 3.7 Flash accepts image input alongside text.
Step 3.7 Flash was built by StepFun.
Yes, its capabilities include reasoning alongside function calling and structured outputs.
Yes, it supports function calling, tool use, and parallel tool calls.