import requests r = requests.post( "https://api.aimlapi.com/v1/chat/completions", headers={"Authorization": "Bearer " + AIMLAPI_KEY}, json={ "model": "inclusionai/ling-3.0-flash-sante", "messages": [ { "role": "user", "content": "Hello!" } ] }, ) print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "inclusionai/ling-3.0-flash-sante", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"inclusionai/ling-3.0-flash-sante","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Cached input |
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
Ling 3.0 Flash Sante This page | Reasoning with tool calling | |||
| Finance-domain reasoning over long documents | ||||
| High-volume, zero-cost chat and reasoning tasks | ||||
| Fast, low-cost reasoning and tool use at scale |
It is a health- and medicine-focused language model designed for medical knowledge reasoning, clinical safety, evidence-based retrieval, and long-horizon medical tasks.
Yes. It is used for medical question answering, including complex and multi-step clinical questions.
Yes. The model is used for evidence-based retrieval and literature summaries.
Yes. It is used for clinical documentation drafts.
Yes. Reasoning is enabled by default, and the model is designed for medical knowledge reasoning and general reasoning tasks.
Yes. It supports function calling for tool-based workflows, and tool calling works whether reasoning is enabled or disabled.
Yes. It retains general agentic capabilities in addition to its medical specialization.
Yes. The model retains general coding capabilities alongside its health and medicine capabilities.
It accepts text input only.
Yes. It uses a mixture-of-experts architecture with 124 billion total parameters and approximately 5.1 billion active parameters per token.
The provider catalogue identifies streaming as a capability for this model.