Gemini 3.5 Flash Lite API

Gemini 3.5 Flash Lite is Google's most cost-efficient GA model, optimized for high-volume agentic tasks, translation and simple data processing, with multimodal text, image, video and audio input.
Context
1.05M tokens
Input
$0.39 / 1M
Output
$3.25 / 1M
Released
Jul 21, 2026

How to use Gemini 3.5 Flash Lite API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to google/gemini-3.5-flash-lite.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "google/gemini-3.5-flash-lite",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "google/gemini-3.5-flash-lite",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemini-3.5-flash-lite","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Gemini 3.5 Flash Lite API Pricing

TypePrice
Input
$0.39 / 1M tokens
Output
$3.25 / 1M tokens

Google Search grounding billed separately at $0.0455 per call. Cached input priced at $0.039/1M.

Gemini 3.5 Flash Lite vs other models

ModelInputOutputContextBest for
$0.39 / 1M
$3.25 / 1M
1.05M tokens
High-volume, cost-efficient agentic tasks
$0.325 / 1M
$1.95 / 1M
1M tokens
Reasoning + agents
$1.95 / 1M
$9.75 / 1M
1.05M tokens
Reasoning + agents
$1.3 / 1M
$7.8 / 1M
1M tokens
Fast, high-volume tasks
$0.975 / 1M
$4.875 / 1M
1.05M tokens
Agentic coding + high-efficiency reasoning

Start building with Gemini 3.5 Flash Lite

Get API Key
1000+ models, one API.