GoogleGoogleText / Chat

Gemini 3.5 Flash Lite API: Google Vertex (Gemini + Imagen)'s chat / LLM model, one key away

Gemini 3.5 Flash Lite via the ApexApi gateway

Gemini 3.5 Flash Lite is Flash-Lite's first step into the agentic space, prioritizing thinking and tool calling to serve as a quick, efficient, and capable subagent for larger, more complex workflows. Compared to Gemini 3.1 Flash-Lite, it delivers stronger agentic performance and tool use, more precise and complete document understanding, Call it through ApexApi's unified, OpenAI-compatible API with one `ak-` key, alongside every other model in the catalog. It supports streaming responses. Pricing is $0.34 input and $2.88 output per 1M tokens, billed pay-per-use from your credit balance with no subscription.

Gemini 3.5 Flash Lite is served through ApexApi's unified, OpenAI-compatible API. One key, one format, with smart routing and automatic failover. Provider: Google Vertex (Gemini + Imagen).

Why Gemini 3.5 Flash Lite

OpenAI-compatible

Drop-in with any OpenAI SDK. Change the base URL and key, keep your code.

Pay-per-use credits

No subscription or minimum. Buy credits and pay only for what you call.

One key, every model

Switch to any other model in the catalog by changing one string. No re-integration.

Specifications

Modality
chat
Context window
Max output
Streaming
Yes
Tool / function calling
Vision input

Gemini 3.5 Flash Lite pricing

  • A million tokens in and a million out costs $3.22 here, all-in.
  • That makes it the 37th cheapest of the 81 chat models we serve, ranked on output price.
  • It runs 5.0× cheaper than Gemini 3.1 Pro Preview, the pricier option from the same maker.

Getting started with Gemini 3.5 Flash Lite

OpenAI-compatible. Point your client at ApexApi, swap your key, and call it. No new SDK to learn.

POST /v1/chat/completions
from openai import OpenAI

client = OpenAI(
    base_url="https://api.apexapi.dev/v1",
    api_key="<YOUR_APEXAPI_KEY>",
)

response = client.chat.completions.create(
    model="google/gemini-3.5-flash-lite",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)

Gemini 3.5 Flash Lite use cases

What teams build with Gemini 3.5 Flash Lite on ApexApi.

Agents & automation

Tool-using agents that orchestrate multi-step workflows and call APIs.

Coding assistants

Code generation, refactoring, reviews, and multi-file reasoning.

RAG & knowledge apps

Summarization and long-document Q&A grounded in your own data.

Gemini 3.5 Flash Lite vs other models

How it compares, and on ApexApi you can switch between any of them with a single string change, same key, same endpoint.

Gemini 3.5 Flash Lite vs Claude Opus 4.7

Compare Gemini 3.5 Flash Lite against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.

Learn more about Claude Opus 4.7 API →

Gemini 3.5 Flash Lite vs GPT-5.5

Compare Gemini 3.5 Flash Lite against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.

Learn more about GPT-5.5 API →

Gemini 3.5 Flash Lite vs Gemini 3.1 Pro Preview

Compare Gemini 3.5 Flash Lite against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.

Learn more about Gemini 3.1 Pro Preview API →

Gemini 3.5 Flash Lite API: FAQ

How do I call Gemini 3.5 Flash Lite through ApexApi?

Create an ApexApi key, point your client at https://api.apexapi.dev/v1, and POST to /v1/chat/completions with the model set to `google/gemini-3.5-flash-lite`. The API is OpenAI-compatible, so an existing OpenAI SDK works once you change the base URL and key.

How much does Gemini 3.5 Flash Lite cost on ApexApi?

$0.34 per million input tokens and $2.88 per million output tokens, all-in with no separate platform fee. You buy credits and pay per call, with no subscription or minimum.

Does Gemini 3.5 Flash Lite support tool calling and image input?

Gemini 3.5 Flash Lite supports streaming responses, and it does not support tool / function calling and image (vision) input.

What should I use instead of Gemini 3.5 Flash Lite?

From the same maker, Gemini 3.1 Flash Lite costs less for lighter work and Gemini 3.6 Flash is the step up when this one falls short. Switching is one string: both run on the same endpoint and the same key, so there is nothing to re-integrate.

Ready to build with Gemini 3.5 Flash Lite?

One API key. Every major AI model. Pay only for what you use.

Get your API key, free to start