ApexApiApexApi
Sign In
GoogleGoogleText / Chat

Gemini 3.1 Flash Lite Preview API: Google AI's chat / LLM model, one key away

Gemini 3.1 Flash Lite Preview via the ApexApi gateway

The preview channel for Gemini 3.1 Flash Lite. Same shape as the stable tier; use it to test upcoming behaviour, not to serve production traffic. Call it through ApexApi's unified, OpenAI-compatible API with one `ak-` key, alongside every other model in the catalog. It supports a 1,048,576-token context window, tool / function calling, image (vision) input, and streaming responses. Pricing is $0.30 input and $1.80 output per 1M tokens, billed pay-per-use from your credit balance with no subscription.

Gemini 3.1 Flash Lite Preview is served through ApexApi's unified, OpenAI-compatible API. One key, one format, with smart routing and automatic failover. Provider: Google AI.

Why Gemini 3.1 Flash Lite Preview

1,048,576-token context

Reason over long documents, codebases, and multi-turn history in a single request.

Tool & function calling

Structured outputs and tool orchestration for agentic and automation workloads.

Vision input

Send images alongside text for document, UI, and visual-understanding tasks.

Specifications

Modality
chat
Context window
1048.6K tokens
Max output
65.5K tokens
Streaming
Yes
Tool / function calling
Yes
Vision input
Yes

Gemini 3.1 Flash Lite Preview: measured reliability

From real calls ApexApi made to Gemini 3.1 Flash Lite Preview during health sweeps over the last 30 days. Our own measurements, not vendor claims.

Success rate
100%
Avg latency
1.9s
Calls measured
1
Last checked
Aug 30

See how Gemini 3.1 Flash Lite Preview compares against every model we serve on the model rankings.

Gemini 3.1 Flash Lite Preview pricing

  • A million tokens in and a million out costs $2.10 here, all-in.
  • That makes it the 26th cheapest of the 81 chat models we serve, ranked on output price.
  • It runs 8× cheaper than Gemini 3.1 Pro Preview, the pricier option from the same maker.

Getting started with Gemini 3.1 Flash Lite Preview

OpenAI-compatible. Point your client at ApexApi, swap your key, and call it. No new SDK to learn.

from openai import OpenAI

client = OpenAI(
    base_url="https://api.apexapi.dev/v1",
    api_key="<YOUR_APEXAPI_KEY>",
)

response = client.chat.completions.create(
    model="google/gemini-3.1-flash-lite-preview",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)

Gemini 3.1 Flash Lite Preview use cases

What teams build with Gemini 3.1 Flash Lite Preview on ApexApi.

Agents & automation

Tool-using agents that orchestrate multi-step workflows and call APIs.

Coding assistants

Code generation, refactoring, reviews, and multi-file reasoning.

RAG & knowledge apps

Summarization and long-document Q&A grounded in your own data.

Gemini 3.1 Flash Lite Preview vs other models

How it compares, and on ApexApi you can switch between any of them with a single string change, same key, same endpoint.

Gemini 3.1 Flash Lite Preview vs Claude Opus 4.7

Compare Gemini 3.1 Flash Lite Preview against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.

Learn more about Claude Opus 4.7 API →

Gemini 3.1 Flash Lite Preview vs GPT-5.5

Compare Gemini 3.1 Flash Lite Preview against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.

Learn more about GPT-5.5 API →

Gemini 3.1 Flash Lite Preview vs Gemini 3.1 Pro Preview

Compare Gemini 3.1 Flash Lite Preview against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.

Learn more about Gemini 3.1 Pro Preview API →

Gemini 3.1 Flash Lite Preview API: FAQ

How do I call Gemini 3.1 Flash Lite Preview through ApexApi?

Create an ApexApi key, point your client at https://api.apexapi.dev/v1, and POST to /v1/chat/completions with the model set to `google/gemini-3.1-flash-lite-preview`. The API is OpenAI-compatible, so an existing OpenAI SDK works once you change the base URL and key.

How much does Gemini 3.1 Flash Lite Preview cost on ApexApi?

$0.30 per million input tokens and $1.80 per million output tokens, all-in with no separate platform fee. You buy credits and pay per call, with no subscription or minimum.

What is the context window for Gemini 3.1 Flash Lite Preview?

1,048,576 tokens. It can return up to 65,536 tokens in a single response.

Does Gemini 3.1 Flash Lite Preview support tool calling and image input?

Gemini 3.1 Flash Lite Preview supports tool / function calling, image (vision) input, and streaming responses.

What should I use instead of Gemini 3.1 Flash Lite Preview?

From the same maker, Gemini 2.5 Flash Lite costs less for lighter work and Gemini 3.5 Flash Lite is the step up when this one falls short. Switching is one string: both run on the same endpoint and the same key, so there is nothing to re-integrate.

Ready to build with Gemini 3.1 Flash Lite Preview?

One API key. Every major AI model. Pay only for what you use.

Get your API key, free to start
Gemini 3.1 Flash Lite Preview API on ApexApi