//
Google's most cost-effective multimodal model, which can provide the fastest performance for high-frequency lightweight tasks. Gemini 3.1 Flash-Lite is most suitable for handling massive agent tasks, simple data extraction tasks, and extremely low-latency applications where budget and speed are the main constraints.
This model is available from multiple providers. Select one to see its pricing and sample code.
Official Channels of Google AI Studio for Gemini
Google Gemini Studio + Vertex Hybrid Channel
Transparent pay-as-you-go pricing by token usage — no subscriptions or hidden fees.
Copy an example below to start calling this model in minutes.
from openai import OpenAI
client = OpenAI(
base_url="https://api.apipod.ai/v1",
api_key="<YOUR_API_KEY>",
)
# Chat Completions
completion = client.chat.completions.create(
model="gemini-3.1-flash-lite-preview",
messages=[
{
"role": "user",
"content": "What is the meaning of life?",
}
],
)
print(completion.choices[0].message.content)