Choose flagship
Use the top model when correctness, reasoning depth, and coding reliability matter more than cost.
Gemini collection
The Gemini API Family collection is a comparison of related Text & Multimodal models on APIAny. Use it to pick a version by quality, cost, and workflow before you integrate.
Use Gemini models for chat, low-latency multimodal tasks, and Google visual generation workflows.
| Model | Best For | Input | Output | Context | Cached |
|---|---|---|---|---|---|
| Gemini 3.5 Flash Fast Chat | Responsive Gemini chat model with an OpenAI-compatible API and token billing. | 1 credit / 1K | 4 credits / 1K | 1M | 0.1 credit / 1K |
| Gemini 2.5 Flash Lite Value Chat | Lower-cost Gemini model for high-throughput chat and automation tasks. | 1 credit / 1K | 2 credits / 1K | 1M | 0.1 credit / 1K |
| Gemini Omni Video | Gemini Omni text-to-video and image-to-video models under one family. | Text / image | Video | 6s / 8s / 10s | Async |
| Nano Banana 2 Visual | Image generation and editing for production creative workflows. | Text / image | Image | 0.5K / 1K / 2K / 4K | Async |
Use the top model when correctness, reasoning depth, and coding reliability matter more than cost.
Use value models for assistants, automation, support, and workflows that still need strong quality.
Use lightweight models for routing, classification, extraction, and high-volume background tasks.
Prototype and ship features on Text & Multimodal models without managing multiple provider SDKs.
Route across the family by changing the model id — pricing, channels, and fallbacks stay on the platform.
Evaluate versions side by side, then move traffic between models without rewriting integration code.
Unified access
Keep product code stable while changing models, providers, pricing, and channel routes from the platform.
curl https://apiany.ai/v1/chat/completions \
-H "Authorization: Bearer $APIANY_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3.5-flash",
"messages": [{ "role": "user", "content": "Plan an agent workflow" }]
}'Send requests to one OpenAI-compatible APIAny endpoint with a single API key, and set the public model id in the request body — no provider-specific SDKs required.
Usage is paid from prepaid APIAny credits at the rates shown in the comparison table. Credits never expire and there is no subscription.
Yes. Change only the model id in your request — endpoint, auth, and response format stay the same across the family.
Use the comparison table and the How to choose guidance above: pick flagship for quality, balanced for everyday production, and budget for high-volume tasks.
Sources
The quotations below are from official API documentation that APIAny implements against.
The Chat Completions API endpoint will generate a model response from a list of messages comprising a conversation.
The Gemini API provides access to Google's most capable generative AI models.
APIAny© 2026 APIAny. All rights reserved.