All modelsGoogle
gemini-2.5-flash
gemini-2.5-flashFast Gemini workhorse for multimodal apps where latency and price matter
Context 1MReasoningToolsFilesVisionAudio
Pricing
Input
S$0.389S$0.350/M
Output
S$3.24S$2.92/M
Cache read
Reusing a prompt already cached upstream
Not offered
Cache write
Storing a prompt for later reuse
Not offered
Struck-through figures are the vendor's published list price.
Performance
Loading throughput, latency, and success rate…
Quick start
Full documentationThe API is OpenAI-compatible — point the base URL here and pass the model id.
cURL
curl https://api.lulutokens.ai/v1/chat/completions \ -H "Authorization: Bearer $LULU_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "gemini-2.5-flash", "messages": [{"role": "user", "content": "Hello"}] }'Python
from openai import OpenAIclient = OpenAI(base_url="https://api.lulutokens.ai/v1", api_key="LULU_API_KEY")response = client.chat.completions.create( model="gemini-2.5-flash", messages=[{"role": "user", "content": "Hello"}],)Endpoints:openaiopenai-responseanthropic