LuluTokens
All models
Gemini

gemini-3-flash-preview

Google

gemini-3-flash-preview

New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs

Context 1MReasoningToolsFilesVisionAudio

Pricing

Input

S$0.648S$0.583/M

Output

S$3.89S$3.50/M

Cache read

Reusing a prompt already cached upstream

S$0.065S$0.058/M

Cache write

Storing a prompt for later reuse

Not offered

Struck-through figures are the vendor's published list price.

Performance

Loading throughput, latency, and success rate…

The API is OpenAI-compatible — point the base URL here and pass the model id.

cURL
curl https://api.lulutokens.ai/v1/chat/completions \  -H "Authorization: Bearer $LULU_API_KEY" \  -H "Content-Type: application/json" \  -d '{    "model": "gemini-3-flash-preview",    "messages": [{"role": "user", "content": "Hello"}]  }'
Python
from openai import OpenAIclient = OpenAI(base_url="https://api.lulutokens.ai/v1", api_key="LULU_API_KEY")response = client.chat.completions.create(    model="gemini-3-flash-preview",    messages=[{"role": "user", "content": "Hello"}],)
Endpoints:openaiopenai-responseanthropic