All models
gpt-5.4-nano
OpenAI
gpt-5.4-nanoCheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
Context 400KReasoningToolsFilesVision
Pricing
Input
S$0.259S$0.249/M
Output
S$1.62S$1.56/M
Cache read
Reusing a prompt already cached upstream
S$0.026S$0.025/M
Cache write
Storing a prompt for later reuse
Not offered
Struck-through figures are the vendor's published list price.
Performance
Loading throughput, latency, and success rate…
Quick start
Full documentationThe API is OpenAI-compatible — point the base URL here and pass the model id.
cURL
curl https://api.lulutokens.ai/v1/chat/completions \ -H "Authorization: Bearer $LULU_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "gpt-5.4-nano", "messages": [{"role": "user", "content": "Hello"}] }'Python
from openai import OpenAIclient = OpenAI(base_url="https://api.lulutokens.ai/v1", api_key="LULU_API_KEY")response = client.chat.completions.create( model="gpt-5.4-nano", messages=[{"role": "user", "content": "Hello"}],)Endpoints:openaiopenai-responseanthropic