128K
text, image
text
Supported
About this model
GPT-4o (Omnimodal) is OpenAI's 1.2 trillion parameter multimodal foundation model featuring unified input processing across text, vision, and audio. Optimized for real-time interaction with enhanced reasoning and cross-modal understanding capabilities.
Input modalities
text
Accepted as model input
image
Accepted as model input
Providers
Available routing options for this model through OneInfer.
Input tokens
$2.500 / 1M tokens
Output tokens
$10.000 / 1M tokens
Routing
OneInfer optimized
Pricing
Current OneInfer pricing for this model.
| Usage | Price |
|---|---|
| Input tokens | $2.500 / 1M tokens |
| Output tokens | $10.000 / 1M tokens |
| Cached input tokens | $1.250 / 1M tokens |
Performance
Published evaluation results associated with this model.
Multimodal Understanding
Reasoning Performance
Efficiency & Latency
Cross-Modal Alignment
API Guide
- 1
Generate your access token
Use your API key to generate the JWT access token required by the Models API.
- 2
Add the JWT token
Copy the generated JWT token and replace
JWT_TOKENin the example below. - 3
Check your credits
If your balance is too low, before calling the model.
- 4
Run the API example
Choose your preferred language, copy the example, and send your first model request.
curl https://api.oneinfer.ai/v1/ula/chat/completions \ -H "Authorization: Bearer $JWT_TOKEN" \ -H "Content-Type: application/json" \ -d '{ "model": "openai/gpt-4o", "messages": [ { "role": "user", "content": "Hello!" } ] }'