Great for internal assistants, workflow orchestration, and knowledge services.
Context
128K tokens
Availability
2/2 available
Reference latency
2.50s
Inferred from the model family and tags; actual calls are authoritative.
Final cost is determined at settlement.
Pricing comes from the unified pricing config; balance and budget are checked before each call.
Rate-limit fields sync from the console config; no guarantees until synced.