OpenAI API Fast mode (formerly Priority processing) SLA
OpenAI · AI models · no public SLA · checked 23h ago
No public SLA, so there is no promise to test. Observed over 14 days: 100%.
Incidents, last 365 days
No incidents on file for this service.
Commitments and credits
Docs FAQ: 'Fast mode for GPT-6 Astra does not include a latency SLA. For GPT-5.6 and earlier models, Fast mode and Scale Tier receive the same service-level agreement treatment, and eligible Enterprise agreements may provide service credits when latency targets aren't met.' No numbers are stated on the Fast mode page itself, so tiers is empty (the Scale Tier numbers may apply by reference but that is not stated explicitly).
SLA changes
- Jul 30, 2026OpenAI API Fast mode (formerly Priority processing)Priority processing renamed Fast mode ('Priority processing was renamed Fast mode on July 30, 2026'); service_tier 'priority' still accepted.Source
More in ai models
Amazon BedrockAWS · AI models · Standard99.9%First credit10%Observed 365d99.178%All incidents 99.852%
Claude API (Standard tier)Anthropic · AI modelsNoneFirst creditn/aObserved 365d98.76%All incidents 94.774%
Claude API Priority TierAnthropic · AI models · Priority Tier (existing commitments only)99.5%First creditn/aObserved 365d98.8%All incidents 94.774%
Claude Enterprise (claude.ai)Anthropic · AI modelsNoneFirst creditn/aObserved 365d98.62%All incidents 94.774%
Cohere API (SaaS: Command, Embed, Rerank)Cohere · AI modelsNoneFirst creditn/aObserved 365d100%All incidents 99.986%
Cohere private deployments / Model Vault / NorthCohere · AI modelsNoneFirst creditn/aObserved 365dn/aAll incidents 99.986%
Vertex AI (Gemini API on Vertex)Google Cloud · AI models · Vertex AI Platform: Training, Deployment, Batch Prediction99.9%First credit10%Observed 365d100%All incidents 99.488%
Hugging Face Inference EndpointsHugging Face · AI modelsNoneFirst creditn/aObserved 365dNo feed
Terms, exclusions and sources
- Measured
- Latency targets referenced but not published; SLA terms live in Enterprise agreements.
- How to claim
- Contact your account director
- Not covered
- GPT-6 Astra: no latency SLA; Fine-tuned models and embeddings not supported; SLA credits only under eligible Enterprise agreements
- Status history
- Status history since Sep 9, 2026
- Weighted uptime
- All incidents, weighted: 99.824% · Downtime
Summaries of published SLAs; the contract you sign governs. Logos via logo.dev; trademarks belong to their owners.