On Oct 8 OpenAI Developers said Ultrafast for GPT-6.1 Sol is rolling out in the API, Codex, and ChatGPT Work. Docs confirm service_tier ultrafast at 6x Standard ($12/$60 short-context), with US/EU residency; Codex/Work access on Pro $500 and eligible Enterprise/Edu. Codex also shipped instant steering the same day.

Key Takeaways

  • ✓Surface: API + Codex + ChatGPT Work same day; OpenAIDevs post ~199K views / 2,398 likes
  • ✓Price: Ultrafast = 6x Standard → $12/M in / $60/M out short-context ($0.60 cached, $15 cache write); long-context 2x those
  • ✓Call: model gpt-6.1-sol with service_tier ultrafast; WebSockets recommended for tool-heavy agents
  • ✓Access: Codex/Work on Pro $500, eligible Enterprise, credit Edu; Enterprise off by default; US/EU residency
  • ✓Companion: Codex Day 4 instant steering for realtime course-correction alongside Ultrafast
OpenAI rolls out GPT-6.1 Sol Ultrafast across API, Codex, and ChatGPT Work: near-Astra intelligence, up to 8x Sol Standard, $12/$60 per 1M tokens
🖼️Official Media
Click to view high-res
🧭

Turn your technical choice into a development budget

Compare 40 dev plans & simulate token costs vs $20/mo subscriptions

🔬

In-Depth Technical Analysis

On Oct 8 OpenAI rolled out Ultrafast for GPT-6.1 Sol across the API, Codex, and ChatGPT Work: near-Astra intelligence at up to 8x Sol Standard speed. Ultrafast is a separate service tier (service_tier: "ultrafast") with its own rate limits. Pricing is 6x Standard: $12/M input and $60/M output short-context ($0.60 cached input, $15 cache write); long-context doubles those rates. Docs confirm US/EU data residency for Sol Ultrafast. In Codex/ChatGPT Work, access is on Pro $500, eligible usage-based Enterprise, and credit-based Edu; Enterprise admins must enable it. Codex lead Tibo also shipped instant steering on Day 4, pairing well with Ultrafast for live course-correction. OpenAI published no independent coding benchmarks beyond the speed/intelligence claims; the OpenAIDevs announcement drew ~199K views and 2,398 likes. Prefer WebSockets for agentic tool loops. Guide: https://developers.openai.com/api/docs/guides/ultrafast-mode

Action HubReady to adopt this in production?

Benchmark side-by-side against alternatives, or calculate monthly token cost vs subscription break-even.