Back to changelog

Claude Sonnet 5.5 and OpenAI's ultrafast tier

Route to Claude Sonnet 5.5, gpt-6.1-sol, and four more models

Three of the six run with zero data retention.

  • Claude Sonnet 5.5 by Anthropic
  • Claude Opus 5.5 through Amazon Bedrock, with zero data retention
  • gpt-6.1-sol and gpt-live-1 by OpenAI, both with zero data retention
  • mimo-v2.6-flash and mimo-v2.6-pro by Xiaomi Mimo

Read more in the Gateway model catalog docs.

OpenAI's ultrafast service tier is available on every Gateway API surface

Ultrafast is OpenAI's fastest processing tier and bills at six times the standard rate. Request it from any surface, alongside the flex and priority tiers. Read more in the Gateway service tiers docs.

previous_response_id chains on /v1/responses resume for 30 days

A stored turn stays available for 30 days, matching OpenAI's expiry, so a conversation picks up after a long pause. DELETE /v1/responses/{id} ends a chain early. Read more in the Gateway multi-turn conversations docs.

Gateway adds OpenAI audio translations, input token counting, and video management

Four more OpenAI endpoints pass through Gateway:

  • Audio translations
  • Responses input items
  • Responses input token counting
  • Video listing and deletion

Read more in the Gateway API reference docs.

Improvements

  • Azure OpenAI - Now serves GPT-6 Luna and GPT-6 Sol
  • Billing - Credit purchases accept SEPA Direct Debit and UPI, with a local-currency estimate and arrival time
  • Claude Sonnet 5.5 - Adds structured output, plus a clear error for unsupported forced tool choice
  • Jev 1.13 - Adds zero data retention
  • Upgraded handling for long-running provider connections.