Route requests to
GPT-OSS 120B
with Merge Gateway

Apply your own routing policies, reduce token costs automatically, and see every routing decision in real time with Merge Gateway.

How GPT-OSS 120B performs*

Intelligence - general reasoning and knowledge
24
Coding - code generation and problem-solving
30

What GPT-OSS 120B costs to run

| Vendor | Input / 1M tokens | Output / 1M tokens | Zero data retention | | --- | ---: | ---: | --- | | Amazon Bedrock | $0.1500 | $0.6000 | Yes | | Baseten | $0.1000 | $0.5000 | Yes | | BytePlus | $0.1000 | $0.5000 | Yes | | Fireworks AI | $0.1500 | $0.6000 | Yes | | Parasail | $0.1000 | $0.7500 | Yes | | Sail | $0.0600 | $0.4000 | No | | Tera | $0.0900 | $0.3600 | Yes | | Together AI | $0.1500 | $0.6000 | Yes |

Test GPT-OSS 120B
with Gateway’s Simulator

See a prompt's output, token spend, latency, and more with GPT-OSS 120B.

Route requests to GPT-OSS 120B in minutes

To get started in seconds, add our Gateway Implementation skill to your project, or pick your preferred SDK below. Check out our other quick start skills here.
Install the Merge Gateway SDK
Python
Copied!
1$ pip install merge-gateway-sdk
Send a request
Python
Copied!
1from merge_gateway import MergeGateway
2
3client = MergeGateway(api_key="YOUR_API_KEY")
4
5response = client.responses.create(
6    model="openai/gpt-5.2",
7    input=[
8        {"type": "message", "role": "system", "content": "You are a helpful programming tutor. Explain the concepts clearly with practical examples."},
9        {"type": "message", "role": "user", "content": "Explain the concept of recursion in programming with a simple set of examples."},
10    ],
11)
12
13print(response.output[0].content[0].text)
Try a diffrent model
Swap the model string to route to a different provider. No other code changes needed.
Anthropic
Copied!
1response = client.responses.create(
2    model="anthropic/claude-sonnet-4-20250514",
3    input=[
4        {"type": "message", "role": "system", "content": "You are a helpful programming tutor. Explain the concepts clearly with practical examples."},
5        {"type": "message", "role": "user", "content": "Explain the concept of recursion in programming with a simple set of examples."},
6    ],
7)
Point to Gateway
Python
Copied!
1from openai import OpenAI
2
3client = OpenAI(
4    api_key="YOUR_API_KEY",
5    base_url="https://api-gateway.merge.dev/v1/openai",
6)
Send a request
Use the standard chat.completions.create method. No provider prefix needed on the model name.
Python
Copied!
1response = client.chat.completions.create(
2    model="gpt-5.2",
3    messages=[
4        {"role": "system", "content": "You are a helpful programming tutor. Explain the concepts clearly with practical examples."},
5        {"role": "user", "content": "Explain the concept of recursion in programming with a simple set of examples."},
6    ],
7)
8
9print(response.choices[0].message.content)
Install packages
Copied!
1npm install merge-gateway-ai-sdk-provider ai
Create the provider
TypeScript
Copied!
1import { createMergeGateway } from "merge-gateway-ai-sdk-provider";
2
3const gateway = createMergeGateway({
4  apiKey: "YOUR_API_KEY",
5});
Send a request
Use generateText to send a request. Model names use the provider/model format.
TypeScript
Copied!
1import { generateText } from "ai";
2
3const { text } = await generateText({
4  model: gateway("openai/gpt-4o"),
5  prompt: "Explain the concept of recursion in programming with a simple set of examples.",
6});
7
8console.log(text);
If you already have @ai-sdk/openai installed, point it at Gateway with a base URL change:
TypeScript
Copied!
1import { createOpenAI } from "@ai-sdk/openai";
2
3const gateway = createOpenAI({
4  apiKey: "YOUR_API_KEY",
5  baseURL: "https://api-gateway.merge.dev/v1/ai-sdk",
6});
7
8// All generateText/streamText calls work unchanged
Install the Merge Gateway SDK
Anthropic SDK
Copied!
1from anthropic import Anthropic
2
3client = Anthropic(
4    api_key="YOUR_API_KEY",
5    base_url="https://api-gateway.merge.dev/v1/anthropic",
6)
7
8message = client.messages.create(
9    model="claude-sonnet-4-20250514",
10    max_tokens=1024,
11    messages=[
12        {"role": "user", "content": "Explain the concept of recursion in programming with a simple set of examples."},
13    ],
14)
15
16print(message.content[0].text)

Explore other models available in Merge Gateway

model logo
GPT-4.1 Mini
model logo
GPT-4.1 Mini
model logo
GPT-4.1 Mini (2025-04-14)
model logo
GPT-4.1 Nano
model logo
GPT-4.1 Nano
model logo
GPT-4.1 Nano (2025-04-14)
model logo
GPT-4o
model logo
GPT-4o
model logo
GPT-4o (2024-05-13)
model logo
GPT-4o (2024-08-06)
model logo
GPT-4o (2024-11-20)
model logo
GPT-4o Mini
model logo
GPT-4o Mini
model logo
GPT-4o Mini (2024-07-18)
model logo
GPT-4o Mini Search Preview
model logo
GPT-4o Mini Search Preview (2025-03-11)
model logo
GPT-4o Search Preview
model logo
GPT-4o Search Preview (2025-03-11)
model logo
GPT-4 Turbo
model logo
GPT-4 Turbo
model logo
GPT-4 Turbo (2024-04-09)
model logo
GPT-5
model logo
GPT-5
model logo
GPT-5.1

GPT-OSS 120B FAQ

Heading

What provider owns GPT-OSS 120B?

GPT-OSS 120B is a OpenAI model.

Which vendors can run GPT-OSS 120B?

Baseten is the default listed vendor, and other active vendors may also be available.

What context window does GPT-OSS 120B support?

GPT-OSS 120B supports 128,000 tokens on the primary listed vendor route.

What capabilities does GPT-OSS 120B support?

Gateway currently lists streaming, structured outputs, tool calling support for GPT-OSS 120B across its available vendor routes.

Try GPT-OSS 120B through Merge Gateway

Route, observe, and control AI requests across providers from one API.