GPT-OSS 120B API: Amazon Bedrock (US)'s chat / LLM model, one key away
GPT-OSS 120B via the ApexApi gateway
OpenAI's large open-weight model, served here so you can call it without hosting it. Tool calling included; no vision. Call it through ApexApi's unified, OpenAI-compatible API with one `ak-` key, alongside every other model in the catalog. It supports a 128,000-token context window, tool / function calling, and streaming responses. Pricing is $0.18 input and $0.72 output per 1M tokens, billed pay-per-use from your credit balance with no subscription.
GPT-OSS 120B is served through ApexApi's unified, OpenAI-compatible API. One key, one format, with smart routing and automatic failover. Provider: Amazon Bedrock (US).
Why GPT-OSS 120B
128,000-token context
Reason over long documents, codebases, and multi-turn history in a single request.
Tool & function calling
Structured outputs and tool orchestration for agentic and automation workloads.
OpenAI-compatible
Drop-in with any OpenAI SDK. Change the base URL and key, keep your code.
Specifications
- Modality
- chat
- Context window
- 128K tokens
- Max output
- —
- Streaming
- Yes
- Tool / function calling
- Yes
- Vision input
- —
GPT-OSS 120B: measured reliability
From real calls ApexApi made to GPT-OSS 120B during health sweeps over the last 30 days. Our own measurements, not vendor claims.
- Success rate
- 100%
- Avg latency
- 3.5s
- Calls measured
- 1
- Last checked
- Aug 30
See how GPT-OSS 120B compares against every model we serve on the model rankings.
GPT-OSS 120B pricing
- A million tokens in and a million out costs $0.90 here, all-in.
- That makes it the 14th cheapest of the 81 chat models we serve, ranked on output price.
- It runs 100× cheaper than GPT-4, the pricier option from the same maker.
Getting started with GPT-OSS 120B
OpenAI-compatible. Point your client at ApexApi, swap your key, and call it. No new SDK to learn.
from openai import OpenAI
client = OpenAI(
base_url="https://api.apexapi.dev/v1",
api_key="<YOUR_APEXAPI_KEY>",
)
response = client.chat.completions.create(
model="openai/gpt-oss-120b",
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)GPT-OSS 120B use cases
What teams build with GPT-OSS 120B on ApexApi.
Agents & automation
Tool-using agents that orchestrate multi-step workflows and call APIs.
Coding assistants
Code generation, refactoring, reviews, and multi-file reasoning.
RAG & knowledge apps
Summarization and long-document Q&A grounded in your own data.
GPT-OSS 120B vs other models
How it compares, and on ApexApi you can switch between any of them with a single string change, same key, same endpoint.
GPT-OSS 120B vs Claude Opus 4.7
Compare GPT-OSS 120B against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.
Learn more about Claude Opus 4.7 API →GPT-OSS 120B vs GPT-5.5
Compare GPT-OSS 120B against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.
Learn more about GPT-5.5 API →GPT-OSS 120B vs Gemini 3.1 Pro Preview
Compare GPT-OSS 120B against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.
Learn more about Gemini 3.1 Pro Preview API →GPT-OSS 120B API: FAQ
How do I call GPT-OSS 120B through ApexApi?
Create an ApexApi key, point your client at https://api.apexapi.dev/v1, and POST to /v1/chat/completions with the model set to `openai/gpt-oss-120b`. The API is OpenAI-compatible, so an existing OpenAI SDK works once you change the base URL and key.
How much does GPT-OSS 120B cost on ApexApi?
$0.18 per million input tokens and $0.72 per million output tokens, all-in with no separate platform fee. You buy credits and pay per call, with no subscription or minimum.
What is the context window for GPT-OSS 120B?
128,000 tokens.
Does GPT-OSS 120B support tool calling and image input?
GPT-OSS 120B supports tool / function calling and streaming responses, and it does not support image (vision) input.
What should I use instead of GPT-OSS 120B?
From the same maker, GPT-5 Nano costs less for lighter work and GPT-5.4 Nano is the step up when this one falls short. Switching is one string: both run on the same endpoint and the same key, so there is nothing to re-integrate.
Ready to build with GPT-OSS 120B?
One API key. Every major AI model. Pay only for what you use.
Get your API key, free to start