gpt-oss-120b

OpenAI
Model overview

gpt-oss-120b

Text-only open-weight reasoning and agent model with 117B total and 5.1B active parameters, designed to run on a single 80GB GPU.

general LLMreasoningcodingagentlocal inferencecode generationsoftware engineering agent
Representative version

gpt-oss-120b

CompanyOpenAI
Release2025-08-05
Release typeModel release
Parameters117B / 5.1B active
ArchitectureMixture-of-experts transformer
Context131,072
Input modalitiestext
Output modalitiestext
Tool use / function callingConfirmed: Tool use, Function calling
Structured output / reasoningConfirmed: Structured output, Reasoning
Vision / audio / videoNot shown: Vision, Audio, Video
Open weights / LicenseWeights available / Apache-2.0 (license source pending verification)
openai-gpt-oss-120b-current · License source pending verification
Vendor APINot hosted on OpenAI API or ChatGPT
gpt-oss Model Card | OpenAIopenai-gpt-oss-120b-current · 2025-08-05
OpenRouterOpenRouter listed
PricingOpenRouter · 2026-08-11: $0.037 / $0.17
input / output, per 1M tokens
BenchmarkNo reliable public source
Public signals1 community discussion / 1 activity snapshot
Use caseslocal reasoning / coding agents / customizable self-hosted workflows / local inference
Claim evidenceRelease event / Model specifications / Weight checkpointgpt-oss Model Card | OpenAIgpt-oss-120b · 2025-08-05 · gpt-oss-120b · gpt-oss-120b · gpt-oss-120b
OpenRouter context / capabilities / pricingOpenRouter public models APIopenai/gpt-oss-120b · 2026-08-11T06:28:09.895Z
OpenRouter

Pricing / context / availability

OpenRouterOpenRouter listed
OpenRouter listingopenai/gpt-oss-120b
Mapped versiongpt-oss-120b
openai-gpt-oss-120b-current
PricingOpenRouter · 2026-08-11: $0.037 / $0.17
input / output, per 1M tokens
Context131072
Speed/LatencyNo reliable public source
Snapshot time2026-08-11T06:28:09.895Z
SourceView source
Source-bounded public ranking

OpenRouter weekly token-usage ranking

SourceModel / versionMetricRankValueCheckedLink
OpenRouterOpenAI: gpt-oss-120b
openai-gpt-oss-120b-current
Weekly token usage on OpenRouter public rankings API28413,836,501,549 tokens2026-08-11View source

OpenRouter weekly prompt-plus-completion token usage only; this is not a model-quality or whole-market ranking.

Verified limitations

Usage boundaries

  • text-only input and output
  • not hosted in the OpenAI API or ChatGPT
Public activity counters

Hugging Face model-card activity snapshot

PlatformEntryLast-month downloadsCurrent likesCurrent SpacesTimeSource
huggingface
Current HF activity snapshot
openai/gpt-oss-120b4,014,5075,095100View source

Downloads are Hugging Face's trailing-month counter; likes and Spaces are snapshots at collection time.

Missing public information

Public information not yet available

No reliable benchmark sourceNo linkable model-specific reviewSpeed/latency not disclosed