← Back to registry
OpenAI

gpt-oss:120b-cloud

Medium usage

gpt-oss-120b is OpenAI's larger open-weight model, designed for powerful reasoning, agentic tasks, and versatile developer use cases, fitting on a single 80GB GPU thanks to MXFP4 quantization.

toolsthinkingcloudCommunity r/OpenAI
Context window128K
ModalitiesText
Size120B
Pulls12.3M
Tags5
Updated10 months ago
ollama run gpt-oss:120b-cloud

Feature highlights — shared with gpt-oss-20b

Quantization

MoE weights are quantized to MXFP4 (4.25 bits per parameter), enabling the 120B model to fit on a single 80GB GPU. Ollama collaborated with OpenAI to benchmark against their reference implementations to ensure the same quality.

Best practices

Specs sourced from ollama.com/library/gpt-oss-120b · Retrieved 2026-08-29