gpt-oss-20b (high)
Description
The gpt-oss-20b model (technically 20.9B parameters) achieves near-parity with OpenAI o4-mini on core reasoning benchmarks, while running efficiently on a single 80 GB GPU. The gpt-oss-20b model delivers similar results to OpenAI o3‑mini on common benchmarks and can run on edge devices with just 16 GB of memory, making it ideal for on-device use cases, local inference, or rapid iteration without costly infrastructure. Both models also perform strongly on tool use, few-shot function calling, CoT reasoning (as seen in results on the Tau-Bench agentic evaluation suite) and HealthBench (even outperforming proprietary models like OpenAI o1 and GPT‑4o). Note: While referred to as '20b' for simplicity, it technically has 20.9B parameters.
Capability Radar
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 289 | 39.0 | AA |
| General Ranking | 214 | 53.0 | AA |
| Science | 252 | 48.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Biology
Communication
Finance
Healthcare
Math
AA Evaluation Indices
(Artificial Analysis)LLM Stats Category Scores
(LLM Stats (zeroeval))Pricing
Speed
Provider Price Ranking
Provider Price Ranking
8 providers
Compare pricing across different API providers for this model.