runaiicloud
ModelsGPUsPricingDocsConnectCompareEnterprisePlayground
Log inGet started
← Model library

OpenAI gpt-oss-120b

OpenAI's open-weights 120B MoE. Efficient reasoning per dollar.

Try in Playground

Input / 1M

$0.15

Output / 1M

$0.60

Context

131K

LLMopenaiFunction callingServerless

Benchmarks

Published evaluation scores (higher is better). Measured output speed on our fleet, Standard tier.

SWE-bench Verified

62.4%

Real GitHub issues

GPQA Diamond

71.2%

PhD-level science

MMLU Pro

80.9%

Knowledge & reasoning

AIME 2025

88.6%

Competition math

IFEval

89.8%

Instruction following

Output speed

260

tok/s on runaii

Ways to serve this model

ServerlessPer-token, zero cold starts. From $0.15/M input.Start serverless →Dedicated deploymentReserved GPUs from $40/hr. Full speed control, scale-to-zero, your region.Create deployment →LoRA fine-tuneThis model is not tunable yet.Start training →

Live model traffic — last 24h

Requests to runaii/gpt-oss-120b across the fleet, from the live meter.

Requests

…

Tokens

…

Spend

…

Quickstart

OpenAI SDK (change the base URL) or the Anthropic SDK — both speak runaii.cloud natively.

curl https://api.runaii.cloud/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $RUNAII_API_KEY" \
  -d '{
    "model": "runaii/gpt-oss-120b",
    "messages": [{"role": "user", "content": "Say hello in Spanish"}]
  }'
liveMaking requestMaking request
runaiicloud

Serverless inference, dedicated GPUs, and training for open models. OpenAI- and Anthropic-compatible APIs.

© 2026 runaii

Platform

Model libraryGPUsPricingCompare providersSavings calculatorDocsServerlessDeploymentsTrainingBatch API

Developers

PlaygroundCookbookCLIAgents / MCPResearch notesUI/UX systemUse casesTutorialsModel advisorBlogCustomersFAQ

Company

EnterpriseStartupsAboutCareersPartnersTrust centerSLAStatusChangelogrunaii chatSupportAPI keysTermsPrivacy