OpenAI

GPT OSS 20B

openai
Chat

GPT OSS 20B running in a Trusted Execution Environment (TEE). OpenAI's compact open-weight 21B MoE model with 3.6B active parameters, optimized for lower-latency inference, with hardware attestation evidence available for independent verification.

Beta

Modalities

In / out price

$0.03 / $0.14

per 1M

Context

131K

Max output

16K

Released

Mar 18, 2026

Updated

Aug 7, 2026

Providers

3 routes

The same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below.

ProviderPrivacyInput / 1MOutput / 1MCache read / 1MContextMax outputUptimeLatencyThroughput
VeniceVenice
E2EE$0.05$0.19128K33K
Private$0.07$0.30128K16K
DeepInfraDeepInfra
Private$0.03$0.14131K16K

Privacy

3 / 3

End-to-end encrypted routing

E2EE tier

  1. Anonymous
  2. Private
  3. E2EE

Your client encrypts the prompt before it leaves your environment. E2EE includes TEE isolation; only the verified enclave can decrypt the request.

Features

3
Streaming
Reasoning
Web search

Uptime

GPT OSS 20B — AnonRouter Models