Claude Opus 4.7 is Anthropic's most capable generally available model for complex reasoning and agentic coding. It features a 1M token context window, 128K max output tokens, adaptive thinking, and strong multimodal capabilities.
AWS Bedrock
inference provider · 47 models
Access 47 models served through AWS Bedrock on AnonRouter's privacy-first gateway, including Claude Opus 4.7, Claude Opus 4.8, and Claude Opus 5. Private Bedrock routes require zero-data-retention mode; anonymous Bedrock routes hide the end user's identity behind AnonRouter's AWS account but may retain payload content.
Models
47
Modalities
1
Text
From (input)
$0.04
per 1M tokens
Max context
1M
Private routes
47 / 47
not anonymous-only
Catalog by modality
47 routesAWS Bedrock models47
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports long-horizon agentic work, complex multi-step coding, and memory-driven tasks where coherence over extended sessions matters. It features a 1M token context window, 128K max output tokens, adaptive thinking, and strong multimodal capabilities.
Claude Opus 5 is Anthropic's most capable model in the Opus family. It delivers major gains over Opus 4.8 in agentic coding, professional knowledge work, and long-horizon reasoning, with a 1M token context window, 128K max output tokens, adaptive thinking, and strong multimodal capabilities.
Claude Sonnet 5 is Anthropic's latest Sonnet model, substantially improving on Sonnet 4.6 in coding and agentic work and reaching near-Opus quality on many tasks. It features a 1M token context window, adaptive thinking, and strong document and vision understanding.
DeepSeek-V3.2 is an efficient large language model with DeepSeek Sparse Attention (DSA) for long contexts. It features strong reasoning and tool-use skills, achieving top results on the 2025 IMO and IOI.
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities, including structured outputs and function calling. Gemma 3 27B is Google's latest open source model, successor to Gemma 2.
Gemma 4 26B A4B is a Mixture-of-Experts model from Google DeepMind with 26B total parameters and only 4B active per token, offering fast inference at high quality. It handles text, image, and video input, supports 256K context, function calling, and reasoning with configurable thinking modes.
Gemma 4 31B is a dense model from Google DeepMind with 31B parameters, delivering frontier-level reasoning performance. It handles text, image, and video input, supports 256K context, function calling, and configurable thinking modes.
MiniMax-M2.5 is a state-of-the-art large language model optimized for coding, agentic workflows, and modern application development with enhanced reasoning capabilities.
Kimi K2.5 is Moonshot AIs most advanced open reasoning model, featuring trillion-parameter Mixture-of-Experts architecture with 32B active parameters and 256K context windows.
NVIDIA Nemotron 3 Nano 30B is a compact and efficient language model from NVIDIA, optimized for fast inference while maintaining strong performance across diverse tasks.
GPT OSS 20B running in a Trusted Execution Environment (TEE). OpenAI's compact open-weight 21B MoE model with 3.6B active parameters, optimized for lower-latency inference, with hardware attestation evidence available for independent verification.
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. The model supports configurable reasoning depth, full chain-of-thought access, and native tool use, including function calling, browsing, and structured output generation
Built for in-depth research and handling long, complex documents. Ideal for technical work, multimodal input, and high-precision tasks.
Turbo variant of Qwen3 Coder 480B, optimized for faster inference on code tasks.
Optimized for speed and efficiency.
Qwen3-VL 235B vision-language model with MoE architecture. The most powerful VL model in the Qwen series with superior visual perception, OCR, and multimodal reasoning.
Grok 4.3 is xAI's most intelligent and fastest reasoning model with function calling, structured outputs, and a 1M-token context window. Suited for agentic workflows, instruction-following tasks, and applications requiring high factual accuracy.
GLM-4.6 is a large language model developed by Zhiyuan AI, featuring strong reasoning capabilities and support for multiple languages. Supports the largest context window for processing extensive text and detailed analysis.
GLM-4.7 is a large language model developed by Zhiyuan AI, featuring strong reasoning capabilities and support for multiple languages. Supports the largest context window for processing extensive text and detailed analysis.
GLM-4.7-Flash is a fast inference variant of GLM-4.7, optimized for speed while maintaining strong reasoning capabilities. Ideal for applications requiring quick responses with good quality.
GLM-5 is the next-generation large language model developed by Zhiyuan AI, featuring significantly enhanced reasoning capabilities, improved instruction following, and support for multiple languages. Supports large context windows for processing extensive text and detailed analysis.
Gemma 3 4B IT through Amazon Bedrock Mantle with zero retention.
Mistral Large 3 675B Instruct through Amazon Bedrock Mantle with zero retention.
Qwen3 Coder Next through Amazon Bedrock Mantle with zero retention.
Claude Haiku 4.5 served through Amazon Bedrock with support for AWS zero-data-retention mode.
DeepSeek V3.1 served through Amazon Bedrock with support for AWS zero-data-retention mode.
Gemma 3 12B IT served through Amazon Bedrock with support for AWS zero-data-retention mode.
Gemma 4 E2B served through Amazon Bedrock with support for AWS zero-data-retention mode.
MiniMax M2 served through Amazon Bedrock with support for AWS zero-data-retention mode.
MiniMax M2.1 served through Amazon Bedrock with support for AWS zero-data-retention mode.
Devstral 2 123B served through Amazon Bedrock with support for AWS zero-data-retention mode.
Magistral Small 2509 served through Amazon Bedrock with support for AWS zero-data-retention mode.
Ministral 3 3B Instruct served through Amazon Bedrock with support for AWS zero-data-retention mode.
Ministral 3 8B Instruct served through Amazon Bedrock with support for AWS zero-data-retention mode.
Ministral 3 14B Instruct served through Amazon Bedrock with support for AWS zero-data-retention mode.
Voxtral Mini 3B 2507 served through Amazon Bedrock with support for AWS zero-data-retention mode.
Voxtral Small 24B 2507 served through Amazon Bedrock with support for AWS zero-data-retention mode.
Kimi K2 Thinking served through Amazon Bedrock with support for AWS zero-data-retention mode.
Nemotron Nano 9B V2 served through Amazon Bedrock with support for AWS zero-data-retention mode.
Nemotron Nano 12B V2 served through Amazon Bedrock with support for AWS zero-data-retention mode.
Nemotron Super 3 120B served through Amazon Bedrock with support for AWS zero-data-retention mode.
GPT-OSS Safeguard 20B served through Amazon Bedrock with support for AWS zero-data-retention mode.
GPT-OSS Safeguard 120B served through Amazon Bedrock with support for AWS zero-data-retention mode.
Qwen3 32B served through Amazon Bedrock with support for AWS zero-data-retention mode.
Qwen3 Coder 30B A3B Instruct served through Amazon Bedrock with support for AWS zero-data-retention mode.
Palmyra Vision 7B served through Amazon Bedrock with support for AWS zero-data-retention mode.