Fireworks

Fireworks AI

inference provider · 7 models

Access 7 models served through Fireworks AI on AnonRouter's privacy-first gateway, including DeepSeek V4 Flash, DeepSeek V4 Pro, and MiniMax M2.7. Fireworks open-model Chat Completions use its default zero-data-retention policy; service metadata is logged and prompt caches may remain briefly in volatile memory.

Models

7

Modalities

1

Text

From (input)

$0.14

per 1M tokens

Max context

1M

Private routes

7 / 7

not anonymous-only

Catalog by modality

7 routes
Text7

Fireworks AI models7

DeepSeek
DeepSeek V4 Flash
deepseek/deepseek-v4-flash
Private

DeepSeek V4 Flash running in a Trusted Execution Environment (TEE). Hardware attestation evidence is available for independent verification of enclave identity and configuration.

Private|1M context|$0.14/M input|$0.28/M output
DeepSeek
DeepSeek V4 Pro
deepseek/deepseek-v4-pro
Private

DeepSeek V4 Pro is a 1.6T-parameter Mixture-of-Experts model with 49B active parameters and a 1M-token context window. Built for advanced reasoning, coding, and long-horizon agentic workflows with a hybrid attention system for efficient long-context processing.

Private|1M context|$1.74/M input|$3.48/M output
Minimax
MiniMax M2.7
minimax/minimax-m2.7
Private

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity with advanced agentic capabilities through multi-agent collaboration.

Private|197K context|$0.30/M input|$1.20/M output
Kimi
Kimi K2.6
moonshotai/kimi-k2.6
Private

Kimi K2.6 is an open-source, native multimodal agentic model from Moonshot AI with 1T total parameters and 32B active parameters. It excels at long-horizon coding, coding-driven design, agent swarm orchestration, and proactive autonomous execution with 256K context windows.

Private|262K context|$0.95/M input|$4.00/M output
Kimi
Kimi K2.7 Code
moonshotai/kimi-k2.7-code
Private

Kimi K2.7 Code is Moonshot AI's coding-focused agentic model built on Kimi K2.6, with 1T total parameters and 32B active parameters. It always operates in thinking mode, supports text and image input, and targets long-horizon software engineering, agentic task decomposition, and multi-turn coding workflows with 256K context.

Private|262K context|$0.95/M input|$4.00/M output
Kimi
Kimi K3
moonshotai/kimi-k3
Private

Kimi K3 is an ultra-large-scale, open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at navigating large repositories, using tools, debugging, and iterating against images, logs, tests, and runtime feedback.

Private|1M context|$3.00/M input|$15.00/M output
Zhipu
GLM 5.2
z-ai/glm-5.2
Private

GLM 5.2 running in a Trusted Execution Environment (TEE). Z.AI's flagship model for long-horizon tasks with enhanced reasoning and project-level engineering context, with hardware attestation evidence available for independent verification.

Private|1M context|$1.40/M input|$4.40/M output
Provider documentation3 sources
Fireworks AI provider — AnonRouter