Meta

Meta

meta-llama/ · 10 models

Access 10 Meta models through AnonRouter's privacy-first gateway, including Hermes 3 Llama 3.1 405b, Llama 3.2 3B, and Llama 3.3 70B. Compare pricing, context windows, and capabilities across Meta's text routes — every request anonymized, with no payload logging.

Models

10

Modalities

1

Text

From (input)

$0.02

per 1M tokens

Max context

1M

Private routes

10 / 10

not anonymous-only

Catalog by modality

10 routes
Text10

Meta models10

Meta
Hermes 3 Llama 3.1 405b
meta-llama/hermes-3-llama-3.1-405b
Private

Hermes 3 405B is a frontier level, full parameter finetune of the Llama-3.1 405B foundation model, focused on aligning LLMs to the user, with powerful steering capabilities and control given to the end user.

Private|128K context|$1.10/M input|$3.00/M output
Meta
Llama 3.2 3B
meta-llama/llama-3.2-3b
Private

Llama 3.2 3B is a text model.

Private|128K context|$0.15/M input|$0.60/M output
Meta
Llama 3.3 70B
meta-llama/llama-3.3-70b
TEE

Llama 3.3 70B is a text model.

TEE|128K context|$0.70/M input|$2.80/M output
Meta
Llama 3.3 70B Instruct Turbo
meta-llama/llama-3.3-70b-instruct-turbo
Private

Llama 3.3-70B Turbo is a highly optimized version of the Llama 3.3-70B model, utilizing FP8 quantization to deliver significantly faster inference speeds with a minor trade-off in accuracy. The model is designed to be helpful, safe, and flexible, with a focus on responsible deployment and mitigating potential risks such as bias, toxicity, and misinformation. It achieves state-of-the-art performance on various benchmarks, including conversational tasks, language translation, and text generation.

Private|131K context|$0.10/M input|$0.32/M output
Meta
Llama 4 Maverick 17B 128E Instruct FP8
meta-llama/llama-4-maverick-17b-128e-instruct-fp8
Private

The Llama 4 collection of models are natively multimodal AI models that enable text and multimodal experiences. These models leverage a mixture-of-experts architecture to offer industry-leading performance in text and image understanding. Llama 4 Maverick, a 17 billion parameter model with 128 experts

Private|1M context|$0.20/M input|$0.80/M output
Meta
Llama 4 Scout 17B 16E Instruct
meta-llama/llama-4-scout-17b-16e-instruct
Private

The Llama 4 collection of models are natively multimodal AI models that enable text and multimodal experiences. These models leverage a mixture-of-experts architecture to offer industry-leading performance in text and image understanding. Llama 4 Scout, a 17 billion parameter model with 16 experts

Private|328K context|$0.10/M input|$0.30/M output
Meta
Llama Guard 4 12B
meta-llama/llama-guard-4-12b
Private

Llama Guard 4 is a natively multimodal safety classifier with 12 billion parameters trained jointly on text and multiple images. Llama Guard 4 is a dense architecture pruned from the Llama 4 Scout pre-trained model and fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM inputs (prompt classification) and in LLM responses (response classification). It itself acts as an LLM: it generates text in its output that indicates whether a given prompt or response is safe or unsafe, and if unsafe, it also lists the content categories violated.

Private|164K context|$0.18/M input|$0.18/M output
Meta
Meta Llama 3.1 70B Instruct Turbo
meta-llama/meta-llama-3.1-70b-instruct-turbo
Private

Meta developed and released the Meta Llama 3.1 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8B, 70B and 405B sizes

Private|131K context|$0.40/M input|$0.40/M output
Meta
Meta Llama 3.1 8B Instruct Turbo
meta-llama/meta-llama-3.1-8b-instruct-turbo
Private

Meta developed and released the Meta Llama 3.1 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8B, 70B and 405B sizes

Private|131K context|$0.02/M input|$0.04/M output
Meta
Muse Glimmer 30B
meta-models/muse-glimmer-30b
Private

Muse Glimmer is a 30B multimodal agentic model distilled from Muse Spark — reasoning, tool use, and failure recovery in a single model that runs locally on consumer hardware.

Private|131K context|$0.30/M input|$1.20/M output
Meta models — AnonRouter