Skip to content

Meta

9 modelsModel creator

Models published by Meta, available through the AnyRouter API. Each can route across multiple upstream providers for availability and price.

131K

Meta's Llama 3.1 70B instruction-tuned model with strong reasoning and multilingual capabilities.

Terms1
TTFT 600ms50 tok/s
131K

Meta's compact Llama 3.1 8B instruction-tuned model optimized for fast inference and edge deployments.

Terms2
TTFT 200ms150 tok/s
131K

Meta's Llama 3.3 70B instruction-tuned model for high-quality chat, reasoning, and coding at a fraction of frontier cost.

131K

Meta's Llama 4 Scout with 17B parameters and 16 experts, featuring native multimodal support with 10M context window via interleaved attention.

Terms2
TTFT 400ms80 tok/s
164KZDR

Llama Guard 4 is a 12B natively multimodal safety classifier for content safety classification on LLM inputs and responses. It acts as an LLM itself: it generates text indicating whether a prompt or response is safe or unsafe and, if unsafe, lists the violated categories (MLCommons hazards taxonomy S1–S14). Use it as a guardrail/judge hop in front of or behind chat models.

131KFree

Meta Muse Glimmer 30B is a ~30B dense multimodal instruct model (text and image in, text out) with a 131K context window, chain-of-thought reasoning, and function calling. Available on NVIDIA NIM for chat and agent workflows.

1M

Meta's Muse Spark 1.1 is a multimodal reasoning model built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text,a 1M-token context window. The model orchestrates multi-agent workflows — acting as a main agent that plans and delegates or as a subagent — and generalizes zero-shot to new tools, MCP servers, and custom skills. It supports structured output, parallel function calling, built-in searchcitations, and configurable reasoning effort. Availablea. NOTE: currently available to users in the United States only.

1M

Meta's Muse Spark 1.2 is a multimodal reasoning model for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context window.

1M

Meta's Muse Spark 1.3 is a multimodal reasoning model for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context window.