Models
21 models
Meta: Llama 3.1 70B Instruct
$0.400/M input
$0.400/M output
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...
Meta: Llama 3.1 8B Instruct
$0.050/M input
$0.080/M output
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
Meta: Llama 3.2 1B Instruct
$0.027/M input
$0.201/M output
Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...
Meta: Llama 3.2 3B Instruct
$0.050/M input
$0.330/M output
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
Meta: Llama 3.3 70B Instruct
$0.100/M input
$0.320/M output
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
Meta: Llama 4 Maverick
$0.200/M input
$0.696/M output
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Meta: Llama 4 Scout
$0.100/M input
$0.300/M output
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
Meta: Llama Guard 4 12B
$0.180/M input
$0.180/M output
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
Meta: Muse Glimmer 30B
$0.350/M input
$1.50/M output
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
Meta: Muse Spark 1.1
$1.25/M input
$4.25/M output
Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...
Meta: Muse Spark 1.2
$1.25/M input
$4.25/M output
Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...
meta/codellama-70b
— input
— output
meta/llama-3.1-70b-instruct
— input
— output
meta/llama-3.1-8b-instruct
— input
— output
meta/llama-3.2-11b-vision-instruct
— input
— output
meta/llama-3.2-1b-instruct
— input
— output
meta/llama-3.2-3b-instruct
— input
— output
meta/llama-3.2-90b-vision-instruct
— input
— output
meta/llama-3.3-70b-instruct
— input
— output
meta/llama-guard-4-12b
— input
— output
meta/llama2-70b
— input
— output