Llama 4 Scout by Meta demonstrates competitive pricing. View detailed benchmark data including scores across coding, math, reasoning, speed, and cost metrics.
Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.
General Benchmarks
Human preference ELO from blind head-to-head votes
LMSYS / HuggingFaceArtificial Analysis composite intelligence score across 10 sub-benchmarks
Artificial Analysis (via OpenRouter)Massive Multitask Language Understanding professional benchmark
HuggingFace LeaderboardCoding Benchmarks
Human preference ELO for building a working web page from a brief
Design Arena (via OpenRouter)Artificial Analysis composite coding score
Artificial Analysis (via OpenRouter)Human preference ELO for building things — websites, UI, games, charts, SVG
Design Arena (via OpenRouter)Math Benchmarks
Reasoning Benchmarks
Artificial Analysis composite score for multi-step tool-using tasks
Artificial Analysis (via OpenRouter)Speed Benchmarks
Cost Benchmarks
Context Benchmarks
Available from 2 providers
The same model costs different amounts depending on who serves it. You can bring your own key for 1 of these — connect it here.
| Provider | Input / 1M | Output / 1M | Cached in | Context | Your key |
|---|---|---|---|---|---|
| NanoGPT | $0.085 | $0.46 | $0.0425 | 328K | — |
| OpenRouter | $0.1 | $0.3 | — | 1.3M | Supported |
Llama 4 Scout — Benchmark Scores Overview
Scores normalized to percentage scale for visual comparison. ELO scores mapped to 0-100 range (1100-1500).
Compare Llama 4 Scout With
Llama 4 Scout — Frequently Asked Questions
How intelligent is Llama 4 Scout?
Llama 4 Scout scores 1240 on the Chatbot Arena ELO rating, making it an entry-level AI model. This score is based on blind head-to-head human preference voting.
How much does Llama 4 Scout cost?
Llama 4 Scout costs $0.10 per 1M input tokens and $0.30 per 1M output tokens. This makes it one of the more affordable models.
How fast is Llama 4 Scout?
Llama 4 Scout generates output at 140 tokens per second, which is moderate compared to other models. The time to first token is 330 ms.
How good is Llama 4 Scout at coding?
Llama 4 Scout achieves 28.0% on SWE-bench Verified, demonstrating basic real-world software engineering capability. This benchmark tests the model's ability to resolve actual GitHub issues.
How good is Llama 4 Scout at math and reasoning?
Llama 4 Scout scores 68.0% on the MATH benchmark (competition-level mathematics). It also achieves 38.0% on GPQA Diamond, a graduate-level science reasoning benchmark.
What is the context window of Llama 4 Scout?
Llama 4 Scout has a context window of 1.3M tokens. This determines how much text, conversation history, and code the model can process in a single request.
Who created Llama 4 Scout?
Llama 4 Scout was created by Meta. It is classified as a open source model in our catalogue.
Is Llama 4 Scout open source?
Yes, Llama 4 Scout is an open-source model. The model weights are publicly available for download and self-hosting.