Skip to content
Dashboard

Llama 4 Scout 17B 16E Instruct

Llama 4 Scout 17B 16E Instruct is a natively multimodal Mixture of Experts (MoE) model with a context window of 131.1K tokens, purpose-built for processing entire codebases, multi-document corpora, and extended user activity logs in a single inference call.

Tool UseVision (Image)
index.ts
import { streamText } from 'ai'
const result = streamText({
model: 'meta/llama-4-scout',
prompt: 'Why is the sky blue?'
})

Providers

Route requests across multiple providers. Copy a provider slug to set your preference. Visit the docs for more info. Using a provider means you agree to their terms, listed under Legal.

Provider
Context
Max Output
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
ZDR
No Training
Release Date
DeepInfra
Legal:Terms
Privacy
131K8K
0.3s
47tps
$0.10/M
$0.30/M
04/05/2025
Groq
Legal:Terms
Privacy
131K8K
$0.11/M
$0.34/M
04/05/2025
Bedrock
Legal:Terms
Privacy
128K8K
0.2s
199tps
$0.17/M
$0.66/M
04/05/2025