Llama 4 Maverick API

Advanced multimodal AI model with superior reasoning and coding capabilities
Output

How to use Llama 4 Maverick API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to .

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Llama 4 Maverick API Pricing

TypePrice
Input
Output

Llama 4 Maverick Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
Intelligence
9.3
Composite score across standardised reasoning, knowledge and problem-solving evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Coding
16.3
Composite score across standardised coding evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026
Math
19.3
Composite score across standardised mathematics evaluations, measured independently by Artificial AnalysisSourceSeptember 12, 2026

Llama 4 Maverick vs other models

ModelInputOutputContextBest for
$6.5 / 1M tokens
$39 / 1M tokens
1.05M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$0.65 / 1M tokens
$3.9 / 1M tokens
1.05M tokens
Reasoning + agents

Frequently asked questions

Yes, Llama 4 Maverick can stream responses as they are generated.

Yes, Llama 4 Maverick supports both function calling and structured outputs.

Llama 4 Maverick was built by Meta.

Llama 4 Maverick is accessed through the v1/chat/completions endpoint.

Start building with Llama 4 Maverick

Get API Key
1000+ models, one API.