Llama 3.1 70B Instruct Turbo (Deprecated) API

Llama 3.1 70B Instruct Turbo: Meta's advanced, multilingual, instruction-tuned language model for diverse, high-accuracy natural language tasks.
Output

How to use Llama 3.1 70B Instruct Turbo (Deprecated) API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to .

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Llama 3.1 70B Instruct Turbo (Deprecated) API Pricing

TypePrice
Input
Output

Llama 3.1 70B Instruct Turbo (Deprecated) vs other models

ModelInputOutputContextBest for
$6.5 / 1M tokens
$39 / 1M tokens
1.05M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$0.65 / 1M tokens
$3.9 / 1M tokens
1.05M tokens
Reasoning + agents

Frequently asked questions

Yes, Llama 3.1 70B Instruct Turbo (Deprecated) can stream responses as they are generated.

Yes, Llama 3.1 70B Instruct Turbo (Deprecated) supports both function calling and structured outputs.

Llama 3.1 70B Instruct Turbo (Deprecated) was built by Meta.

No. The listed capabilities are function calling, streaming, and structured outputs; vision is not included.

This model is marked as deprecated.

Start building with Llama 3.1 70B Instruct Turbo (Deprecated)

Get API Key
1000+ models, one API.