Llama 3.3 70B Instruct

llama-3.3-70b-instruct · Llama

API Pricing

Input$0.6 / 1M tokens
Output$1.2 / 1M tokens

Specifications

Context131K tokens
Modalitiestext
Featureslong context

Frequently asked questions

What is the context length of Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct has a 131,072 token context window.

How much does Llama 3.3 70B Instruct cost?

On AIHubMix, Llama 3.3 70B Instruct costs $0.6 per million input tokens and $1.2 per million output tokens.

What modalities does Llama 3.3 70B Instruct support?

Llama 3.3 70B Instruct accepts text input.

What capabilities does Llama 3.3 70B Instruct support?

Llama 3.3 70B Instruct supports long context. Per-protocol parameter support is listed in the capability table on this page.

How do I call Llama 3.3 70B Instruct via API?

Llama 3.3 70B Instruct is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to llama-3.3-70b-instruct — no other code changes needed.

Who created Llama 3.3 70B Instruct?

Llama 3.3 70B Instruct is developed by Llama. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was Llama 3.3 70B Instruct released?

Llama 3.3 70B Instruct was released on December 6, 2024 by Llama.

More models from Llama

See all Llama models →

Llama 4 Maverick

by Llama

Llama 4 Maverick is a high-capacity Mixture-of-Experts (MoE) model from Meta, featuring…

$0.2/1M in · $0.2/1M out
1,048,576 tokens context

Llama 4 Scout

by Llama

Llama 4 Scout is a highly efficient Mixture-of-Experts (MoE) model from Meta, activating…

$0.2/1M in · $0.2/1M out
131,072 tokens context

Llama 3.3 70B

by Llama

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and…

$0.6/1M in · $0.6/1M out
65,536 tokens context

Llama 3.1 70B

by Llama
$0.44/1M in · $0.44/1M out

Llama3.1 8B

by Llama

cerebras

$0.3/1M in · $0.6/1M out

Cerebras Llama 3.3 70B

by Llama
$0.6/1M in · $0.6/1M out

Use Llama 3.3 70B Instruct via the AIHubMix unified API — one interface for every major LLM.