Meta Llama

Llama 3.1 405B Instruction

Specifications, pricing details, deployment targets, and capability limits compiled from the canonical database.

Technical Overview

Model FamilyLlama 3.1
Modalitytext-to-text
Release Generation405B
Context Window128,000 tokens
Max Output Tokens4,096 tokens

Token Pricing (per 1M tokens)

Input Cost$5.30
Output Cost$16.00

Pricing metrics are calibrated for standard server instances. Local deployments (e.g. Ollama) represent $0.00 infrastructure costs.

Supported Capabilities

💻 Coding Assistant🧩 Logical Reasoning🔧 Tool Calling📄 Structured JSON

Deployment & Availability

Hosting ChannelsAWS Bedrock, GCP Vertex AI, Azure AI Studio
Geographic RegionsUS East (N. Virginia), US West (Oregon), US Central 1 (Iowa), East US

Best Use Cases

Ideal for automated code refactoring, logic generation, syntax checks, and CI/CD pipelines. Matches large repository analysis and multi-document parsing workloads.

Other Models from Meta Llama

Llama 3.1 70B InstructionLlama 3.1
⚖ Compare with other modelsLaunch AI Model Selector