AI Pulse by Inblix

Llama 3.1 8B

By Meta · Visit provider →

Llama 3.1 8B is the most affordable model in the lineup, perfect for high-throughput, simple tasks. At just $0.05 per 1M input tokens, it's ideal when cost is the primary concern and tasks are straightforward.

Input Price
$0.05
per 1M tokens
Output Price
$0.08
per 1M tokens
Context Window
128K
tokens
Calculate cost for Llama 3.1 8B

Key Strengths

  • Extremely affordable pricing
  • Very high throughput
  • Good enough for simple tasks
  • Wide ecosystem support for fine-tuning

Best Use Cases

High-volume text classificationSimple Q&A and FAQ systemsPre-filtering and routing pipelinesEdge and on-device deployments

Pricing Details

Metric Value
Input (per 1M tokens) $0.05
Output (per 1M tokens) $0.08
Context window 128K tokens
Provider Meta
Architecture 8B-parameter dense transformer, 128K context, optimized for maximum throughput

Ultra-cheap, via Groq

Benchmark Performance

MMLU-Pro 36.5 / 100
GPQA Diamond 27 / 100
MATH-500 21.8 / 100
LiveCodeBench 8.5 / 100
Humanity's Last Exam 4.3 / 100
Intelligence Index 1.9 / 100
AIME 2025 0 / 100

Source: Artificial Analysis (independent benchmarking). Updated 2026-08-08.

Recent AI News

Get smarter about AI

The sharpest AI news, curated daily. Delivered free to your inbox.

Pricing data is sourced from official API documentation and updated weekly. Actual costs may vary based on prompt caching, batch processing, service tiers, and regional pricing. Inblix is not affiliated with Meta.