AI Pulse by Inblix

Topic: Llama

4 articles

Explore our coverage of Llama — 4 curated articles, summaries, and related resources from the Inblix archive.

Related Terms

AMD buys Taalas to bake Llama 3.1 into silicon at 16,000 tokens per second — Inblix summary
AI News

AMD buys Taalas to bake Llama 3.1 into silicon at 16,000 tokens per second

The Decoder · Aug 7, 2026 · 2 min read

AMD is acquiring Toronto-based startup Taalas, a company that came out of stealth just last February with a genuinely w...

NVIDIA drops Llama Nemotron 8B VLM to dethrone OCR leaders on Hugging Face — Inblix summary
Research

NVIDIA drops Llama Nemotron 8B VLM to dethrone OCR leaders on Hugging Face

Hugging Face Blog · Jun 27, 2025 · 3 min read

NVIDIA just dropped a focused bomb on the Vision Language Model space, and it’s aimed squarely at the unglamorous but m...

How Llama 3.2's RoPE encoding evolved from a broken integer trick — Inblix summary
Research

How Llama 3.2's RoPE encoding evolved from a broken integer trick

Hugging Face Blog · Nov 25, 2024 · 2 min read

Most people use Rotary Positional Encoding (RoPE) without thinking twice. It's baked into Llama 3.2, Mistral, and pract...

Hugging Face and Intel squeeze Llama 3.1 into 4-bit for edge devices with OpenVINO — Inblix summary
Research

Hugging Face and Intel squeeze Llama 3.1 into 4-bit for edge devices with OpenVINO

Hugging Face Blog · Sep 20, 2024 · 2 min read

Running a model like Meta-Llama-3.1-8B on a laptop or edge device is moving from 'technically possible' to genuinely pr...