Research
AMD's 192-core Turin chip doubles Llama 3 speed, no GPU required
Hugging Face Blog · Oct 10, 2024 · 2 min read
The era of running serious LLM inference without a GPU just got a lot more interesting. AMD's 5th Gen EPYC "Turin" proc...
2 articles
Explore our coverage of Llama-3 — 2 curated articles, summaries, and related resources from the Inblix archive.
Hugging Face Blog · Oct 10, 2024 · 2 min read
The era of running serious LLM inference without a GPU just got a lot more interesting. AMD's 5th Gen EPYC "Turin" proc...
Hugging Face Blog · Sep 18, 2024 · 2 min read
Microsoft Research’s BitNet architecture promised a world where LLM parameters are just -1, 0, or 1, slashing memory an...