Unsloth crushes MoE training with 7.3x speedup, Axolotl hits 1.45x on Qwen3.5
MarkTechPost · Jul 22, 2026 · 2 min read
The fine-tuning wars aren't about which library you use anymore. They're about where each project places its engineerin...
35 articles
Explore our coverage of fine-tuning — 35 curated articles, summaries, and related resources from the Inblix archive.
MarkTechPost · Jul 22, 2026 · 2 min read
The fine-tuning wars aren't about which library you use anymore. They're about where each project places its engineerin...
OpenAI Blog · Jul 19, 2026 · 2 min read
OpenAI has published new research demonstrating that a language model's behavior can be significantly steered by fine-t...
OpenAI Blog · Jul 18, 2026 · 2 min read
OpenAI is cracking open GPT-3's black box, letting any API customer train custom versions of the model on their own dat...
OpenAI Blog · Jul 18, 2026 · 2 min read
You train a dog by rewarding good behavior, not by programming every single command it will ever hear. That's the analo...
OpenAI Blog · Jul 17, 2026 · 2 min read
OpenAI just handed developers the keys to one of its most popular models. Fine-tuning for GPT-3.5 Turbo is live, lettin...
OpenAI Blog · Jul 17, 2026 · 2 min read
OpenAI delivered a one-two punch to developers Thursday: a suite of long-requested fine-tuning API improvements landed,...
OpenAI Blog · Jul 17, 2026 · 2 min read
OpenAI just pulled the trigger on two moves that will reshape how developers build with its tech: GPT-4 is now generall...
OpenAI Blog · Jul 16, 2026 · 2 min read
The math behind job boards has always been a bit of a black box. You upload a resume, you get a list of openings, and y...
OpenAI Blog · Jul 16, 2026 · 2 min read
OpenAI finally gave developers what they've been asking for: fine-tuning for GPT-4o. The announcement on August 20, 202...
OpenAI Blog · Jul 15, 2026 · 2 min read
OpenAI just stitched together its entire model distillation pipeline into a single, integrated workflow, and it’s a mov...
OpenAI Blog · Jul 15, 2026 · 2 min read
OpenAI finally plugged the glaring hole in its fine-tuning API. Starting today, you can feed images—not just text—into...
OpenAI Blog · Jul 15, 2026 · 2 min read
Decagon, the 2023-founded startup that already counts Eventbrite, Notion, and Duolingo among its customers, just droppe...
OpenAI Blog · Jul 15, 2026 · 2 min read
OpenAI just made its most capable reasoning model, o1, available to developers on usage tier 5, packing in production-r...
OpenAI Blog · Jul 15, 2026 · 2 min read
The grind of junior banking analysts — the endless nights combing through SEC filings, pitch decks, and private data ro...
OpenAI Blog · Jul 14, 2026 · 2 min read
OpenAI just announced the Pioneers Program, a hands-on initiative that pairs its research teams directly with startups...
Hugging Face Blog · Jun 24, 2026 · 2 min read
The gap between what a general-purpose library offers and what frontier models actually need has become a chasm. NVIDIA...
Hugging Face Blog · Jun 18, 2026 · 2 min read
If you've fine-tuned an open-source model recently, you almost certainly used LoRA. And you're not alone. An analysis o...
Hugging Face Blog · Mar 31, 2026 · 2 min read
The team behind TRL just tagged v1.0, but don't call it a finished product. It's an acknowledgment of a reality that's...
Hugging Face Blog · Feb 20, 2026 · 2 min read
The barrier to training your own customized AI model just collapsed. Hugging Face and Unsloth are teaming up to effecti...
Hugging Face Blog · Dec 11, 2025 · 2 min read
OpenAI is handing its coding agent, Codex, the keys to the Hugging Face kingdom, and the implications are a bit dizzyin...
Hugging Face Blog · Dec 4, 2025 · 2 min read
Anthropic's Claude just got a practical new ability that sidesteps the usual notebook grind: it can now fine-tune open-...
Hugging Face Blog · Sep 10, 2025 · 2 min read
The gap between finding a useful model and actually making it your own has been a persistent headache. Together AI and...
Hugging Face Blog · Jun 19, 2025 · 2 min read
For all the staggering image generation capabilities of models like black-forest-labs/FLUX.1-dev, the hardware wall for...
Hugging Face Blog · Jun 11, 2025 · 2 min read
NVIDIA just lowered the barrier to entry for generalist robot AI in a big way. The company announced Isaac GR00T N1.5,...
Hugging Face Blog · May 25, 2025 · 2 min read
Another day, another obscure shape mismatch tearing through a perfectly good training run. This time, the culprit is a...
Hugging Face Blog · May 15, 2025 · 2 min read
The Falcon-Edge series tackles the central headache of extreme model compression head-on: you usually have to pick betw...
Hugging Face Blog · Apr 22, 2025 · 2 min read
Pipeline-based OCR engines have a dirty secret: they often butcher the reading order, especially on layout-rich pages....
Hugging Face Blog · Mar 26, 2025 · 2 min read
You don't need a massive, general-purpose reranker to get state-of-the-art search results. A new technical deep-dive an...
Hugging Face Blog · Jan 31, 2025 · 2 min read
It’s almost jarring to remember that just a year ago, we were all laughing at AI’s inability to count fingers. The pace...
Hugging Face Blog · Dec 3, 2024 · 2 min read
Capital Fund Management, the $15.5 billion quantitative hedge fund, faced a messy but critical problem: the company tag...
Hugging Face Blog · Nov 4, 2024 · 2 min read
Argilla 2.4 just dropped, and it rips out the last big barrier between a domain expert and a high-quality fine-tuning d...
Hugging Face Blog · Oct 21, 2024 · 2 min read
If you've been waiting to use Meta's new Llama 3.2 models in Keras, you can stop. Martin Görner from Google confirmed t...
Hugging Face Blog · Sep 18, 2024 · 2 min read
Microsoft Research’s BitNet architecture promised a world where LLM parameters are just -1, 0, or 1, slashing memory an...
Hugging Face Blog · Jul 18, 2024 · 3 min read
The team behind Idefics2 just released Docmatix, a document visual question answering dataset that makes the previous s...
Hugging Face Blog · Jul 11, 2024 · 2 min read
The Numina team has just pulled off something remarkable: a fine-tuned 7-billion-parameter model that out-reasoned much...