AI Pulse by Inblix

Topic: prompt-caching

2 articles

Explore our coverage of prompt-caching — 2 curated articles, summaries, and related resources from the Inblix archive.

LoRA fine-tuning slashes agent costs by 95%, but prompt caching wins on latency — Inblix summary
Research

LoRA fine-tuning slashes agent costs by 95%, but prompt caching wins on latency

Machine Learning Mastery · Aug 10, 2026 · 2 min read

The push to move AI agents from prototypes to production has collided with two unforgiving walls: spiraling API costs a...

OpenAI ships GPT-5.1: faster thinking, cheaper tokens, and tools that actually ship code — Inblix summary
Product

OpenAI ships GPT-5.1: faster thinking, cheaper tokens, and tools that actually ship code

OpenAI Blog · Jul 12, 2026 · 2 min read

OpenAI is rolling out GPT-5.1 in the API platform today, and the headline isn't just about a bigger model—it's about a...