AI Pulse by Inblix

Topic: CUDA

7 articles

Explore our coverage of CUDA — 7 curated articles, summaries, and related resources from the Inblix archive.

OpenAI releases Triton 1.0, a Python-like language for GPU kernels — Inblix summary
Product

OpenAI releases Triton 1.0, a Python-like language for GPU kernels

OpenAI Blog · Jul 19, 2026 · 2 min read

OpenAI is shipping Triton 1.0 to the public, and it's the kind of tool that could quietly rewrite how a lot of neural n...

PyTorch profiling isn't just for experts: Here's how to read your first trace — Inblix summary
Research

PyTorch profiling isn't just for experts: Here's how to read your first trace

Hugging Face Blog · May 29, 2026 · 2 min read

Most PyTorch users know they should profile their models. Few actually do it. The barrier isn't a lack of tools — it's...

Synchronous batching wastes 24% of GPU time: Here's how async cuts it to zero — Inblix summary
Research

Synchronous batching wastes 24% of GPU time: Here's how async cuts it to zero

Hugging Face Blog · May 14, 2026 · 2 min read

If you're running inference on an H200 at $5 an hour, every second the GPU sits idle is money you're lighting on fire....

AMD's new Helios rack out-specs Nvidia's Vera Rubin on paper, targets CUDA with ROCm — Inblix summary
Industry

AMD's new Helios rack out-specs Nvidia's Vera Rubin on paper, targets CUDA with ROCm

The Register AI · May 5, 2026 · 2 min read

AMD is no longer just nipping at Nvidia's heels—it's swinging for the fences, and doing it on two fronts. First, the co...

Anthropic's Opus 5 ships at half price, drops mandatory data retention — Inblix summary
Industry

Anthropic's Opus 5 ships at half price, drops mandatory data retention

The Register AI · Apr 28, 2026 · 2 min read

Anthropic made a move this week that should have enterprise buyers paying close attention. They quietly launched Opus 5...

Hugging Face Taught Claude and Codex to Write Production CUDA Kernels — Inblix summary
Research

Hugging Face Taught Claude and Codex to Write Production CUDA Kernels

Hugging Face Blog · Feb 13, 2026 · 2 min read

Writing fast CUDA kernels is a dark art. It's not just about knowing C++ — it's about wrangling GPU memory hierarchies,...

Hugging Face's Kernel Builder Turns Solo CUDA Code Into Shared, Production-Ready Python Packages — Inblix summary
Research

Hugging Face's Kernel Builder Turns Solo CUDA Code Into Shared, Production-Ready Python Packages

Hugging Face Blog · Aug 18, 2025 · 2 min read

Writing a fast CUDA kernel for a specific GPU is one thing. Making it build cleanly for multiple architectures, survive...