AI Pulse by Inblix

Topic: multimodal

14 articles

Explore our coverage of multimodal — 14 curated articles, summaries, and related resources from the Inblix archive.

Related Terms

ChatGPT gets eyes and ears, rolling out now to paid users — Inblix summary
Product

ChatGPT gets eyes and ears, rolling out now to paid users

OpenAI Blog · Jul 17, 2026 · 2 min read

OpenAI is giving ChatGPT the ability to see, hear, and speak, starting with Plus and Enterprise subscribers over the ne...

OpenAI launches GPT-4o, gives free ChatGPT users the good stuff — Inblix summary
Product

OpenAI launches GPT-4o, gives free ChatGPT users the good stuff

OpenAI Blog · Jul 17, 2026 · 2 min read

OpenAI just made its most powerful move since the original ChatGPT launch, and it's not just about a new model. The com...

OpenAI's GPT-4o safety report: Persuasion gets a 'medium' risk rating — Inblix summary
Product

OpenAI's GPT-4o safety report: Persuasion gets a 'medium' risk rating

OpenAI Blog · Jul 16, 2026 · 2 min read

OpenAI just dropped the System Card for GPT-4o, its new omni model that handles text, audio, and images natively. The h...

OpenAI drops free multimodal moderation model with 42% accuracy bump — Inblix summary
Product

OpenAI drops free multimodal moderation model with 42% accuracy bump

OpenAI Blog · Jul 15, 2026 · 2 min read

OpenAI just put a new sheriff in town, and it works across 40 languages. The company released omni-moderation-latest, a...

OpenAI lets you fine-tune GPT-4o with images, and yes, it's a big deal — Inblix summary
Product

OpenAI lets you fine-tune GPT-4o with images, and yes, it's a big deal

OpenAI Blog · Jul 15, 2026 · 2 min read

OpenAI finally plugged the glaring hole in its fine-tuning API. Starting today, you can feed images—not just text—into...

GPT-4o can now generate images — and it's a whole different beast — Inblix summary
Product

GPT-4o can now generate images — and it's a whole different beast

OpenAI Blog · Jul 14, 2026 · 2 min read

OpenAI just tore up the playbook on AI image generation. They've embedded a new imaging engine directly into the archit...

OpenAI bakes image gen into GPT-4o, and it can actually read — Inblix summary
Product

OpenAI bakes image gen into GPT-4o, and it can actually read

OpenAI Blog · Jul 14, 2026 · 2 min read

OpenAI just made image generation a native feature of GPT-4o, and the early details suggest it's a genuine step beyond...

NVIDIA's new AI safety model can now enforce custom rules, not just a one-size-fits-all policy — Inblix summary
Research

NVIDIA's new AI safety model can now enforce custom rules, not just a one-size-fits-all policy

Hugging Face Blog · Jun 4, 2026 · 2 min read

NVIDIA just dropped Nemotron 3.5 Content Safety, and the biggest upgrade isn't a better accuracy score — it's that the...

Google's Gemma 4 packs GPT-rivaling power into a 4B-parameter model — Inblix summary
Research

Google's Gemma 4 packs GPT-rivaling power into a 4B-parameter model

Hugging Face Blog · Apr 2, 2026 · 2 min read

A new open model family from Google is making a very specific kind of promise: frontier-level performance without the d...

Google's Gemma 3n Crams 5B-Parameter Brains Into Just 2GB of VRAM — Inblix summary
Research

Google's Gemma 3n Crams 5B-Parameter Brains Into Just 2GB of VRAM

Hugging Face Blog · Jun 26, 2025 · 2 min read

Google just dropped a pair of models that flip the script on what “small” AI means. The Gemma 3n series, now fully inte...

Meta drops Llama Guard 4: a 12B safety model that runs on a single GPU — Inblix summary
Research

Meta drops Llama Guard 4: a 12B safety model that runs on a single GPU

Hugging Face Blog · Apr 29, 2025 · 2 min read

Meta just released Llama Guard 4, and the headline feature isn't just about better detection — it's about practicality....

Hugging Face shrinks video AI to 256M params — and it actually works — Inblix summary
Research

Hugging Face shrinks video AI to 256M params — and it actually works

Hugging Face Blog · Feb 20, 2025 · 2 min read

Hugging Face just dropped a family of video-understanding models so small you can run them on a phone, and they're not...

The world’s smallest VLM is here at 256M parameters, beating 80B models from 17 months ago — Inblix summary
Research

The world’s smallest VLM is here at 256M parameters, beating 80B models from 17 months ago

Hugging Face Blog · Jan 23, 2025 · 2 min read

Hugging Face just shrank its vision language models down to sizes that would have sounded like a joke a year and a half...

Google drops PaliGemma 2 vision models up to 28B parameters with 896px resolution — Inblix summary
Research

Google drops PaliGemma 2 vision models up to 28B parameters with 896px resolution

Hugging Face Blog · Dec 5, 2024 · 2 min read

Google just significantly expanded its open vision-language model lineup with PaliGemma 2, a family of models that now...