Uploaded January 2026 | Updated September 2026, 2 hours ago
Day 26/42: Modality & Multimodal Models
Yesterday, we compressed intelligence.
Today, we expand senses.
Text is a modality.
So are images, audio, and video.
Multimodal models process several at once.
That’s why modern systems can:
see images,
read text,
and listen.
This changes what AI can understand, not just generate.
Missed Day 25? Start there.
Tomorrow, we talk thinking speed: reasoning models.
I’m Louis-François, PhD dropout, now CTO & co-founder at Towards AI. Follow me for tomorrow’s no-BS AI roundup 🚀
#MultimodalAI #LLM #AIExplained #GenerativeAI #LearnAI #WhatsAI #short
Day 26/42: Modality & Multimodal Models
Yesterday, we compressed intelligence.
Today, we expand senses.
Text is a modality.
So are images, audio, and video.
Multimodal models process several at once.
That’s why modern systems can:
see images,
read text,
and listen.
This changes what AI can understand, not just generate.
Missed Day 25? Start there.
Tomorrow, we talk thinking speed: reasoning models.
I’m Louis-François, PhD dropout, now CTO & co-founder at Towards AI. Follow me for tomorrow’s no-BS AI roundup 🚀
#MultimodalAI #LLM #AIExplained #GenerativeAI #LearnAI #WhatsAI #short










