Multimodal AI
Multimodal AI breakthroughs and products
Articles
Knowledge-guided Disentanglement with Atomic Actions for Actionβ9
Proposes KDA, using LLMs to decompose actions into atomic actions with spatial-temporal semantics for enhanced, disentangled action recognition, achieving
How an Overlooked Geothermal Plant Got a Second Chanceβ9
An abandoned geothermal plant in New Mexico is revived by AI-driven startup Zanskar, now generating 25 MW of clean baseload power for thousands of homes.
The Download: A Chip Talent Battle, and Deflating AI Hypeβ8
Samsung engineers defect to SK Hynix for record bonuses amid an AI chip talent war, plus a reality check on deflating AI hype in 2026.
The AI Hype Index: Unsexy AIβ8
MIT Technology Reviewβs AI Hype Index ranks the most overhyped and underrated AI trends, celebrating practical βunsexyβ AI over flashy promises.
DisasterTD: Disaster Toponym Disambiguation Using Multimodalβ7
DisasterTD integrates multimodal LLMs and cross-view geolocalization to disambiguate disaster toponyms, achieving 71.62% accuracy within 1 km on Hurricane Harve...
Gradient-Based Latent Decomposition Reveals Mechanisms ofβ7
A new method reveals why fine-grained malignancy cues degrade in weakly supervised mammography, with 95.6% of latent features lying orthogonal to supervisory si...
The Download: OpenAIβs Predictable Hack, and an AI Stock Sell-Offβ10
OpenAIβs predictable hack and a major AI stock sell-off highlight growing concerns over security and market volatility.
Samsung Chip Workers Defect to Rival SK Hynix as AI Talent Warβ9
As AI chip competition heats up, Samsung chip workers defect to SK Hynix for better pay, HBM-dominated growth, and career advancement, intensifying the global t...
MegaSlide-DiT: Memory-Centric Adaptation and Deformable Localβ7
MegaSlide-DiT adapts 105B-parameter video diffusion on single H200 via memory-centric host-GPU streaming and 3D deformable local attention, cutting
OpenAI called the Hugging Face attack unprecedented. But weβveβ9
OpenAI called the Hugging Face attack unprecedented, but a decade-old AI safety experiment shows similar vulnerabilities have long existed.
