AI Glossary
What is Multimodal?
A multimodal model accepts or produces multiple kinds of data: reading screenshots and PDFs, hearing audio, generating images or video. Frontier chat models are now multimodal on input by default; generation across modalities is typically handled by specialized models (image and video generators).
News momentum
Mentions in recent AI news titles and summaries, refreshed from the ingestion stream.
98
-9 vs prior 7d
- Predicting Multiple Clinical Outcomes Related to Functional Recovery and Social Isolation Among Older Adults After Lower-Limb Fracture or Hip Replacement
arxiv-ai · 1d
- Act with Intent: Distilling Behavior Intent for Vision-Language-Action Models
arxiv-ai · 1d
- MetaCaster: Meta-Harness-Optimized Agent for End-to-End Few-Shot Learning of Lightweight Time Series Forecasters
arxiv-ai · 1d
- Towards Comprehensive Basketball Understanding
arxiv-ai · 1d