How I Study AI - Learn AI Papers & Lectures the Easy Way

FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs

Intermediate

Qian Chen, Jinlan Fu et al.Jan 20arXiv

FutureOmni is the first benchmark that tests if multimodal AI models can predict what happens next from both sound and video, not just explain what already happened.

#multimodal LLM#audio-visual reasoning#future forecasting

Papers1

FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs