Advancing Vision, Language, and Multimodal Intelligence for Bharat and Beyond.
Recognition, detection, segmentation and scene understanding.
Multimodal reasoning and vision-language learning.
Large Language Models and multilingual AI.
Image generation and restoration.
Large-scale multimodal learning.
Research addressing Indian societal challenges.