IMG

Total Publications: 553
PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction
Naman Mishra,Shankar Gangisetty,Jawahar C V
International Conference on Robotics and Automation, ICRA, 2026
Core Rank : A* Google Rank : 122
V-Align: Visual Forced Alignment via Phoneme to Video Optimal Path Traversal
Souvik Ghosh,Jawahar C V,Vinay Namboodiri
Annual Conference of the International Speech Communication Association, INTERSPEECH, 2026
Core Rank : A Google Rank : 111
LipAdapter: Text-to-Video Alignment is All You Need for Lip-to-Speech
Souvik Ghosh,Jawahar C V,Vinay Namboodiri
Annual Conference of the International Speech Communication Association, INTERSPEECH, 2026
Core Rank : A Google Rank : 111
MOTOR: A Multimodal Dataset for Two-Wheeler Rider Behavior Understanding
Varun A Paturkar,Shankar Gangisetty,Jawahar C V
International Conference on Robotics and Automation, ICRA, 2026
Core Rank : A* Google Rank : 122
Unifying Scientific Communication: Fine-Grained Correspondence Across Scientific Media
Megha Mariam K M,Vineeth N Balasubramanian,Jawahar C V
Computer Vision and Pattern Recognition - Findings, CVPR-F, 2026
Learning Beyond Labels: Self-Supervised Handwritten Text Recognition
Shree Mitra,Ajoy Mondal,Jawahar C V
Winter Conference on Applications of Computer Vision, WACV, 2026
Core Rank : A Google Rank : 109
PhyEduVideo: A Benchmark for Evaluating Text-to-Video Models for Physics Education
Megha Mariam K M,Aditya Arun,Zakaria Laskar,Jawahar C V
Winter Conference on Applications of Computer Vision, WACV, 2026
Core Rank : A Google Rank : 109
Towards Scalable Sign Production: Leveraging Co-Articulated Gloss Dictionary for Fluid Sign Synthesis
Agrawal Aparna Nitin,Seshadri Mazumder,Jawahar C V,Vinay P Namboodiri
Indian Conference on Computer Vision, Graphics and Image Processing, ICVGIP, 2025
Core Rank : - Google Rank : -
From pixels to tables: reconstructing complex tables from document images
Sachin Raja,Ajoy Mondal,Jawahar C V
International Journal on Document Analysis and Recognition, IJDAR, 2025
Core Rank : - Google Rank : 23
EviFiVQA: A Benchmark for Evidence-Grounded Multi-hop Reasoning in Financial VQA
Sachin Raja,Ajoy Mondal,Jawahar C V
International Conference on Document Analysis and Recognition, ICDAR, 2025
Core Rank : A Google Rank : 50