PPaperPicks

C. V. Jawahar

IIIT Hyderabad, Centre for Visual Information Technology (CVIT), India

24 papers at tracked venues · 4 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Can VLMs Understand Handwritten Mathematical Documents?
  2. Distilling What and Why: Enhancing Driver Intention Prediction with MLLMs
  3. Learning Beyond Labels: Self-Supervised Handwritten Text Recognition
  4. MIST: Multilingual Incidental Dataset for Scene Text Detection
  5. PhyEduVideo: A Benchmark for Evaluating Text-to-Video Models for Physics Education
  6. UniTabBank: A Large Scale Multi-Lingual, Multi-Layout, Multi-Type, Multi-Format Dataset for Table Detection
  7. A Dataset for Semantic Segmentation in the Presence of Unknowns
  8. AI-Generated Lecture Slides for Improving Slide Element Detection and Retrieval
  9. Adapting Vision-Language Models for Hindi OCR
  10. Attend to What I Say: Highlighting Relevant Content on Slides
  11. DashGaze: Driver Gaze Through Dashcam
  12. EviFiVQA: A Benchmark for Evidence-Grounded Multi-hop Reasoning in Financial VQA
  13. ICDAR 2025 Handwritten Notes Understanding Challenge
  14. Multilingual Query-by-Example KWS for Indian Languages using Transliteration
  15. Pedestrian Intention and Trajectory Prediction in Unstructured Traffic Using IDD-PeD
  16. Towards Safer and Understandable Driver Intention Prediction
  17. Treading Towards Privacy-Preserving Table Structure Recognition
  18. UniLayDet: Simple Multi-dataset Document Layout Analysis
  19. Can Reasons Help Improve Pedestrian Intent Estimation? A Cross-Modal Approach
  20. Early Anticipation of Driving Maneuvers
  21. Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives
  22. IDD-X: A Multi-View Dataset for Ego-relative Important Object Localization and Explanation in Dense and Unstructured Traffic
  23. Understanding the Generalization of Pretrained Diffusion Models on Out-of-Distribution Data
    AAAI 2024 ·
    Sai Niranjan Ramachandran
  24. Visual Place Recognition in Unstructured Driving Environments