P
PaperPicks
Conferences
Shih-Fu Chang
14 papers at tracked venues · 9 at CORE A* · active 2024–2025
DBLP profile ↗
ORCID search ↗
Venues
EMNLP
×3
ACL
×2
CVPR
×2
ICLR
×2
AAAI
×1
ACM MM
×1
ECCV
×1
NAACL
×1
NeurIPS
×1
Frequent coauthors
Hammad A. Ayyubi
DBLP profile ↗
ORCID search ↗
×3
Xudong Lin
DBLP profile ↗
ORCID search ↗
×2
Mingyang Zhou
DBLP profile ↗
ORCID search ↗
×1
Junzhang Liu
DBLP profile ↗
ORCID search ↗
×1
Kung-Hsiang Huang
DBLP profile ↗
ORCID search ↗
×1
Haoxuan You
DBLP profile ↗
ORCID search ↗
×1
Zhecan Wang
DBLP profile ↗
ORCID search ↗
×1
Jiawei Ma
DBLP profile ↗
ORCID search ↗
×1
Ali Zare
DBLP profile ↗
ORCID search ↗
×1
Yulei Niu
DBLP profile ↗
ORCID search ↗
×1
Brian Chen
DBLP profile ↗
ORCID search ↗
×1
Papers
M²-TabFact: Multi-Document Multi-Modal Fact Verification with Visual and Textual Representations of Tabular Data
ACL 2025
·
Mingyang Zhou
DBLP profile ↗
ORCID search ↗
PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction
NAACL 2025
·
Hammad A. Ayyubi
DBLP profile ↗
ORCID search ↗
Beyond Grounding: Extracting Fine-Grained Event Hierarchies across Modalities
AAAI 2024
·
Hammad A. Ayyubi
DBLP profile ↗
ORCID search ↗
Detecting Multimodal Situations with Insufficient Context and Abstaining from Baseless Predictions
ACM MM 2024
·
Junzhang Liu
DBLP profile ↗
ORCID search ↗
Do LVLMs Understand Charts? Analyzing and Correcting Factual Errors in Chart Captioning
ACL 2024
·
Kung-Hsiang Huang
DBLP profile ↗
ORCID search ↗
Ferret: Refer and Ground Anything Anywhere at Any Granularity
ICLR 2024
·
Haoxuan You
DBLP profile ↗
ORCID search ↗
JourneyBench: A Challenging One-Stop Vision-Language Understanding Benchmark of Generated Images
NeurIPS 2024
·
Zhecan Wang
DBLP profile ↗
ORCID search ↗
MoDE: CLIP Data Experts via Clustering
CVPR 2024
·
Jiawei Ma
DBLP profile ↗
ORCID search ↗
Personalized Video Comment Generation
EMNLP 2024
·
Xudong Lin
DBLP profile ↗
ORCID search ↗
RAP: Retrieval-Augmented Planner for Adaptive Procedure Planning in Instructional Videos
ECCV 2024
·
Ali Zare
DBLP profile ↗
ORCID search ↗
SCHEMA: State CHangEs MAtter for Procedure Planning in Instructional Videos
ICLR 2024
·
Yulei Niu
DBLP profile ↗
ORCID search ↗
Training-free Deep Concept Injection Enables Language Models for Video Question Answering
EMNLP 2024
·
Xudong Lin
DBLP profile ↗
ORCID search ↗
VIEWS: Entity-Aware News Video Captioning
EMNLP 2024
·
Hammad A. Ayyubi
DBLP profile ↗
ORCID search ↗
What, When, and Where? Self-Supervised Spatio- Temporal Grounding in Untrimmed Multi-Action Videos from Narrated Instructions
CVPR 2024
·
Brian Chen
DBLP profile ↗
ORCID search ↗