P
PaperPicks
Conferences
Manling Li
23 papers at tracked venues · 20 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
ACL
×6
NeurIPS
×5
ICML
×4
EMNLP
×3
AAAI
×2
CVPR
×2
ICLR
×1
Frequent coauthors
Shiqi Chen
DBLP profile ↗
ORCID search ↗
×2
Keshigeyan Chandrasegaran
DBLP profile ↗
ORCID search ↗
×2
Zheyu Fan
DBLP profile ↗
ORCID search ↗
×1
Ziyi Wang
DBLP profile ↗
ORCID search ↗
×1
Prabhu Prakash Kagitha
DBLP profile ↗
ORCID search ↗
×1
Chi Wan
DBLP profile ↗
ORCID search ↗
×1
Zhenyu Pan
DBLP profile ↗
ORCID search ↗
×1
Rui Yang
DBLP profile ↗
ORCID search ↗
×1
Sina J. Semnani
DBLP profile ↗
ORCID search ↗
×1
Fan-Yun Sun
DBLP profile ↗
ORCID search ↗
×1
Jinhui Ye
DBLP profile ↗
ORCID search ↗
×1
Xuehang Guo
DBLP profile ↗
ORCID search ↗
×1
Papers
EMCompress: Video-LLMs with Endomorphic Multimodal Compression
ACL 2026
·
Zheyu Fan
DBLP profile ↗
ORCID search ↗
Trajectory2Task: Training Robust Tool-Calling Agents with Synthesized Yet Verifiable Data for Complex User Intents
ACL 2026
·
Ziyi Wang
DBLP profile ↗
ORCID search ↗
Unifying Inference-Time Planning Language Generation
ACL 2026
·
Prabhu Prakash Kagitha
DBLP profile ↗
ORCID search ↗
WorldAgen: Unified State-Action Prediction with Test-Time World Model Training
AAAI 2026
·
Chi Wan
DBLP profile ↗
ORCID search ↗
Bring Reason to Vision: Understanding Perception and Reasoning through Model Merging
ICML 2025
·
Shiqi Chen
DBLP profile ↗
ORCID search ↗
Chain-of-Action: Faithful and Multimodal Question Answering through Large Language Models
ICLR 2025
·
Zhenyu Pan
DBLP profile ↗
ORCID search ↗
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
ICML 2025
·
Rui Yang
DBLP profile ↗
ORCID search ↗
Exploring Diffusion Transformer Designs via Grafting
NeurIPS 2025
·
Keshigeyan Chandrasegaran
DBLP profile ↗
ORCID search ↗
From Large Language Models to Large Action Models: Reasoning and Planning with Physical World Knowledge
AAAI 2025
·
Manling Li
LEMONADE: A Large Multilingual Expert-Annotated Abstractive Event Dataset for the Real World
ACL 2025
·
Sina J. Semnani
DBLP profile ↗
ORCID search ↗
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
CVPR 2025
·
Fan-Yun Sun
DBLP profile ↗
ORCID search ↗
Re-thinking Temporal Search for Long-Form Video Understanding
CVPR 2025
·
Jinhui Ye
DBLP profile ↗
ORCID search ↗
SyncMind: Measuring Agent Out-of-Sync Recovery in Collaborative Software Engineering
ICML 2025
·
Xuehang Guo
DBLP profile ↗
ORCID search ↗
The Law of Knowledge Overshadowing: Towards Understanding, Predicting and Preventing LLM Hallucination
ACL 2025
·
Yuji Zhang
DBLP profile ↗
ORCID search ↗
VAGEN: Reinforcing World Model Reasoning for Multi-Turn VLM Agents
NeurIPS 2025
·
Kangrui Wang
DBLP profile ↗
ORCID search ↗
Why Is Spatial Reasoning Hard for VLMs? An Attention Mechanism Perspective on Focus Areas
ICML 2025
·
Shiqi Chen
DBLP profile ↗
ORCID search ↗
Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
NeurIPS 2024
·
Manling Li
HourVideo: 1-Hour Video-Language Understanding
NeurIPS 2024
·
Keshigeyan Chandrasegaran
DBLP profile ↗
ORCID search ↗
IKEA Manuals at Work: 4D Grounding of Assembly Instructions on Internet Videos
NeurIPS 2024
·
Yunong Liu
DBLP profile ↗
ORCID search ↗
Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing: EMNLP 2024 - System Demonstrations, Miami, Florida, USA, November 12-16, 2024
EMNLP 2024
·
Delia Irazú Hernández Farías
DBLP profile ↗
ORCID search ↗
Training-free Deep Concept Injection Enables Language Models for Video Question Answering
EMNLP 2024
·
Xudong Lin
DBLP profile ↗
ORCID search ↗
Why Does New Knowledge Create Messy Ripple Effects in LLMs?
EMNLP 2024
·
Jiaxin Qin
DBLP profile ↗
ORCID search ↗
Word Embeddings Are Steers for Language Models
ACL 2024
·
Chi Han
DBLP profile ↗
ORCID search ↗