P
PaperPicks
Conferences
Yuliang Liu
23 papers at tracked venues · 20 at CORE A* · active 2024–2025
DBLP profile ↗
ORCID search ↗
Venues
ICCV
×5
CVPR
×4
NeurIPS
×4
ACL
×3
EMNLP
×2
ICML
×2
AAAI
×1
ECCV
×1
ICLR
×1
Frequent coauthors
Zhang Li
DBLP profile ↗
ORCID search ↗
×2
Mingxin Huang
DBLP profile ↗
ORCID search ↗
×2
Wenwen Yu
DBLP profile ↗
ORCID search ↗
×1
Zhiyuan Hu
DBLP profile ↗
ORCID search ↗
×1
Liang Yin
DBLP profile ↗
ORCID search ↗
×1
Jingqun Tang
DBLP profile ↗
ORCID search ↗
×1
Yang Liu
DBLP profile ↗
ORCID search ↗
×1
Ling Fu
DBLP profile ↗
ORCID search ↗
×1
Dongliang Luo
DBLP profile ↗
ORCID search ↗
×1
Linger Deng
DBLP profile ↗
ORCID search ↗
×1
Enming Zhang
DBLP profile ↗
ORCID search ↗
×1
Hanshen Zhu
DBLP profile ↗
ORCID search ↗
×1
Papers
AdaptiveStep: Automatically Dividing Reasoning Step through Model Confidence
ICML 2025
·
Yuliang Liu
DocThinker: Explainable Multimodal Large Language Models with Rule-Based Reinforcement Learning for Document Understanding
ICCV 2025
·
Wenwen Yu
DBLP profile ↗
ORCID search ↗
LIRA: Inferring Segmentation in Large Multi-Modal Models with Local Interleaved Region Assistance
ICCV 2025
·
Zhang Li
DBLP profile ↗
ORCID search ↗
LongRecipe: Recipe for Efficient Long Context Generalization in Large Language Models
ACL 2025
·
Zhiyuan Hu
DBLP profile ↗
ORCID search ↗
MSTAR: Box-free Multi-query Scene Text Retrieval with Attention Recycling
NeurIPS 2025
·
Liang Yin
DBLP profile ↗
ORCID search ↗
MTVQA: Benchmarking Multilingual Text-Centric Visual Question Answering
ACL 2025
·
Jingqun Tang
DBLP profile ↗
ORCID search ↗
Mini-Monkey: Alleviating the Semantic Sawtooth Effect for Lightweight MLLMs via Complementary Image Pyramid
ICLR 2025
·
Mingxin Huang
DBLP profile ↗
ORCID search ↗
Multi-Scenario Overlapping Text Segmentation with Depth Awareness
ICCV 2025
·
Yang Liu
DBLP profile ↗
ORCID search ↗
OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning
NeurIPS 2025
·
Ling Fu
DBLP profile ↗
ORCID search ↗
SemiETS: Integrating Spatial and Content Consistencies for Semi-Supervised End-to-end Text Spotting
CVPR 2025
·
Dongliang Luo
DBLP profile ↗
ORCID search ↗
Theorem-Validated Reverse Chain-of-Thought Problem Generation for Geometric Reasoning
EMNLP 2025
·
Linger Deng
DBLP profile ↗
ORCID search ↗
Towards Comprehensive Lecture Slides Understanding: Large-Scale Dataset and Effective Method
ICCV 2025
·
Enming Zhang
DBLP profile ↗
ORCID search ↗
Training-Free Geometric Image Editing on Diffusion Models
ICCV 2025
·
Hanshen Zhu
DBLP profile ↗
ORCID search ↗
WildDoc: How Far Are We from Achieving Comprehensive and Robust Document Understanding in the Wild?
EMNLP 2025
·
An-Lan Wang
DBLP profile ↗
ORCID search ↗
AP-Adapter: Improving Generalization of Automatic Prompts on Unseen Text-to-Image Diffusion Models
NeurIPS 2024
·
Yuchen Fu
DBLP profile ↗
ORCID search ↗
Bridging the Gap Between End-to-End and Two-Step Text Spotting
CVPR 2024
·
Mingxin Huang
DBLP profile ↗
ORCID search ↗
Deciphering Oracle Bone Language with Diffusion Models
ACL 2024
·
Haisu Guan
DBLP profile ↗
ORCID search ↗
MoE Jetpack: From Dense Checkpoints to Adaptive Mixture of Experts for Vision Tasks
NeurIPS 2024
·
Xingkui Zhu
DBLP profile ↗
ORCID search ↗
Monkey: Image Resolution and Text Label are Important Things for Large Multi-Modal Models
CVPR 2024
·
Zhang Li
DBLP profile ↗
ORCID search ↗
OMNIPARSER: A Unified Framework for Text Spotting, Key Information Extraction and Table Recognition
CVPR 2024
·
Jianqiang Wan
DBLP profile ↗
ORCID search ↗
ViTEraser: Harnessing the Power of Vision Transformers for Scene Text Removal with SegMIM Pretraining
AAAI 2024
·
Dezhi Peng
DBLP profile ↗
ORCID search ↗
Video-LaVIT: Unified Video-Language Pre-training with Decoupled Visual-Motional Tokenization
ICML 2024
·
Yang Jin
DBLP profile ↗
ORCID search ↗
Well Begun is Half Done: The Importance of Initialization in Dataset Distillation
ECCV 2024
·
Yiran Guan
DBLP profile ↗
ORCID search ↗