P
PaperPicks
Conferences
Rongrong Ji
Xiamen University, Xiamen, China
97 papers at tracked venues · 85 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0001-9163-2932 ↗
Google Scholar ↗
Homepage ↗
Venues
NeurIPS
×18
ICML
×17
ACM MM
×16
CVPR
×12
ECCV
×8
ICLR
×8
ICCV
×6
ACL
×5
AAAI
×3
EMNLP
×2
IROS
×1
NAACL
×1
Frequent coauthors
Xudong Li
DBLP profile ↗
ORCID search ↗
×4
Zhihang Lin
DBLP profile ↗
ORCID search ↗
×3
Weizhong Huang
DBLP profile ↗
ORCID search ↗
×3
Yunshan Zhong
DBLP profile ↗
ORCID search ↗
×3
Mingrui Wu
DBLP profile ↗
ORCID search ↗
×3
Qiong Wu
DBLP profile ↗
ORCID search ↗
×2
Ziyin Zhou
DBLP profile ↗
ORCID search ↗
×2
Zhen Sun
DBLP profile ↗
ORCID search ↗
×2
Xudong Wang
DBLP profile ↗
ORCID search ↗
×2
You Huang
DBLP profile ↗
ORCID search ↗
×2
Chaoyou Fu
DBLP profile ↗
ORCID search ↗
×2
Wenhao Li
DBLP profile ↗
ORCID search ↗
×2
Papers
ALGOGEN: Tool-Generated Verifiable Traces for Reliable Algorithm Visualization
ACL 2026
·
Kunpeng Liao
DBLP profile ↗
ORCID search ↗
EMA: An Episodic Memory Agent for Efficient and Selective Memory
ACL 2026
·
Hongyi Lan
DBLP profile ↗
ORCID search ↗
Relaxing the Constraints: A Dual-Importance Projection Mechanism for Lifelong Model Editing
ACL 2026
·
Zhenghai Chen
DBLP profile ↗
ORCID search ↗
Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
NeurIPS 2025
·
Qiong Wu
DBLP profile ↗
ORCID search ↗
Aigi-Holmes: Towards Explainable and Generalizable AI-Generated Image Detection via Multimodal Large Language Models
ICCV 2025
·
Ziyin Zhou
DBLP profile ↗
ORCID search ↗
Automated Fine-Grained Mixture-of-Experts Quantization
ACL 2025
·
Zhanhao Xie
DBLP profile ↗
ORCID search ↗
Automated Manipulation of Magnetic Microswarms for Temporal Logic Cargo Delivery Tasks in Complex Environments
IROS 2025
·
Naifu Zhang
DBLP profile ↗
ORCID search ↗
BAME: Block-Aware Mask Evolution for Efficient N: M Sparse Training
ICML 2025
·
Chenyi Yang
DBLP profile ↗
ORCID search ↗
Benchmarking Abstract and Reasoning Abilities Through A Theoretical Perspective
ICML 2025
·
Qingchuan Ma
DBLP profile ↗
ORCID search ↗
Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference
AAAI 2025
·
Zhihang Lin
DBLP profile ↗
ORCID search ↗
CPPO: Accelerating the Training of Group Relative Policy Optimization-Based Reasoning Models
NeurIPS 2025
·
Zhihang Lin
DBLP profile ↗
ORCID search ↗
DAMamba: Vision State Space Model with Dynamic Adaptive Scan
NeurIPS 2025
·
Tanzhe Li
DBLP profile ↗
ORCID search ↗
DS-VLM: Diffusion Supervision Vision Language Model
ICML 2025
·
Zhen Sun
DBLP profile ↗
ORCID search ↗
Determining Layer-wise Sparsity for Large Language Models Through a Theoretical Perspective
ICML 2025
·
Weizhong Huang
DBLP profile ↗
ORCID search ↗
Discovering Important Experts for Mixture-of-Experts Models Pruning Through a Theoretical Perspective
NeurIPS 2025
·
Weizhong Huang
DBLP profile ↗
ORCID search ↗
Dynamic Low-Rank Sparse Adaptation for Large Language Models
ICLR 2025
·
Weizhong Huang
DBLP profile ↗
ORCID search ↗
EasyInv: Toward Fast and Better DDIM Inversion
ICML 2025
·
Ziyue Zhang
DBLP profile ↗
ORCID search ↗
Enhancing Language Model Hypernetworks with Restart: A Study on Optimization
NAACL 2025
·
Yihan Zhang
DBLP profile ↗
ORCID search ↗
Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
ICLR 2025
·
Gen Luo
DBLP profile ↗
ORCID search ↗
Few-Shot Image Quality Assessment via Adaptation of Vision-Language Models
ICCV 2025
·
Xudong Li
DBLP profile ↗
ORCID search ↗
FlashSloth : Lightning Multimodal Large Language Models via Embedded Visual Compression
CVPR 2025
·
Bo Tong
DBLP profile ↗
ORCID search ↗
FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-Identification
ICML 2025
·
Zhen Sun
DBLP profile ↗
ORCID search ↗
From Objects to Events: Unlocking Complex Visual Understanding in Object Detectors Via LLM-guided Symbolic Reasoning
ICCV 2025
·
Yuhui Zeng
DBLP profile ↗
ORCID search ↗
GPT-ReID: Learning Fine-grained Representation with GPT for Text-based Person Retrieval
ACM MM 2025
·
Xudong Wang
DBLP profile ↗
ORCID search ↗
GS-Bias: Global-Spatial Bias Learner for Single-Image Test-Time Adaptation of Vision-Language Models
ICML 2025
·
Zhaohong Huang
DBLP profile ↗
ORCID search ↗
Generate Aligned Anomaly: Region-Guided Few-Shot Anomaly Image-Mask Pair Synthesis for Industrial Inspection
ACM MM 2025
·
Yilin Lu
DBLP profile ↗
ORCID search ↗
HRSeg: High-Resolution Visual Perception and Enhancement for Reasoning Segmentation
ACM MM 2025
·
Weihuang Lin
DBLP profile ↗
ORCID search ↗
Inter2Former: Dynamic Hybrid Attention for Efficient High-Precision Interactive Segmentation
ICCV 2025
·
You Huang
DBLP profile ↗
ORCID search ↗
LTD-Bench: Evaluating Large Language Models by Letting Them Draw
NeurIPS 2025
·
Liuhao Lin
DBLP profile ↗
ORCID search ↗
Learning Interleaved Image-Text Comprehension in Vision-Language Large Models
ICLR 2025
·
Chenyu Zhou
DBLP profile ↗
ORCID search ↗
MIHBench: Benchmarking and Mitigating Multi-Image Hallucinations in Multimodal Large Language Models
ACM MM 2025
·
Jiale Li
DBLP profile ↗
ORCID search ↗
MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
NeurIPS 2025
·
Chaoyou Fu
DBLP profile ↗
ORCID search ↗
OracleFusion: Assisting the Decipherment of Oracle Bone Script with Structurally Constrained Semantic Typography
ICCV 2025
·
Caoshuo Li
DBLP profile ↗
ORCID search ↗
Routing Experts: Learning to Route Dynamic Experts in Existing Multi-modal Large Language Models
ICLR 2025
·
Qiong Wu
DBLP profile ↗
ORCID search ↗
SVFR: A Unified Framework for Generalized Video Face Restoration
CVPR 2025
·
Zhiyao Wang
DBLP profile ↗
ORCID search ↗
Semantic Alignment and Reinforcement for Data-Free Quantization of Vision Transformers
ICCV 2025
·
Yunshan Zhong
DBLP profile ↗
ORCID search ↗
Spotlight Attention: Towards Efficient LLM Generation via Non-linear Hashing-based KV Cache Retrieval
NeurIPS 2025
·
Wenhao Li
DBLP profile ↗
ORCID search ↗
Towards General Visual-Linguistic Face Forgery Detection
CVPR 2025
·
Ke Sun
DBLP profile ↗
ORCID search ↗
Training Long-Context LLMs Efficiently via Chunk-wise Optimization
ACL 2025
·
Wenhao Li
DBLP profile ↗
ORCID search ↗
VISA: Group-wise Visual Token Selection and Aggregation via Graph Summarization for Efficient MLLMs Inference
ACM MM 2025
·
Pengfei Jiang
DBLP profile ↗
ORCID search ↗
VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction
NeurIPS 2025
·
Chaoyou Fu
DBLP profile ↗
ORCID search ↗
VITA-Audio: Fast Interleaved Audio-Text Token Generation for Efficient Large Speech-Language Model
NeurIPS 2025
·
Zuwei Long
DBLP profile ↗
ORCID search ↗
VTON-HandFit: Virtual Try-on for Arbitrary Hand Pose Guided by Hand Priors Embedding
CVPR 2025
·
Yujie Liang
DBLP profile ↗
ORCID search ↗
Video-RAG: Visually-aligned Retrieval-Augmented Long Video Comprehension
NeurIPS 2025
·
Yongdong Luo
DBLP profile ↗
ORCID search ↗
What You Perceive Is What You Conceive: A Cognition-Inspired Framework for Open Vocabulary Image Segmentation
ACM MM 2025
·
Jianghang Lin
DBLP profile ↗
ORCID search ↗
Zooming from Context to Cue: Hierarchical Preference Optimization for Multi-Image MLLMs
NeurIPS 2025
·
Xudong Li
DBLP profile ↗
ORCID search ↗
polybasic Speculative Decoding Through a Theoretical Perspective
ICML 2025
·
Ruilin Wang
DBLP profile ↗
ORCID search ↗
γ-MoD: Exploring Mixture-of-Depth Adaptation for Multimodal Large Language Models
ICLR 2025
·
Yaxin Luo
DBLP profile ↗
ORCID search ↗
3D-GRES: Generalized 3D Referring Expression Segmentation
ACM MM 2024
·
Changli Wu
DBLP profile ↗
ORCID search ↗
AccDiffusion: An Accurate Method for Higher-Resolution Image Generation
ECCV 2024
·
Zhihang Lin
DBLP profile ↗
ORCID search ↗
Adaptive Feature Selection for No-Reference Image Quality Assessment by Mitigating Semantic Noise Sensitivity
ICML 2024
·
Xudong Li
DBLP profile ↗
ORCID search ↗
Adaptive Selection based Referring Image Segmentation
ACM MM 2024
·
Pengfei Yue
DBLP profile ↗
ORCID search ↗
Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation
ACM MM 2024
·
Jingjing Xie
DBLP profile ↗
ORCID search ↗
AffineQuant: Affine Transformation Quantization for Large Language Models
ICLR 2024
·
Yuexiao Ma
DBLP profile ↗
ORCID search ↗
Aligning and Prompting Everything All at Once for Universal Visual Perception
CVPR 2024
·
Yunhang Shen
DBLP profile ↗
ORCID search ↗
AnyTrans: Translate AnyText in the Image with Large Scale Models
EMNLP 2024
·
Zhipeng Qian
DBLP profile ↗
ORCID search ↗
Autoregressive Queries for Adaptive Tracking with Spatio-Temporal Transformers
CVPR 2024
·
Jinxia Xie
DBLP profile ↗
ORCID search ↗
CaM: Cache Merging for Memory-efficient LLMs Inference
ICML 2024
·
Yuxin Zhang
DBLP profile ↗
ORCID search ↗
CamoTeacher: Dual-Rotation Consistency Learning for Semi-supervised Camouflaged Object Detection
ECCV 2024
·
Xunfa Lai
DBLP profile ↗
ORCID search ↗
Cantor: Inspiring Multimodal Chain-of-Thought of MLLM
ACM MM 2024
·
Timin Gao
DBLP profile ↗
ORCID search ↗
Code Membership Inference for Detecting Unauthorized Data Use in Code Pre-trained Language Models
EMNLP 2024
·
Sheng Zhang
DBLP profile ↗
ORCID search ↗
ControlMLLM: Training-Free Visual Prompt Learning for Multimodal Large Language Models
NeurIPS 2024
·
Mingrui Wu
DBLP profile ↗
ORCID search ↗
Cross-Modality Perturbation Synergy Attack for Person Re-identification
NeurIPS 2024
·
Yunpeng Gong
DBLP profile ↗
ORCID search ↗
Deep Instruction Tuning for Segment Anything Model
ACM MM 2024
·
Xiaorui Huang
DBLP profile ↗
ORCID search ↗
DiffAgent: Fast and Accurate Text-to-Image API Selection with Large Language Model
CVPR 2024
·
Lirui Zhao
DBLP profile ↗
ORCID search ↗
DiffuMatting: Synthesizing Arbitrary Objects with Matting-Level Annotation
ECCV 2024
·
Xiaobin Hu
DBLP profile ↗
ORCID search ↗
DiffusionFake: Enhancing Generalization in Deepfake Detection via Guided Stable Diffusion
NeurIPS 2024
·
Ke Sun
DBLP profile ↗
ORCID search ↗
Director3D: Real-world Camera Trajectory and 3D Scene Generation from Text
NeurIPS 2024
·
Xinyang Li
DBLP profile ↗
ORCID search ↗
Dynamic Sparse No Training: Training-Free Fine-tuning for Sparse LLMs
ICLR 2024
·
Yuxin Zhang
DBLP profile ↗
ORCID search ↗
ERQ: Error Reduction for Post-Training Quantization of Vision Transformers
ICML 2024
·
Yunshan Zhong
DBLP profile ↗
ORCID search ↗
Enhancing Tampered Text Detection Through Frequency Feature Fusion and Decomposition
ECCV 2024
·
Zhongxi Chen
DBLP profile ↗
ORCID search ↗
Evaluating and Analyzing Relationship Hallucinations in Large Vision-Language Models
ICML 2024
·
Mingrui Wu
DBLP profile ↗
ORCID search ↗
Exploring Phrase-Level Grounding with Text-to-Image Diffusion Model
ECCV 2024
·
Danni Yang
DBLP profile ↗
ORCID search ↗
Exploring Target Representations for Masked Autoencoders
ICLR 2024
·
Xingbin Liu
DBLP profile ↗
ORCID search ↗
Fast Text-to-3D-Aware Face Generation and Manipulation via Direct Cross-modal Mapping and Geometric Regularization
ICML 2024
·
Jinlu Zhang
DBLP profile ↗
ORCID search ↗
FocSAM: Delving Deeply into Focused Objects in Segmenting Anything
CVPR 2024
·
You Huang
DBLP profile ↗
ORCID search ↗
GOI: Find 3D Gaussians of Interest with an Optimizable Open-vocabulary Semantic-space Hyperplane
ACM MM 2024
·
Yansong Qu
DBLP profile ↗
ORCID search ↗
GraCo: Granularity-Controllable Interactive Segmentation
CVPR 2024
·
Yian Zhao
DBLP profile ↗
ORCID search ↗
I2EBench: A Comprehensive Benchmark for Instruction-based Image Editing
NeurIPS 2024
·
Yiwei Ma
DBLP profile ↗
ORCID search ↗
Integrating Global Context Contrast and Local Sensitivity for Blind Image Quality Assessment
ICML 2024
·
Xudong Li
DBLP profile ↗
ORCID search ↗
Learning Image Demoiréing from Unpaired Real Data
AAAI 2024
·
Yunshan Zhong
DBLP profile ↗
ORCID search ↗
Multi-branch Collaborative Learning Network for 3D Visual Grounding
ECCV 2024
·
Zhipeng Qian
DBLP profile ↗
ORCID search ↗
Multimodal Inplace Prompt Tuning for Open-set Object Detection
ACM MM 2024
·
Guilin Li
DBLP profile ↗
ORCID search ↗
Outlier-aware Slicing for Post-Training Quantization in Vision Transformer
ICML 2024
·
Yuexiao Ma
DBLP profile ↗
ORCID search ↗
PortraitBooth: A Versatile Portrait Model for Fast Identity-Preserved Personalization
CVPR 2024
·
Xu Peng
DBLP profile ↗
ORCID search ↗
Prompting to Adapt Foundational Segmentation Models
ACM MM 2024
·
Jie Hu
DBLP profile ↗
ORCID search ↗
QueryMatch: A Query-based Contrastive Learning Framework for Weakly Supervised Visual Grounding
ACM MM 2024
·
Shengxin Chen
DBLP profile ↗
ORCID search ↗
RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation
NeurIPS 2024
·
Changli Wu
DBLP profile ↗
ORCID search ↗
RLE: A Unified Perspective of Data Augmentation for Cross-Spectral Re-Identification
NeurIPS 2024
·
Lei Tan
DBLP profile ↗
ORCID search ↗
Rotated Multi-Scale Interaction Network for Referring Remote Sensing Image Segmentation
CVPR 2024
·
Sihan Liu
DBLP profile ↗
ORCID search ↗
SAM as the Guide: Mastering Pseudo-Label Refinement in Semi-Supervised Referring Expression Segmentation
ICML 2024
·
Danni Yang
DBLP profile ↗
ORCID search ↗
StealthDiffusion: Towards Evading Diffusion Forensic Detection through Diffusion Model
ACM MM 2024
·
Ziyin Zhou
DBLP profile ↗
ORCID search ↗
TF-FAS: Twofold-Element Fine-Grained Semantic Guidance for Generalizable Face Anti-spoofing
ECCV 2024
·
Xudong Wang
DBLP profile ↗
ORCID search ↗
Textual Grounding for Open-Vocabulary Visual Information Extraction in Layout-Diversified Documents
ECCV 2024
·
Mengjun Cheng
DBLP profile ↗
ORCID search ↗
Toward Open-Set Human Object Interaction Detection
AAAI 2024
·
Mingrui Wu
DBLP profile ↗
ORCID search ↗
UniPTS: A Unified Framework for Proficient Post-Training Sparsity
CVPR 2024
·
Jingjing Xie
DBLP profile ↗
ORCID search ↗
X-Oscar: A Progressive Framework for High-quality Text-guided 3D Animatable Avatar Generation
ICML 2024
·
Yiwei Ma
DBLP profile ↗
ORCID search ↗