PPaperPicks

Linli Xu

University of Science and Technology of China (USTC), State Key Laboratory of Cognitive Intelligence, China

19 papers at tracked venues · 18 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Large Reasoning Embedding Models: Towards Next-Generation Dense Retrieval Paradigm
  2. Multimodal Table Understanding with Difficulty-aware Reinforcement Learning
  3. Addressing Representation Collapse in Vector Quantized Models with One Linear Layer
  4. BASIC: Boosting Visual Alignment with Intrinsic Refined Embeddings in Multimodal Large Language Models
  5. CROP: Integrating Topological and Spatial Structures via Cross-View Prefixes for Molecular LLMs
  6. Dynamic Prefix as Instructor for Incremental Named Entity Recognition: A Unified Seq2Seq Generation Framework
  7. ImageScope: Unifying Language-Guided Image Retrieval via Large Multimodal Model Collective Reasoning
  8. Input Domain Aware MoE: Decoupling Routing Decisions from Task Optimization in Mixture of Experts
  9. S²MILE: Semantic-and-Structure-Aware Music-Driven Lyric Generation
  10. Tracking the Copyright of Large Vision-Language Models through Parameter Learning Adversarial Images
  11. Video In-context Learning: Autoregressive Transformers are Zero-Shot Video Imitators
  12. Break the Visual Perception: Adversarial Attacks Targeting Encoded Visual Tokens of Large Vision-Language Models
  13. Bridging Gaps in Content and Knowledge for Multimodal Entity Linking
  14. Empowering Diffusion Models on the Embedding Space for Text Generation
  15. Generative Pre-trained Speech Language Model with Efficient Hierarchical Transformer
  16. HRVDA: High-Resolution Visual Document Assistant
  17. Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective
  18. Talk With Human-like Agents: Empathetic Dialogue Through Perceptible Acoustic Reception and Reaction
  19. Visual Hallucination Elevates Speech Recognition