PPaperPicks

Rui Hu

14 papers at tracked venues · 11 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. LENS: Learning to Segment Anything with Unified Reinforced Reasoning
  2. Q Cache: Visual Attention Is Valuable in Less than Half of Decode Layers for Multimodal Large Language Model
  3. Adaptive Markup Language Generation for Contextually-Grounded Visual Document Understanding
  4. Beyond Squared Error: Exploring Loss Design for Enhanced Training of Generative Flow Networks
    ICLR 2025 · Rui Hu
  5. BlueLM-V-3B: Algorithm and System Co-Design for Multimodal Large Language Models on Mobile Devices
  6. Finite-Time Analysis of Discrete-Time Stochastic Interpolants
  7. GroundingSuite: Measuring Complex Multi-Granular Pixel Grounding
    ICCV 2025 · Rui Hu
  8. MEraser: An Effective Fingerprint Erasure Approach for Large Language Models
  9. ST3: Accelerating Multimodal Large Language Model by Spatial-Temporal Visual Token Trimming
  10. Transcript-Prompted Whisper with Dictionary-Enhanced Decoding for Japanese Speech Annotation
  11. UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents
  12. FALIP: Visual Prompt as Foveal Attention Boosts CLIP Zero-Shot Performance
  13. FashionR2R: Texture-preserving Rendered-to-Real Image Translation with Diffusion Models
    NeurIPS 2024 · Rui Hu
  14. Volumetric Conditional Score-Based Residual Diffusion Model for PET/MR Denoising