PPaperPicks

Haoji Hu

10 papers at tracked venues · 8 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Q Cache: Visual Attention Is Valuable in Less than Half of Decode Layers for Multimodal Large Language Model
  2. Orientation Matters: Making 3D Generative Models Orientation-Aligned
  3. Recammaster: Camera-Controlled Generative Rendering From a Single Video
  4. ST3: Accelerating Multimodal Large Language Model by Spatial-Temporal Visual Token Trimming
  5. SynCamMaster: Synchronizing Multi-Camera Video Generation from Diverse Viewpoints
  6. UniEdit: A Unified Tuning-Free Framework for Video Motion and Appearance Editing
  7. UrbanCAD: Towards Highly Controllable and Photorealistic 3D Vehicles for Urban Scene Simulation
  8. FALIP: Visual Prompt as Foveal Attention Boosts CLIP Zero-Shot Performance
  9. Robustness-Guided Image Synthesis for Data-Free Quantization
  10. Unified Medical Image Pre-training in Language-Guided Common Semantic Space