PPaperPicks

Fan Ma

19 papers at tracked venues · 17 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. Autonomous LLM-Enhanced Adversarial Attack for Text-to-Motion
  2. BrainGuard: Privacy-Preserving Multisubject Image Reconstructions from Brain Activities
  3. DreamDPO: Aligning Text-to-3D Generation with Human Preferences via Direct Preference Optimization
  4. From Trial to Triumph: Advancing Long Video Understanding via Visual Context Sample Scaling and Self-Reward Alignment
  5. Image Regeneration: Evaluating Text-to-Image Model via Generating Identical Image with Multimodal Large Language Models
  6. Imagine and Seek: Improving Composed Image Retrieval with an Imagined Proxy
  7. InfiniDreamer: Arbitrarily Long Human Motion Generation Via Segment Score Distillation
  8. Long-horizon Visual Instruction Generation with Logic and Attribute Self-reflection
  9. Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion
  10. CapHuman: Capture Your Moments in Parallel Universes
  11. Clustering for Protein Representation Learning
  12. HeadStudio: Text to Animatable Head Avatars with 3D Gaussian Splatting
  13. Knowledge-Enhanced Dual-Stream Zero-Shot Composed Image Retrieval
  14. LSK3DNet: Towards Effective and Efficient 3D Perception with Large Sparse Kernels
  15. MIGC: Multi-Instance Generation Controller for Text-to-Image Synthesis
  16. Psychometry: An Omnifit Model for Image Reconstruction from Human Brain Activity
  17. Stitching Segments and Sentences towards Generalization in Video-Text Pre-training
    AAAI 2024 · Fan Ma
  18. Vista-llama: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens
    CVPR 2024 · Fan Ma
  19. VividDreamer: Invariant Score Distillation for Hyper-Realistic Text-to-3D Generation