PPaperPicks

Hao Tan

Adobe Research

25 papers at tracked venues · 23 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. 4D-LRM: Large Space-Time Reconstruction Model From and To Any View at Any Time
  2. Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors
  3. DiffTell: A High-Quality Dataset for Describing Image Manipulation Changes
  4. Gaussian Mixture Flow Matching Models
  5. Generating 3D-Consistent Videos from Unposed Internet Photos
  6. LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias
  7. LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers
  8. Long-LRM: Long-Sequence Large Reconstruction Model for Wide-Coverage Gaussian Splats
  9. MegaSynth: Scaling Up 3D Scene Reconstruction with Synthesized Data
  10. Progressive Autoregressive Video Diffusion Models
  11. RandAR: Decoder-only Autoregressive Visual Generation in Random Orders
  12. Rayzer: a Self-Supervised Large View Synthesis Model
  13. RelitLRM: Generative Relightable Radiance for Large Reconstruction Models
  14. Turbo3D: Ultra-fast Text-to-3D Generation
  15. VEGGIE: Instructional Editing and Reasoning Video Concepts with Grounded Generation
  16. Building Vision-Language Models on Solid Foundations with Masked Distillation
  17. Carve3D: Improving Multi-view Reconstruction Consistency for Diffusion Models with RL Finetuning
  18. DMV3D: Denoising Multi-view Diffusion Using 3D Large Reconstruction Model
  19. GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
  20. Identifying Speakers in Dialogue Transcripts: A Text-based Approach Using Pretrained Language Models
  21. Instant3D: Fast Text-to-3D with Sparse-view Generation and Large Reconstruction Model
  22. LRM-Zero: Training Large Reconstruction Models with Synthesized Data
  23. LRM: Large Reconstruction Model for Single Image to 3D
  24. PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction
  25. SOHES: Self-supervised Open-world Hierarchical Entity Segmentation