PPaperPicks

Kai Zhang

Cornell University, Cornell Tech, Ithaca, IA, USA

24 papers at tracked venues · 21 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. 4D-LRM: Large Space-Time Reconstruction Model From and To Any View at Any Time
  2. Baking Gaussian Splatting Into Diffusion Denoiser for Fast and Scalable Single-Stage Image-to-3D Generation and Reconstruction
  3. Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors
  4. Gaussian Mixture Flow Matching Models
  5. Generating 3D-Consistent Videos from Unposed Internet Photos
  6. LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias
  7. LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers
  8. Long-LRM: Long-Sequence Large Reconstruction Model for Wide-Coverage Gaussian Splats
  9. MegaSynth: Scaling Up 3D Scene Reconstruction with Synthesized Data
  10. RandAR: Decoder-only Autoregressive Visual Generation in Random Orders
  11. Rayzer: a Self-Supervised Large View Synthesis Model
  12. RelitLRM: Generative Relightable Radiance for Large Reconstruction Models
  13. Turbo3D: Ultra-fast Text-to-3D Generation
  14. DATENeRF: Depth-Aware Text-Based Editing of NeRFs
  15. DMV3D: Denoising Multi-view Diffusion Using 3D Large Reconstruction Model
  16. GPT-4V(ision) is a Human-Aligned Evaluator for Text-to-3D Generation
  17. GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
    ECCV 2024 · Kai Zhang
  18. Instant3D: Fast Text-to-3D with Sparse-view Generation and Large Reconstruction Model
  19. LRM-Zero: Training Large Reconstruction Models with Synthesized Data
  20. LRM: Large Reconstruction Model for Single Image to 3D
  21. MegaScenes: Scene-Level View Synthesis at Scale
  22. Neural Directional Encoding for Efficient and Accurate View-Dependent Appearance Modeling
  23. Neural Gaffer: Relighting Any Object via Diffusion
  24. PF-LRM: Pose-Free Large Reconstruction Model for Joint Pose and Shape Prediction