PPaperPicks

Heng Wang

8 papers at tracked venues · 5 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Gotta Hear Them All: Towards Sound Source Aware Audio Generation
  2. ChoreoMuse: Robust Music-to-Dance Video Generation with Style Transfer and Beat-Adherent Motion
  3. Dance any Beat: Blending Beats with Visuals in Dance Video Generation
  4. Multimodal Causal Reasoning Benchmark: Challenging Multimodal Large Language Models to Discern Causal Links Across Modalities
  5. Advancements in 3D Lane Detection Using LiDAR Point Clouds: From Data Collection to Model Development
  6. Enhancing Advanced Visual Reasoning Ability of Large Language Models
  7. LaneCMKT: Boosting Monocular 3D Lane Detection with Cross-Modal Knowledge Transfer
  8. V2A-Mapper: A Lightweight Solution for Vision-to-Audio Generation by Connecting Foundation Models
    AAAI 2024 · Heng Wang