P
PaperPicks
Conferences
Shentong Mo
14 papers at tracked venues · 10 at CORE A* · active 2024–2025
DBLP profile ↗
ORCID search ↗
Venues
ECCV
×4
CVPR
×3
NeurIPS
×3
AAAI
×2
ICLR
×1
ICML
×1
Frequent coauthors
Lin Zhang
DBLP profile ↗
ORCID search ↗
×1
Weiguo Pian
DBLP profile ↗
ORCID search ↗
×1
Jiantao Wu
DBLP profile ↗
ORCID search ↗
×1
Tanvir Mahmud
DBLP profile ↗
ORCID search ↗
×1
Papers
Foley-Flow: Coordinated Video-to-Audio Generation with Masked Audio-Visual Alignment and Dynamic Conditional Flows
CVPR 2025
·
Shentong Mo
GMAIL: Generative Modality Alignment for generated Image Learning
ICML 2025
·
Shentong Mo
Scaling Diffusion Mamba with Bidirectional SSMs for Efficient 3D Shape Generation
AAAI 2025
·
Shentong Mo
The Dynamic Duo of Collaborative Masking and Target for Advanced Masked Autoencoder Learning
AAAI 2025
·
Shentong Mo
pMoE: Prompting Diverse Experts Together Wins More in Visual Adaptation
ICLR 2025
·
Shentong Mo
Aligning Audio-Visual Joint Representations with an Agentic Workflow
NeurIPS 2024
·
Shentong Mo
Audio-Synchronized Visual Animation
ECCV 2024
·
Lin Zhang
DBLP profile ↗
ORCID search ↗
Audio-Visual Generalized Zero-Shot Learning the Easy Way
ECCV 2024
·
Shentong Mo
Connecting Joint-Embedding Predictive Architecture with Contrastive Self-supervised Learning
NeurIPS 2024
·
Shentong Mo
Continual Audio-Visual Sound Separation
NeurIPS 2024
·
Weiguo Pian
DBLP profile ↗
ORCID search ↗
DailyMAE: Towards Pretraining Masked Autoencoders in One Day
ECCV 2024
·
Jiantao Wu
DBLP profile ↗
ORCID search ↗
Fast Training of Diffusion Transformer with Extreme Masking for 3D Point Clouds Generation
ECCV 2024
·
Shentong Mo
MA-AVT: Modality Alignment for Parameter-Efficient Audio-Visual Transformers
CVPR 2024
·
Tanvir Mahmud
DBLP profile ↗
ORCID search ↗
Unveiling the Power of Audio-Visual Early Fusion Transformers with Dense Interactions Through Masked Modeling
CVPR 2024
·
Shentong Mo