PPaperPicks

Jiajun Deng

23 papers at tracked venues · 13 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. 3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer
    CVPR 2025 · Jiajun Deng
  2. Exploring SSL Discrete Speech Features for Zipformer-based Contextual ASR
  3. GraspCoT: Integrating Physical Property Reasoning for 6-DoF Grasping Under Flexible Language Instructions
  4. Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots
  5. MOPSA: Mixture of Prompt-Experts Based Speaker Adaptation for Elderly Speech Recognition
  6. On-the-fly Routing for Zero-shot MoE Speaker Adaptation of Speech Foundation Models for Dysarthric Speech Recognition
  7. PGOV3D: Open-Vocabulary 3D Semantic Segmentation with Partial-to-Global Curriculum
  8. RaCFormer: Towards High-Quality 3D Object Detection via Query-based Radar-Camera Fusion
  9. S3R-GS: Streamlining the Pipeline for Large-Scale Street Scene Reconstruction
  10. Self-Classification Enhancement and Correction for Weakly Supervised Object Detection
  11. SpatialSplat: Efficient Semantic 3D from Sparse Unposed Images
  12. SynTag: Enhancing the Geometric Robustness of Inversion-Based Generative Image Watermarking
  13. VLMPlanner: Integrating Visual Language Models with Motion Planning
  14. Agent3D-Zero: An Agent for Zero-Shot 3D Understanding
  15. Cycle-Consistency Learning for Captioning and Grounding
  16. End-to-End Rate-Distortion Optimized 3D Gaussian Representation
  17. FARFusion V2: A Geometry-based Radar-Camera Fusion Method on the Ground for Roadside Far-Range 3D Object Detection
  18. Hierarchical Temporal Context Learning for Camera-Based Semantic Scene Completion
  19. Joint Speaker Features Learning for Audio-visual Multichannel Speech Separation and Recognition
  20. One-pass Multiple Conformer and Foundation Speech Systems Compression and Quantization Using An All-in-one Neural Model
  21. RayFormer: Improving Query-Based Multi-Camera 3D Object Detection via Ray-Centric Strategies
  22. Revisiting Open-Set Panoptic Segmentation
  23. Towards Effective and Efficient Non-autoregressive Decoding Using Block-based Attention Mask