PPaperPicks

Yuhao Wang

35 papers at tracked venues · 23 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. CADTrack: Learning Contextual Aggregation with Deformable Alignment for Robust RGBT Tracking
  2. Cross-Modal Coreference Alignment: Enabling Reliable Information Transfer in Omni-LLMs
  3. Do Large Language Models Reason About Uncertainty Like Humans? A Benchmark on Hurricane Forecast Visualization Comprehension
  4. Miner: Mining Intrinsic Mastery for Data-Efficient RL in Large Reasoning Models
  5. PEOD: A Pixel-Aligned Event-RGB Benchmark for Object Detection Under Challenging Conditions
  6. SAS-VPReID: A Scale-Adaptive Framework with Shape Priors for Video-Based Person Re-IDentification at Extreme Far Distances
  7. STMI: Segmentation-Guided Token Modulation with Cross-Modal Hypergraph Interaction for Multi-Modal Object Re-Identification
  8. Signal: Selective Interaction and Global-local Alignment for Multi-Modal Object Re-Identification
  9. Uncovering Pretraining Code in LLMs: A Syntax-Aware Attribution Approach
  10. VReID-XFD: Video-Based Person Re-Identification at Extreme Far Distance Challenge Results
  11. W2S-AlignTree: Weak-to-Strong Inference-Time Alignment for Large Language Models via Monte Carlo Tree Search
  12. When Seeing Is not Enough: Revealing the Limits of Active Reasoning in MLLMs
  13. Breaking Down Power Barriers in On-Device Streaming ASR: Insights and Solutions
  14. CLIMB-ReID: A Hybrid CLIP-Mamba Framework for Person Re-Identification
  15. DeMo: Decoupled Feature-Based Mixture of Experts for Multi-Modal Object Re-Identification
    AAAI 2025 · Yuhao Wang
  16. Dynamic Analysis and Adaptive Discriminator for Fake News Detection
  17. EfficientVLA: Training-Free Acceleration and Compression for Vision-Language-Action Models
  18. EvolveBench: A Comprehensive Benchmark for Assessing Temporal Awareness in LLMs on Evolving Knowledge
  19. Gaussian Mean Testing under Truncation
  20. IDEA: Inverted Text with Cooperative Deformable Aggregation for Multi-modal Object Re-Identification
    CVPR 2025 · Yuhao Wang
  21. Infrared and Visible Image Fusion with Language-Driven Loss in CLIP Embedding Space
    ACM MM 2025 · Yuhao Wang
  22. Learning High-dimensional Gaussians from Censored Data
  23. MambaPro: Multi-Modal Object Re-identification with Mamba Aggregation and Synergistic Prompt
    AAAI 2025 · Yuhao Wang
  24. SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant
  25. Sigma: Siamese Mamba Network for Multi-Modal Semantic Segmentation
  26. Subtyping Breast Lesions via Generative Augmentation Based Long-Tailed Recognition in Ultrasound
  27. TRRG: Towards Truthful Radiology Report Generation With Cross-Modal Disease Clue Enhanced Large Language Models
    MICCAI 2025 · Yuhao Wang
  28. Toward Universal Laws of Outlier Propagation
  29. UniConvNet: Expanding Effective Receptive Field While Maintaining Asymptotically Gaussian Distribution for ConvNets of Any Scale
    ICCV 2025 · Yuhao Wang
  30. Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement
  31. VocalNet: Speech LLMs with Multi-Token Prediction for Faster and High-Quality Generation
    EMNLP 2025 · Yuhao Wang
  32. MM-SAP: A Comprehensive Benchmark for Assessing Self-Awareness of Multimodal Large Language Models in Perception
    ACL 2024 · Yuhao Wang
  33. Magic Tokens: Select Diverse Tokens for Multi-modal Object Re-Identification
  34. Optimal estimation of Gaussian (poly)trees
    AISTATS 2024 · Yuhao Wang
  35. TOP-ReID: Multi-Spectral Object Re-identification with Token Permutation
    AAAI 2024 · Yuhao Wang