PPaperPicks

Mingkui Tan

27 papers at tracked venues · 22 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Action-and-object Aware Alignment for Partially Relevant Video Retrieval
  2. FAM: Fine-Grained Alignment Matters in Multimodal Embedding Learning with Large Vision-Language Models
  3. Latent-Condensed Transformer for Efficient Long Context Modeling
  4. NaVLA$2$: A Vision-Language-Audio-Action Model for Multimodal Instruction Navigation
  5. SUGAR: Learning Skeleton Representation with Visual-Motion Knowledge for Action Recognition
  6. Continual Knowledge Adaptation for Reinforcement Learning
  7. Core Context Aware Transformers for Long Context Language Modeling
  8. Curse of High Dimensionality Issue in Transformer for Long Context Modeling
  9. Dataset, Baseline and Evaluation Design for GAVE Challenge
  10. Deep Electromagnetic Structure Design Under Limited Evaluation Budgets
  11. Efficient Dynamic Ensembling for Multiple LLM Experts
  12. Enhancing User-Oriented Proactivity in Open-Domain Dialogues with Critic Guidance
  13. Frequency-Aware Autoregressive Modeling for Efficient High-Resolution Image Synthesis
  14. Generating Long-form Story Using Dynamic Hierarchical Outlining with Memory-Enhancement
  15. LSceneLLM: Enhancing Large 3D Scene Understanding Using Adaptive Visual Preferences
  16. Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
  17. Open-World Drone Active Tracking with Goal-Centered Rewards
  18. Physics-Driven Spatiotemporal Modeling for AI-Generated Video Detection
  19. Test-Time Learning for Large Language Models
  20. Test-Time Model Adaptation for Quantized Neural Networks
  21. Test-time Adapted Reinforcement Learning with Action Entropy Regularization
  22. Understanding Emotional Body Expressions via Large Language Models
  23. Cross-Device Collaborative Test-Time Adaptation
  24. Detecting Machine-Generated Texts by Multi-Population Aware Optimization for Maximum Mean Discrepancy
  25. G-NeRF: Geometry-enhanced Novel View Synthesis from Single-View Images
  26. HiLo: Detailed and Robust 3D Clothed Human Reconstruction with High-and Low-Frequency Information of Parametric Models
  27. Towards Robust and Efficient Cloud-Edge Elastic Model Adaptation via Selective Entropy Distillation