PPaperPicks

Guanbin Li

53 papers at tracked venues · 44 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Collaborative Multi-Agent Scripts Generation for Enhancing Imperfect-Information Reasoning in Murder Mystery Games
  2. Mobile-Agent-RAG: Driving Smart Multi-Agent Coordination with Contextual Knowledge Empowerment for Long-Horizon Mobile Automation
  3. 3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians
  4. Adadrive: Self-Adaptive Slow-Fast System for Language-Grounded Autonomous Driving
  5. Beyond the Destination: A Novel Benchmark for Exploration-Aware Embodied Question Answering
  6. Bridging Knowledge Gap Between Image Inpainting and Large-Area Visible Watermark Removal
  7. DAGSM: Disentangled Avatar Generation with GS-enhanced Mesh
  8. DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering
  9. DeepShield: Fortifying Deepfake Video Detection with Local and Global Forgery Analysis
  10. DreamFuse: Adaptive Image Fusion with Diffusion Transformer
  11. DreamLayer: Simultaneous Multi-Layer Generation via Diffusion Model
  12. Empowering Large Language Models with 3D Situation Awareness
  13. FakeRadar: Probing Forgery Outliers to Detect Unknown Deepfake Videos
  14. Free-Moref: Instantly Multiplexing Context Perception Capabilities of Video-Mllms Within Single Inference
  15. GUIDED: Granular Understanding via Identification, Detection, and Discrimination for Fine-Grained Open-Vocabulary Object Detection
  16. GeoSplatting: Towards Geometry Guided Gaussian Splatting for Physically-Based Inverse Rendering
  17. GlassWizard: Harvesting Diffusion Priors for Glass Surface Detection
  18. Hierarchically Controlled Deformable 3D Gaussians for Talking Head Synthesis
  19. LLM-driven Multimodal and Multi-Identity Listening Head Generation
  20. LaneDiffusion: Improving Centerline Graph Learning via Prior Injected BEV Feature Generation
  21. PDC-Net: Pattern Divide-and-Conquer Network for Pelvic Radiation Injury Segmentation
  22. PatchWiper: Leveraging Dynamic Patch-Wise Parameters for Real-World Visible Watermark Removal
  23. Pattern-Anchored Adaptive Prototype Learning for Gastroscopic Lesion Detection and Beyond
  24. Pseudo-Label Reconstruction for Partial Multi-Label Learning
  25. ReferSplat: Referring Segmentation in 3D Gaussian Splatting
  26. Rethinking Query-based Transformer for Continual Image Segmentation
  27. Screening, Rectifying, and Re-Screening: A Unified Framework for Tuning Vision-Language Models with Noisy Labels
  28. Sim-DETR: Unlock DETR for Temporal Sentence Grounding
  29. Towards Long-Horizon Vision-Language Navigation: Platform, Benchmark and Method
  30. VLDrive: Vision-Augmented Lightweight MLLMs for Efficient Language-Grounded Autonomous Driving
  31. VTON 360: High-Fidelity Virtual Try-On from Any Viewing Direction
  32. AlignSAM: Aligning Segment Anything Model to Open Context via Reinforcement Learning
  33. Cell Graph Transformer for Nuclei Classification
  34. Customize your NeRF: Adaptive Source Driven 3D Scene Editing via Local-Global Iterative Training
  35. Decoupled Pseudo-Labeling for Semi-Supervised Monocular 3D Object Detection
  36. FedDiv: Collaborative Noise Filtering for Federated Learning with Noisy Labels
  37. GraphVAE: Unveiling Dynamic Stock Relationships with Variational Autoencoder-based Factor Modeling
  38. Interactive 3D Object Detection with Prompts
  39. Learning Background Prompts to Discover Implicit Knowledge for Open Vocabulary Object Detection
  40. MarvelOVD: Marrying Object Recognition and Vision-Language Models for Robust Open-Vocabulary Object Detection
  41. Mask-Enhanced Segment Anything Model for Tumor Lesion Semantic Segmentation
  42. Multi-modal Denoising Diffusion Pre-training for Whole-Slide Image Classification
  43. NeRF-HuGS: Improved Neural Radiance Fields in Non-static Scenes Using Heuristics-Guided Segmentation
  44. OVER-NAV: Elevating Iterative Vision-and-Language Navigation with Open-Vocabulary Detection and StructurEd Representation
  45. Open-Vocabulary Segmentation with Semantic-Assisted Calibration
  46. Removing Interference and Recovering Content Imaginatively for Visible Watermark Removal
  47. TreeReward: Improve Diffusion Model via Tree-Structured Feedback Learning
  48. UniCell: Universal Cell Nucleus Classification via Prompt Learning
  49. UniFL: Improve Latent Diffusion Model via Unified Feedback Learning
  50. Variance-Insensitive and Target-Preserving Mask Refinement for Interactive Image Segmentation
  51. VersVideo: Leveraging Enhanced Temporal Diffusion Models for Versatile Video Generation
  52. WhodunitBench: Evaluating Large Multimodal Agents via Murder Mystery Games
  53. WildVidFit: Video Virtual Try-On in the Wild via Image-Based Controlled Diffusion Models