PPaperPicks

Qibin Hou

29 papers at tracked venues · 28 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Easy Samples Are All You Need: Self-Evolving LLMs via Data-Efficient Reinforcement Learning
  2. SM3Det: A Unified Model for Multi-Modal Remote Sensing Object Detection
  3. Strip R-CNN: Large Strip Convolution for Remote Sensing Object Detection
  4. AR-1-to-3: Single Image to Consistent 3D Object via Next-View Prediction
  5. DFormerv2: Geometry Self-Attention for RGBD Semantic Segmentation
  6. Docopilot: Improving Multimodal Models for Document-Level Understanding
  7. K-LoRA: Unlocking Training-Free Fusion of Any Subject and Style LoRAs
  8. KAC: Kolmogorov-Arnold Classifier for Continual Learning
  9. Multi-Task Dense Predictions via Unleashing the Power of Diffusion
  10. OmniSegmentor: A Flexible Multi-Modal Learning Framework for Semantic Segmentation
  11. Re-Aligning Language to Visual Objects with an Agentic Workflow
  12. Revisiting Efficient Semantic Segmentation: Learning Offsets for Better Spatial and Class Feature Alignment
  13. SE-GUI: Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning
  14. TAR3D: Creating High-Quality 3D Assets Via Next-Part Prediction
  15. TempSamp-R1: Effective Temporal Sampling with Reinforcement Fine-Tuning for Video LLMs
  16. Unbiased Region-Language Alignment for Open-Vocabulary Dense Prediction
  17. Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation
  18. CorrMatch: Label Propagation via Correlation Matching for Semi-Supervised Semantic Segmentation
  19. CrossKD: Cross-Head Knowledge Distillation for Object Detection
  20. DFormer: Rethinking RGBD Representation Learning for Semantic Segmentation
  21. Get What You Want, Not What You Don't: Image Content Suppression for Text-to-Image Diffusion Models
  22. Multi-Task Dense Prediction via Mixture of Low-Rank Experts
  23. OPUS: Occupancy Prediction Using a Sparse Set
  24. Polyper: Boundary Sensitive Polyp Segmentation
  25. SARDet-100K: Towards Open-Source Benchmark and ToolKit for Large-Scale SAR Object Detection
  26. StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation
  27. TeMO: Towards Text-Driven 3D Stylization for Multi-Object Meshes
  28. Towards Stable 3D Object Detection
  29. Traffic Scene Parsing Through the TSP6K Dataset