PPaperPicks

Chao Xu

24 papers at tracked venues · 17 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Deeply Seeking Boundary for Lunar Regolith Segmentation
  2. MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
  3. MoEC: A Memory-Routed Mixture-of-Experts Controller for Adaptive Minecraft Control
  4. An insect-scale multimodal amphibious piezoelectric robot
  5. AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
  6. BokehDiff: Neural Lens Blur with One-Step Diffusion
  7. Design and Implement of Large-scale Tail-sitter VTOL UAV
  8. Exploring the Frontiers of Animation Video Generation in the Sora Era: Method, Dataset and Benchmark
  9. InterAnimate: Taming Region-Aware Diffusion Model for Realistic Human Interaction Animation
  10. Large-Scale Trade-Off Curve Computation for Incentive Allocation with Cardinality and Matroid Constraints
  11. OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text
  12. OmniDocBench: Benchmarking Diverse PDF Document Parsing with Comprehensive Annotations
  13. PlaNet: Learning to Mitigate Atmospheric Turbulence in Planetary Images
  14. Polarization Guided Mask-Free Shadow Removal
  15. Synchronized Video-to-Audio Generation via Mel Quantization-Continuum Decomposition
  16. Transformer-Guided Genetic Programming for Symbolic Regression
    GECCO 2025 · Chao Xu
  17. Complementing Event Streams and RGB Frames for Hand Mesh Reconstruction
  18. MeshAvatar: Learning High-Quality Triangular Human Avatars from Multi-view Videos
  19. NB-GTR: Narrow-Band Guided Turbulence Removal
  20. Neural Underwater Scene Representation
  21. PSC: Extending Context Window of Large Language Models via Phase Shift Calibration
  22. Quality-Improved and Property-Preserved Polarimetric Imaging via Complementarily Fusing
  23. RAM-Avatar: Real-time Photo-Realistic Avatar from Monocular Videos with Full-body Control
  24. SPHINX-X: Scaling Data and Parameters for a Family of Multi-modal Large Language Models