PPaperPicks

Weijia Li

24 papers at tracked venues · 20 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Beyond Surface-Level Pattern Trap: LLM Agents for Faster and Smarter Cross-Architecture Code Migration
    ACL 2026 · Weijia Li
  2. MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
  3. RoZO: Geometry-Aware Zeroth-Order Fine-Tuning on Low-Rank Adapters for Black-Box Large Language Models
  4. BLINK-Twice: You see, but do you observe? A Reasoning Benchmark on Visual Perception
  5. Can Large Multimodal Models Understand Agricultural Scenes? Benchmarking with AgroMind
  6. Efficient Multi-modal Large Language Models via Progressive Consistency Distillation
  7. LEGION: Learning to Ground and Explain for Synthetic Image Detection
  8. LOKI: A Comprehensive Synthetic Data Detection Benchmark using Large Multimodal Models
  9. Leveraging BEV Paradigm for Ground-to-Aerial Image Synthesis
  10. Modeling Higher-order Human Beliefs Using the Justified Perspective Model
  11. Scene4U: Hierarchical Layered 3D Scene Reconstruction from Single Panoramic Image for Your Immerse Exploration
  12. Spot the Fake: Large Multimodal Model-Based Synthetic Image Detection with Artifact Explanation
  13. Stop Looking for "Important Tokens" in Multimodal Language Models: Duplication Matters More
  14. Token Pruning in Multimodal Large Language Models: Are We Solving the Right Problem?
  15. UrBench: A Comprehensive Benchmark for Evaluating Large Multimodal Models in Multi-View Urban Scenarios
  16. VHM: Versatile and Honest Vision Language Model for Remote Sensing Image Analysis
  17. Where am I? Cross-View Geo-localization with Natural Language Descriptions
  18. 3D Building Reconstruction from Monocular Remote Sensing Images with Multi-level Supervisions
    CVPR 2024 · Weijia Li
  19. AutoOS: Make Your OS More Powerful by Exploiting Large Language Models
  20. Building Bridges Across Spatial and Temporal Resolutions: Reference-Based Super-Resolution via Change Priors and Conditional Diffusion Model
  21. Cross-View Image Geo-Localization with Panorama-BEV Co-retrieval Network
  22. Parrot Captions Teach CLIP to Spot Text
  23. SG-BEV: Satellite-Guided BEV Fusion for Cross-View Semantic Segmentation
  24. VIGC: Visual Instruction Generation and Correction