PPaperPicks

Wenwei Zhang

35 papers at tracked venues · 27 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. SciExplore: Evaluating Autonomous Agents from Scientific Navigation to Information Integration
  2. Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives
  3. Are Your LLMs Capable of Stable Reasoning?
  4. Calib3D: Calibrating Model Preferences for Reliable 3D Scene Understanding
  5. CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward
  6. F-LMM: Grounding Frozen Large Multimodal Models
  7. Harmonizing Visual Representations for Unified Multimodal Understanding and Generation
  8. InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model
  9. LLaVA-3D: A Simple Yet Effective Pathway to Empowering LMMs with 3D Capabilities
  10. Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs
  11. MindSearch: Mimicking Human Minds Elicits Deep AI Searcher
  12. Pre-Trained Policy Discriminators are General Reward Models
  13. Rethinking Verification for LLM Code Generation: From Generation to Testing
  14. SLAM Assisted 3D Tracking System for Laparoscopic Surgery
  15. Semi-off-Policy Reinforcement Learning for Vision-Language Slow-Thinking Reasoning
  16. Training Language Models to Critique With Multi-agent Feedback
  17. 4D Contrastive Superflows are Dense 3D Representation Learners
  18. ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models
  19. ANAH: Analytical Annotation of Hallucinations in Large Language Models
  20. Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models
  21. AlchemistCoder: Harmonizing and Eliciting Code Capability by Hindsight Tuning on Multi-source Data
  22. CLIM: Contrastive Language-Image Mosaic for Region Representation
  23. CLIPSelf: Vision Transformer Distills Itself for Open-Vocabulary Dense Prediction
  24. Can AI Assistants Know What They Don't Know?
  25. Code Needs Comments: Enhancing Code LLMs with Comment Augmentation
  26. CriticEval: Evaluating Large-scale Language Model as Critic
  27. EmbodiedScan: A Holistic Multi-Modal 3D Perception Suite Towards Embodied AI
  28. Fake Alignment: Are LLMs Really Aligned Well?
  29. GPT4RoI: Instruction Tuning Large Language Model on Region-of-Interest
  30. InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD
  31. MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark
  32. OMG-Seg: Is One Model Good Enough for all Segmentation?
  33. ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities
  34. T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
  35. Unified Human-Scene Interaction via Prompted Chain-of-Contacts