PPaperPicks

Bowen Zhou

Tsinghua University, Department of Electronic Engineering, Beijing, China

44 papers at tracked venues · 37 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. GenPRM: Scaling Test-Time Compute of Process Reward Models via Generative Reasoning
  2. I2E: From Image Pixels to Actionable Interactive Environments for Text-Guided Image Editing
  3. MARS²: Scaling Multi-Agent Tree Search via Reinforcement Learning for Code Generation
  4. MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
  5. Nirvana: A Specialized Generalist Model With Task-Aware Memory Mechanism
  6. SDAR-VL: Stable and Efficient Block-wise Diffusion for Vision-Language Understanding
  7. SDAR: A Synergistic Diffusion-AutoRegression Paradigm for Scalable Sequence Generation
  8. V-GameGym: Visual Game Generation for Code Large Language Models
  9. AdsQA: Towards Advertisement Video Understanding
  10. Advancing LLM Reasoning Generalists with Preference Trees
  11. DePass: Unified Feature Attributing by Simple Decomposed Forward Pass
  12. Dolphin: Moving Towards Closed-loop Auto-research through Thinking, Practice, and Feedback
  13. Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization
  14. Free Process Rewards without Process Labels
  15. Fusing Highly Specialized Language Models for Comprehensive Expertise
  16. How to Synthesize Text Data without Model Collapse?
  17. Intuitive Fine-Tuning: Towards Simplifying Alignment into a Single Process
  18. Less is More: Efficient Model Merging with Binary Task Switch
  19. MARS2 2025 Challenge on Multimodal Reasoning: Datasets, Methods, Results, Discussion, and Outlook
  20. Many Heads Are Better Than One: Improved Scientific Idea Generation by A LLM-Based Multi-Agent System
  21. MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding
  22. Memory Decoder: A Pretrained, Plug-and-Play Memory for Large Language Models
  23. OpenPRM: Building Open-domain Process-based Reward Models with Preference Trees
  24. Retrieval-Augmented Visual Question Answering via Built-in Autoregressive Search Engines
  25. ReviewRL: Towards Automated Scientific Review with RL
  26. TTRL: Test-Time Reinforcement Learning
  27. AdapEdit: Spatio-Temporal Guided Adaptive Editing Algorithm for Text-Based Continuity-Sensitive Image Editing
  28. CoGenesis: A Framework Collaborating Large and Small Language Models for Secure Context-Aware Instruction Following
  29. Empowering Private Tutoring by Chaining Large Language Models
  30. Exploring Adversarial Robustness of Deep State Space Models
  31. Generative Multi-Modal Knowledge Retrieval with Large Language Models
  32. Interactive Continual Learning: Fast and Slow Thinking
  33. LAKE-RED: Camouflaged Images Generation by Latent Background Knowledge Retrieval-Augmented Diffusion
  34. LMD: Faster Image Reconstruction with Latent Masking Diffusion
  35. MSI-Agent: Incorporating Multi-Scale Insight into Embodied Agents for Superior Planning and Decision-Making
  36. Neural Residual Diffusion Models for Deep Scalable Vision Generation
  37. On Large Language Models' Hallucination with Regard to Known Facts
  38. On the token distance modeling ability of higher RoPE attention dimension
  39. PaD: Program-aided Distillation Can Teach Small Models Reasoning Better than Chain-of-thought Fine-tuning
  40. SMR: State Memory Replay for Long Sequence Modeling
  41. Safe-SD: Safe and Traceable Stable Diffusion with Text Prompt Trigger for Invisible Generative Watermarking
  42. Scalable Efficient Training of Large Language Models with Low-dimensional Projected Attention
  43. Trust in Internal or External Knowledge? Generative Multi-Modal Entity Linking with Knowledge Retriever
  44. UltraMedical: Building Specialized Generalists in Biomedicine