PPaperPicks

Yu Cheng

Chinese University of Hong Kong, Department of Computer Science and Engineering, Shatin, Hong Kong

54 papers at tracked venues · 44 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. CoTEvol: Self-Evolving Chain-of-Thoughts for Data Synthesis in Mathematical Reasoning
  2. Each Rank Could be an Expert: Single-Ranked Mixture of Experts LoRA for Multi-task Learning
  3. Less Is More: Vision Representation Compression for Efficient Video Generation with Large Language Models
  4. Multi-LLM Collaborative Search for Complex Problem Solving
  5. Native Hybrid Attention for Efficient Sequence Modeling
  6. Nirvana: A Specialized Generalist Model With Task-Aware Memory Mechanism
  7. One Refiner to Unlock Them All: Inference-Time Reasoning Elicitation via Reinforcement Query Refinement
  8. Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models
  9. The Best of Both Worlds: Combining Parallel and Sequential Inference Scaling via Aggregation Fine-Tuning
  10. Bit-Flip Error Resilience in LLMs: A Comprehensive Analysis and Defense Framework
  11. CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
  12. Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
  13. Cooperative or Competitive? Understanding the Interaction between Attention Heads From A Game Theory Perspective
  14. Divide and Conquer: Grounding LLMs as Efficient Decision-Making Agents via Offline Hierarchical Reinforcement Learning
  15. Diving into Self-Evolving Training for Multimodal Reasoning
  16. Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs
  17. Dynamic Data Mixing Maximizes Instruction Tuning for Mixture-of-Experts
  18. Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think
  19. From Head to Tail: Towards Balanced Representation in Large Vision-Language Models through Adaptive Data Calibration
  20. ImageGen-CoT: Enhancing Text-to-Image in-context Learning with Chain-of-Thought Reasoning
  21. LangBridge: Interpreting Image as a Combination of Language Embeddings
  22. Learning to Reason under Off-Policy Guidance
  23. Liger: Linearizing Large Language Models to Gated Recurrent Structures
  24. Make LoRA Great Again: Boosting LoRA with Adaptive Singular Values and Mixture-of-Experts Optimization Alignment
  25. Occult: Optimizing Collaborative Communications across Experts for Accelerated Parallel MoE Training and Inference
  26. PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models
  27. SEE: Continual Fine-tuning with Sequential Ensemble of Experts
  28. Scaling Physical Reasoning with the PHYSICS Dataset
  29. Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback
  30. Text-to-Decision Agent: Offline Meta-Reinforcement Learning from Natural Language Supervision
  31. Towards Stabilized and Efficient Diffusion Transformers Through Long-Skip-Connections With Spectral Constraints
  32. Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation
  33. Unveiling Attractor Cycles in Large Language Models: A Dynamical Systems View of Successive Paraphrasing
  34. VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models
  35. Weak to Strong Generalization for Large Language Models with Multi-capabilities
  36. Confidence is not Timeless: Modeling Temporal Validity for Rule-based Temporal Knowledge Graph Forecasting
  37. ConflictBank: A Benchmark for Evaluating the Influence of Knowledge Conflicts in LLMs
  38. Enhancing Low-Resource Relation Representations through Multi-View Decoupling
  39. LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-Training
  40. Learning the Unlearned: Mitigating Feature Suppression in Contrastive Learning
  41. Living in the Moment: Can Large Language Models Grasp Co-Temporal Reasoning?
  42. MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
  43. Merge, Then Compress: Demystify Efficient SMoE with Hints from Its Routing Policy
  44. Mitigating Boundary Ambiguity and Inherent Bias for Text Classification in the Era of Large Language Models
  45. MoE-RBench: Towards Building Reliable Language Models with Sparse Mixture-of-Experts
  46. Not All Inputs Are Valid: Towards Open-Set Video Moment Retrieval using Language
  47. On Giant's Shoulders: Effortless Weak to Strong by Dynamic Logits Fusion
  48. On the Universal Truthfulness Hyperplane Inside LLMs
  49. Reinforcement Learning with Token-level Feedback for Controllable Text Generation
  50. Rethinking Weakly-Supervised Video Temporal Grounding From a Game Perspective
  51. SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information
  52. Sparse MoE with Language Guided Routing for Multilingual Machine Translation
  53. Twin-Merging: Dynamic Integration of Modular Expertise in Model Merging
  54. Unsupervised Domain Adaptative Temporal Sentence Localization with Mutual Information Maximization