PPaperPicks

Tianyi Zhou

University of Maryland, College Park, MD, USA

67 papers at tracked venues · 48 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?
  2. Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction
  3. FedMerge: Federated Model Merging for Personalization
  4. Optimizing Length Compression in Large Reasoning Models
  5. Schoenfeld's Anatomy of Mathematical Reasoning by Language Models
  6. Skill Discovery for Software Scripting Automation via Offline Simulations with LLMs
  7. ATLAS: Agent Tuning via Learning Critical Steps
  8. BenTo: Benchmark Reduction with In-Context Transferability
  9. ColorBench: Can VLMs See and Understand the Colorful World? A Comprehensive Benchmark for Color Perception, Reasoning, and Robustness
  10. Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
  11. DataGen: Unified Synthetic Dataset Generation via Large Language Models
  12. Diffusion Curriculum: Synthetic-to-Real Data Curriculum via Image-Guided Diffusion
  13. Don't Think Longer, Think Wisely: Optimizing Thinking Dynamics for Large Reasoning Models
  14. Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement
  15. Florence-VL: Enhancing Vision-Language Models with Generative Vision Encoder and Depth-Breadth Fusion
  16. From Lists to Emojis: How Format Bias Affects Model Alignment
  17. GUI Agents: A Survey
  18. Generative Models for Synthetic Data: Transforming Data Mining in the GenAI Era
  19. Is Your Multimodal Language Model Oversensitive to Safe Queries?
  20. Many-Objective Multi-Solution Transport
  21. Mosaic-IT: Cost-Free Compositional Data Synthesis for Instruction Tuning
  22. Multiple LLM Agents Debate for Equitable Cultural Alignment
  23. OmnixR: Evaluating Omni-modality Language Models on Reasoning across Modalities
  24. Personalized Federated Collaborative Filtering: A Variational AutoEncoder Approach
  25. Preference Controllable Reinforcement Learning with Advanced Multi-Objective Optimization
  26. Quantifying and Modeling Driving Styles in Trajectory Forecasting
  27. R2-T2: Re-Routing in Test-Time for Multimodal Mixture-of-Experts
  28. RuleR: Improving LLM Controllability by Rule-based Data Recycling
  29. The Crystal Ball Hypothesis in diffusion models: Anticipating object positions from initial noise
  30. Tilted Sharpness-Aware Minimization
  31. Understanding the Thinking Process of Reasoning Models: A Perspective from Schoenfeld's Episode Theory
  32. VideoHallu: Evaluating and Mitigating Multi-modal Hallucinations on Synthetic Video Understanding
  33. WALL-E: World Alignment by NeuroSymbolic Learning improves World Model-based LLM Agents
  34. Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency
  35. Your Mixture-of-Experts LLM Is Secretly an Embedding Model for Free
  36. 1+1\textgreater2: Can Large Language Models Serve as Cross-Lingual Knowledge Aggregators?
  37. Adaptive Regularization of Representation Rank as an Implicit Constraint of Bellman Equation
  38. Algorithm and Hardness for Dynamic Attention Maintenance in Large Language Models
  39. AlpaGasus: Training a Better Alpaca with Fewer Data
  40. An End-to-End Submodular Framework for Data-Efficient In-Context Learning
  41. AutoHallusion: Automatic Generation of Hallucination Benchmarks for Vision-Language Models
  42. Automatic Curriculum for Unsupervised Reinforcement Learning
  43. Can LLMs Speak For Diverse People? Tuning LLMs via Debate to Generate Controllable Controversial Statements
  44. Do Text-Free Diffusion Models Learn Discriminative Visual Representations?
  45. Do great minds think alike? Investigating Human-AI Complementarity in Question Answering with CAIMIRA
  46. DrAttack: Prompt Decomposition and Reconstruction Makes Powerful LLMs Jailbreakers
  47. Easy2Hard-Bench: Standardized Difficulty Labels for Profiling LLM Performance and Generalization
  48. Embodied Multi-Modal Agent trained by an LLM from a Parallel TextWorld
  49. Federated Recommendation with Additive Personalization
  50. From Quantity to Quality: Boosting LLM Performance with Self-Guided Data Selection for Instruction Tuning
  51. GPFedRec: Graph-Guided Personalization for Federated Recommendation
  52. Hallusionbench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models
  53. InstructZero: Efficient Instruction Optimization for Black-Box Large Language Models
  54. Meta-Task Prompting Elicits Embeddings from Large Language Models
  55. Multi-Objective Linguistic Control of Large Language Models
  56. ODIN: Disentangled Reward Mitigates Hacking in RLHF
  57. One Prompt is not Enough: Automated Construction of a Mixture-of-Expert Prompts
  58. Position: TrustLLM: Trustworthiness in Large Language Models
  59. Retrieval-Augmented Retrieval: Large Language Models are Strong Zero-Shot Retriever
  60. Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
  61. SpecHub: Provable Acceleration to Multi-Draft Speculative Decoding
  62. Superfiltering: Weak-to-Strong Data Filtering for Fast Instruction-Tuning
  63. Task Adaptation from Skills: Information Geometry, Disentanglement, and New Objectives for Unsupervised Reinforcement Learning
  64. Task-Driven Domain-Agnostic Learning with Information Bottleneck for Autonomous Steering
  65. The Closeness of In-Context Learning and Weight Shifting for Softmax Regression
  66. Understanding the Impact of Negative Prompts: When and How Do They Take Effect?
  67. When Federated Recommendation Meets Cold-Start Problem: Separating Item Attributes and User Interactions