PPaperPicks

Zhongyu Wei

41 papers at tracked venues · 30 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AutoJudger: An Agent-Driven Framework for Efficient Benchmarking of MLLMs
  2. AutoLink: Autonomous Schema Exploration and Expansion for Scalable Schema Linking in Text-to-SQL at Scale
  3. Autoregressive Semantic Visual Reconstruction Helps VLMs Understand Better
  4. CT-FineBench: A Diagnostic Fidelity Benchmark for Fine-Grained Evaluation of CT Report Generation
  5. Doc-V*: Coarse-to-Fine Interactive Visual Reasoning for Multi-Page Document VQA
  6. Each Rank Could be an Expert: Single-Ranked Mixture of Experts LoRA for Multi-task Learning
  7. Interleaved Latent Visual Reasoning with Selective Perceptual Modeling
  8. LifeSim: Long-Horizon User Life Simulator for Personalized Assistant Evaluation
  9. Listen, Pause, and Reason: Toward Perception-Grounded Hybrid Reasoning for Audio Understanding
  10. MAGNET: Towards Adaptive GUI Agents with Memory-Driven Knowledge Evolution
  11. Ready Jurist One: Benchmarking Language Agents for Legal Intelligence in Dynamic Environments
  12. Simple-VGC: Enhancing Visual Grounding in Multimodal Reasoning via Adaptive Tool Composition
  13. SpeechMedAssist: Efficiently and Effectively Adapting Speech Language Models for Medical Consultation
  14. Strong Reasoning Isn't Enough: Evaluating Evidence Elicitation in Interactive Diagnosis
  15. Activating Distributed Visual Region within LLMs for Efficient and Effective Vision-Language Training and Inference
  16. AgentSense: Benchmarking Social Intelligence of Language Agents through Interactive Scenarios
  17. EMGLLM: Data-to-Text Alignment for Electromyogram Diagnosis Generation with Medical Numerical Data Encoding
  18. EcoLANG: Efficient and Effective Agent Communication Language Induction for Social Simulation
  19. FinEval: A Chinese Financial Domain Knowledge Evaluation Benchmark for Large Language Models
  20. HAF-RM: A Hybrid Alignment Framework for Reward Model Training
  21. ITFormer: Bridging Time Series and Natural Language for Multi-Modal QA with Large-Scale Multitask Dataset
  22. Multi-Agent Simulator Drives Language Models for Legal Intensive Interaction
  23. Multi-agent KTO: Enhancing Strategic Interactions of Large Language Model in Language Game
  24. SFMSS: Service Flow aware Medical Scenario Simulation for Conversational Data Generation
  25. SocioBench: Modeling Human Behavior in Sociological Surveys with Large Language Models
  26. Synergistic Multi-Agent Framework with Trajectory Learning for Knowledge-Intensive Tasks
  27. UI-Hawk: Unleashing the Screen Stream Understanding for Mobile GUI Agents
  28. VoCoT: Unleashing Visually Grounded Multi-Step Reasoning in Large Multi-Modal Models
  29. Word Form Matters: LLMs' Semantic Reconstruction under Typoglycemia
  30. ALaRM: Align Language Models via Hierarchical Rewards Modeling
  31. Android in the Zoo: Chain-of-Action-Thought for GUI Agents
  32. Can LLMs Reason with Rules? Logic Scaffolding for Stress-Testing and Improving LLMs
  33. Debatrix: Multi-dimensional Debate Judge with Iterative Chronological Analysis Based on LLM
  34. EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models
  35. From LLMs to MLLMs: Exploring the Landscape of Multimodal Jailbreaking
  36. ReForm-Eval: Evaluating Large Vision Language Models via Unified Re-Formulation of Task-Oriented Benchmarks
  37. SoMeLVLM: A Large Vision Language Model for Social Media Processing
  38. Symbolic Working Memory Enhances Language Models for Complex Rule Application
  39. Unifying Local and Global Knowledge: Empowering Large Language Models as Political Experts with Knowledge Graphs
  40. Unveiling the Truth and Facilitating Change: Towards Agent-based Large-scale Social Movement Simulation
  41. Value at Adversarial Risk: A Graph Defense Strategy against Cost-Aware Attacks