PPaperPicks

Hai Zhao

Shanghai Jiao Tong University, Department of Computer Science and Engineering, China

63 papers at tracked venues · 44 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. BoYaEval: Evaluating Multimodal Large Language Models on Understanding Ancient Chinese Musical Scores
  2. Faster MoE LLM Inference for Extremely Large Models
  3. GoT-R1: Internalizing Graph-of-Thought via Structural Reinforcement for High-Density Reasoning
  4. PAR: Training-Free Positional Perturbation and Attention Recycling for Faithful OCR
  5. ParaCook: On Time-Efficient Planning for Multi-Agent Systems
  6. RACER: Retrieval-Augmented Contextual Rapid Speculative Decoding
  7. Remember Me, Refine Me: A Dynamic Procedural Memory Framework for Experience-Driven Agent Evolution
  8. Scaling LLM Speculative Decoding: Non-Autoregressive Forecasting in Large-Batch Scenarios
  9. TrigReason: Trigger-Based Collaboration between Small and Large Reasoning Models
  10. Can Large Language Models Be Good Language Teachers?
  11. Caution for the Environment: Multimodal LLM Agents are Susceptible to Environmental Distractions
  12. CoViPAL: Layer-wise Contextualized Visual Token Pruning for Large Vision-Language Models
  13. DAC: A Dynamic Attention-aware Approach for Task-Agnostic Prompt Compression
  14. Dialogue-RAG: Enhancing Retrieval for LLMs via Node-Linking Utterance Rewriting
  15. Evolving Chinese Spelling Correction with Corrector-Verifier Collaboration
  16. Faster In-Context Learning for LLMs via N-Gram Trie Speculative Decoding
  17. From Parameters to Performance: A Data-Driven Study on LLM Structure and Development
  18. Game Development as Human-LLM Interaction
  19. IAM: Efficient Inference through Attention Mapping between Different-scale LLMs
  20. KV-Latent: Dimensional-level KV Cache Reduction with Frequency-aware Rotary Positional Embedding
  21. LESA: Learnable LLM Layer Scaling-Up
  22. MEGen: Generative Backdoor into Large Language Models via Model Editing
  23. Open-Theatre: An Open-Source Toolkit for LLM-based Interactive Drama
  24. PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization
  25. SCANS: Mitigating the Exaggerated Safety for LLMs via Safety-Conscious Activation Steering
  26. Segment First or Comprehend First? Explore the Limit of Unsupervised Word Segmentation with Large Language Models
  27. SmallKV: Small Model Assisted Compensation of KV Cache Compression for Efficient LLM Inference
  28. ToM: Leveraging Tree-oriented MapReduce for Long-Context Reasoning in Large Language Models
  29. Towards Enhanced Immersion and Agency for LLM-based Interactive Drama
  30. Unfolding the Headline: Iterative Self-Questioning for News Retrieval and Timeline Summarization
  31. What Limits Bidirectional Model's Generative Capabilities? A Uni-Bi-Directional Mixture-of-Expert Method For Bidirectional Fine-tuning
  32. Wide-Horizon Thinking and Simulation-Based Evaluation for Real-World LLM Planning with Multifaceted Constraints
  33. X-TURING: Towards an Enhanced and Efficient Turing Test for Long-Term Dialogue Agents
  34. XQuant: Achieving Ultra-Low Bit KV Cache Quantization with Cross-Layer Compression
  35. A Coin Has Two Sides: A Novel Detector-Corrector Framework for Chinese Spelling Correction
  36. A Novel Energy Based Model Mechanism for Multi-Modal Aspect-Based Sentiment Analysis
  37. Are LLMs Aware that Some Questions are not Open-ended?
  38. CMMLU: Measuring massive multitask language understanding in Chinese
  39. Chinese Spelling Correction as Rephrasing Language Model
  40. Chinese Spelling Corrector Is Just a Language Learner
  41. CoCo-Agent: A Comprehensive Cognitive MLLM Agent for Smartphone GUI Automation
  42. Dissecting Human and LLM Preferences
  43. Fact-Driven Logical Reasoning for Machine Reading Comprehension
  44. From Role-Play to Drama-Interaction: An LLM Solution
  45. GKT: A Novel Guidance-Based Knowledge Transfer Framework For Efficient Cloud-edge Collaboration LLM Deployment
  46. GLaPE: Gold Label-agnostic Prompt Evaluation for Large Language Models
  47. Generative Judge for Evaluating Alignment
  48. GoT: Effective Graph-of-Thought Reasoning in Language Models
  49. Head-wise Shareable Attention for Large Language Models
  50. Hypergraph based Understanding for Document Semantic Entity Recognition
  51. Instruction-Driven Game Engine: A Poker Case Study
  52. LaCo: Large Language Model Pruning via Layer Collapse
  53. Multi-modal Auto-regressive Modeling via Visual Tokens
  54. On the Robustness of Editing Large Language Models
  55. PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
  56. Reference Trustable Decoding: A Training-Free Augmentation Paradigm for Large Language Models
  57. Selective Prefix Tuning for Pre-trained Language Models
  58. Self-Prompting Large Language Models for Zero-Shot Open-Domain QA
  59. SirLLM: Streaming Infinite Retentive LLM
  60. Sparse is Enough in Fine-tuning Pre-trained Large Language Models
  61. The Music Maestro or The Musically Challenged, A Massive Music Evaluation Benchmark for Large Language Models
  62. VHASR: A Multimodal Speech Recognition System With Vision Hotwords
  63. Vript: A Video Is Worth Thousands of Words