PPaperPicks

Zhaozhuo Xu

32 papers at tracked venues · 19 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Collision to Cognition: Hash-Driven Graph Construction for Efficient RAG
  2. Copyright Detective: A Forensic System to Evidence LLMs Flickering Copyright Leakage Risks
  3. Query-Aware Knowledge Retrieval via Hyperbolic Structuring
  4. ALinFiK: Learning to Approximate Linearized Future Influence Kernel for Scalable Third-Parity LLM Data Valuation
  5. Building Self-Awareness of LLMs Over Weight Quantization
  6. Compression-Aware Computing for Scalable and Sustainable AI
    AAAI 2025 · Zhaozhuo Xu
  7. DEL-ToM: Inference-Time Scaling for Theory-of-Mind Reasoning via Dynamic Epistemic Logic
  8. Dynamic Maintenance of Kernel Density Estimation Data Structure: From Practice to Theory
  9. Position: Iterative Online-Offline Joint Optimization is Needed to Manage Complex LLM Copyright Risks
  10. Profiling LLM's Copyright Infringement Risks under Adversarial Persuasive Prompting
  11. Rescorla-Wagner Steering of LLMs for Undesired Behaviors over Disproportionate Inappropriate Context
  12. Retrieval Augmented Zero-Shot Enzyme Generation for Specified Substrate
  13. Sketch to Adapt: Fine-Tunable Sketches for Efficient LLM Adaptation
  14. Taming Language Models for Text-attributed Graph Learning with Decoupled Aggregation
  15. Word Salad Chopper: Reasoning Models Waste A Ton Of Decoding Budget On Useless Repetitions, Self-Knowingly
  16. Zeroth-Order Fine-Tuning of LLMs with Transferable Static Sparsity
  17. Do LLMs Know to Respect Copyright Notice?
  18. FinCon: A Synthesized LLM Multi-Agent System with Conceptual Verbal Reinforcement for Enhanced Financial Decision Making
  19. GNNs Also Deserve Editing, and They Need It More Than Once
  20. In Defense of Structural Sparse Adapters for Concurrent LLM Serving
  21. KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache
  22. KV Cache Compression, But What Must We Give in Return? A Comprehensive Benchmark of Long Context Capable Approaches
  23. KV Cache is 1 Bit Per Channel: Efficient Large Language Model Inference with Coupled Quantization
  24. Knowledge Graphs Can be Learned with Just Intersection Features
  25. NoMAD-Attention: Efficient LLM Inference on CPUs Through Multiply-add-free Attention
  26. QUEST: Efficient Extreme Multi-Label Text Classification with Large Language Models on Commodity Hardware
  27. SIRIUS : Contexual Sparisty with Correction for Efficient LLMs
  28. ScaleLLM: A Resource-Frugal LLM Serving Framework by Optimizing End-to-End Efficiency
  29. Soft Prompt Recovers Compressed LLMs, Transferably
    ICML 2024 · Zhaozhuo Xu
  30. TVE: Learning Meta-attribution for Transferable Vision Explainer
  31. TensorOpera Router: A Multi-Model Router for Efficient LLM Inference
  32. Token-wise Influential Training Data Retrieval for Large Language Models