PPaperPicks

Dong Li

AMD Inc., Advanced Micro Devices Inc., Beijing, China

17 papers at tracked venues · 14 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. DiffBench Meets DiffAgent: End-to-End LLM-Driven Diffusion Acceleration Code Generation
  2. Learnable Permutation for Structured Sparsity on Transformer Models
  3. MoEC: A Memory-Routed Mixture-of-Experts Controller for Adaptive Minecraft Control
  4. SparK: Query-Aware Unstructured Sparsity with Recoverable KV Cache Channel Pruning
  5. Theory-optimal Quantization Based on Flatness
  6. Amphista: Bi-directional Multi-head Decoding for Accelerating LLM Inference
  7. BaWA: Automatic Optimizing Pruning Metric for Large Language Models with Balanced Weight and Activation
  8. EGSRAL: An Enhanced 3D Gaussian Splatting Based Renderer with Automated Labeling for Large-Scale Driving Scene
  9. Gumiho: A Hybrid Architecture to Prioritize Early Tokens in Speculative Decoding
  10. ReNeg: Learning Negative Embedding with Reward Guidance
  11. Týr-the-Pruner: Structural Pruning LLMs via Global Sparsity Distribution Optimization
  12. DL-QAT: Weight-Decomposed Low-Rank Quantization-Aware Training for Large Language Models
  13. DiP-GO: A Diffusion Pruner via Few-step Gradient Optimization
  14. Enhancing Vision Transformer: Amplifying Non-Linearity in Feedforward Network Module
  15. MonoGS++: Fast and Accurate Monocular RGB Gaussian SLAM
  16. QT-ViT: Improving Linear Attention in ViT with Quadratic Taylor Expansion
  17. UPDP: A Unified Progressive Depth Pruner for CNN and Vision Transformer