PPaperPicks

Yunhe Wang

Noah's Ark Lab, Huawei Technologie, Beijing, China

35 papers at tracked venues · 29 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. GenVidBench: A 6-Million Benchmark for AI-Generated Video Detection
  2. Multi-Granularity Semantic Revision for Large Language Model Distillation
  3. CBQ: Cross-Block Quantization for Large Language Models
  4. CFinBench: A Comprehensive Chinese Financial Benchmark for Large Language Models
  5. DECO: Unleashing the Potential of ConvNets for Query-based Detection and Segmentation
  6. DenseSSM: State Space Models with Dense Hidden Connection for Efficient Large Language Models
  7. EMS-SD: Efficient Multi-sample Speculative Decoding for Accelerating Large Language Models
  8. Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
  9. Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning
  10. GPT4Image: Large Pre-trained Models Help Vision Models Learn Better on Perception Task
  11. LLM Data Selection and Utilization via Dynamic Bi-level Optimization
  12. Linear Multistep Solver Distillation for Fast Sampling of Diffusion Models
  13. Mixture of Lookup Experts
  14. SlimLLM: Accurate Structured Pruning for Large Language Models
  15. TinySAM: Pushing the Envelope for Efficient Segment Anything Model
  16. U-REPA: Aligning Diffusion U-Nets to ViTs
  17. Adapt Without Forgetting: Distill Proximity from Dual Teachers in Vision-Language Models
  18. An Empirical Study of Scaling Law for Scene Text Recognition
  19. Context-Guided Spatial Feature Reconstruction for Efficient Semantic Segmentation
  20. DiJiang: Efficient Large Language Models through Compact Kernelization
  21. Distilling Semantic Priors from SAM to Efficient Image Restoration Models
  22. Enhancing Large Language Models through Adaptive Tokenizers
  23. ExCP: Extreme LLM Checkpoint Compression via Weight-Momentum Joint Shrinking
  24. Image Processing GNN: Breaking Rigidity in Super-Resolution
  25. Kangaroo: Lossless Self-Speculative Decoding for Accelerating LLMs via Double Early Exiting
  26. Memory-Space Visual Prompting for Efficient Vision-Language Fine-Tuning
  27. MemoryFormer : Minimize Transformer Computation by Removing Fully-Connected Layers
  28. Multiscale Positive-Unlabeled Detection of AI-Generated Texts
  29. ParameterNet: Parameters are All You Need for Large-Scale Visual Pretraining of Mobile Networks
  30. Rethinking Optimization and Architecture for Tiny Language Models
  31. SLAB: Efficient Transformers with Simplified Linear Attention and Progressive Re-parameterized Batch Normalization
  32. Star-Agents: Automatic Data Optimization with LLM Agents for Instruction Tuning
  33. Token Compensator: Altering Inference Cost of Vision Transformer Without Re-tuning
  34. U-DiTs: Downsample Tokens in U-Shaped Diffusion Transformers
  35. UFineBench: Towards Text-based Person Retrieval with Ultra-fine Granularity