PPaperPicks

Wei Hu

22 papers at tracked venues · 15 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. How Do Answer Tokens Read Reasoning Traces? Self-Reading Patterns in Thinking LLMs for Quantitative Reasoning
  2. Skill-Aware Data Selection and Fine-Tuning for Data-Efficient Reasoning Distillation
  3. To Diff or Not to Diff? Structure-Aware and Adaptive Output Formats for Efficient LLM-based Code Editing
  4. Benign Overfitting in Single-Head Attention
  5. ChatSOP: An SOP-Guided MCTS Planning Framework for Controllable LLM Dialogue Agents
  6. HyperGCT: A Dynamic Hyper-GNN-Learned Geometric Constraint for 3D Registration
  7. Let Me Grok for You: Accelerating Grokking via Embedding Transfer from a Weaker Model
  8. Mixture of LoRA Experts for Continual Information Extraction with LLMs
  9. Swing-by Dynamics in Concept Learning and Compositional Generalization
  10. What Happens During the Loss Plateau? Understanding Abrupt Learning in Transformers
  11. A Multi-Node Multi-GPU Distributed GNN Training Framework for Large-Scale Online Advertising
  12. Abrupt Learning in Transformers: A Case Study on Matrix Completion
  13. Benign Overfitting and Grokking in ReLU Networks for XOR Cluster Data
  14. DFlow: A Generative Model Combining Denoising AutoEncoder and Normalizing Flow for High Fidelity Waveform Generation
  15. DHGCN: Dynamic Hop Graph Convolution Network for Self-Supervised Point Cloud Learning
  16. E-Paraformer: A Faster and Better Parallel Transformer for Non-autoregressive End-to-End Mandarin Speech Recognition
  17. Geospatial Topological Relation Extraction from Text with Knowledge Augmentation
    SDM 2024 · Wei Hu
  18. How Do Transformers Learn In-Context Beyond Simple Functions? A Case Study on Learning with Representations
  19. Improving Multilingual Text-to-Speech with Mixture-of-Language-Experts and Accent Disentanglement
  20. Multi-frequency Attention Approach for Enhanced Ultrasound Image Segmentation
  21. Near-Interpolators: Rapid Norm Growth and the Trade-Off between Interpolation and Generalization
  22. Understanding Surprising Generalization Phenomena in Deep Learning
    AAAI 2024 · Wei Hu