P
PaperPicks
Conferences
Wei Hu
22 papers at tracked venues · 15 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
ACL
×4
ICLR
×4
NeurIPS
×3
AAAI
×2
InterSpeech
×2
AISTATS
×1
CIKM
×1
EMNLP
×1
ICCV
×1
ICML
×1
MICCAI
×1
SDM
×1
Frequent coauthors
Zhiwei Xu
DBLP profile ↗
ORCID search ↗
×2
Pulkit Gopalani
DBLP profile ↗
ORCID search ↗
×2
Haoyang Chen
DBLP profile ↗
ORCID search ↗
×1
Lechen Zhang
DBLP profile ↗
ORCID search ↗
×1
Wei Cheng
DBLP profile ↗
ORCID search ↗
×1
Roey Magen
DBLP profile ↗
ORCID search ↗
×1
Zhigen Li
DBLP profile ↗
ORCID search ↗
×1
Xiyu Zhang
DBLP profile ↗
ORCID search ↗
×1
Zitao Wang
DBLP profile ↗
ORCID search ↗
×1
Yongyi Yang
DBLP profile ↗
ORCID search ↗
×1
Xuewu Jiao
DBLP profile ↗
ORCID search ↗
×1
Chenfeng Miao
DBLP profile ↗
ORCID search ↗
×1
Papers
How Do Answer Tokens Read Reasoning Traces? Self-Reading Patterns in Thinking LLMs for Quantitative Reasoning
ACL 2026
·
Haoyang Chen
DBLP profile ↗
ORCID search ↗
Skill-Aware Data Selection and Fine-Tuning for Data-Efficient Reasoning Distillation
ACL 2026
·
Lechen Zhang
DBLP profile ↗
ORCID search ↗
To Diff or Not to Diff? Structure-Aware and Adaptive Output Formats for Efficient LLM-based Code Editing
ACL 2026
·
Wei Cheng
DBLP profile ↗
ORCID search ↗
Benign Overfitting in Single-Head Attention
NeurIPS 2025
·
Roey Magen
DBLP profile ↗
ORCID search ↗
ChatSOP: An SOP-Guided MCTS Planning Framework for Controllable LLM Dialogue Agents
ACL 2025
·
Zhigen Li
DBLP profile ↗
ORCID search ↗
HyperGCT: A Dynamic Hyper-GNN-Learned Geometric Constraint for 3D Registration
ICCV 2025
·
Xiyu Zhang
DBLP profile ↗
ORCID search ↗
Let Me Grok for You: Accelerating Grokking via Embedding Transfer from a Weaker Model
ICLR 2025
·
Zhiwei Xu
DBLP profile ↗
ORCID search ↗
Mixture of LoRA Experts for Continual Information Extraction with LLMs
EMNLP 2025
·
Zitao Wang
DBLP profile ↗
ORCID search ↗
Swing-by Dynamics in Concept Learning and Compositional Generalization
ICLR 2025
·
Yongyi Yang
DBLP profile ↗
ORCID search ↗
What Happens During the Loss Plateau? Understanding Abrupt Learning in Transformers
NeurIPS 2025
·
Pulkit Gopalani
DBLP profile ↗
ORCID search ↗
A Multi-Node Multi-GPU Distributed GNN Training Framework for Large-Scale Online Advertising
CIKM 2024
·
Xuewu Jiao
DBLP profile ↗
ORCID search ↗
Abrupt Learning in Transformers: A Case Study on Matrix Completion
NeurIPS 2024
·
Pulkit Gopalani
DBLP profile ↗
ORCID search ↗
Benign Overfitting and Grokking in ReLU Networks for XOR Cluster Data
ICLR 2024
·
Zhiwei Xu
DBLP profile ↗
ORCID search ↗
DFlow: A Generative Model Combining Denoising AutoEncoder and Normalizing Flow for High Fidelity Waveform Generation
ICML 2024
·
Chenfeng Miao
DBLP profile ↗
ORCID search ↗
DHGCN: Dynamic Hop Graph Convolution Network for Self-Supervised Point Cloud Learning
AAAI 2024
·
Jincen Jiang
DBLP profile ↗
ORCID search ↗
E-Paraformer: A Faster and Better Parallel Transformer for Non-autoregressive End-to-End Mandarin Speech Recognition
InterSpeech 2024
·
Kun Zou
DBLP profile ↗
ORCID search ↗
Geospatial Topological Relation Extraction from Text with Knowledge Augmentation
SDM 2024
·
Wei Hu
How Do Transformers Learn In-Context Beyond Simple Functions? A Case Study on Learning with Representations
ICLR 2024
·
Tianyu Guo
DBLP profile ↗
ORCID search ↗
Improving Multilingual Text-to-Speech with Mixture-of-Language-Experts and Accent Disentanglement
InterSpeech 2024
·
Jing Wu
DBLP profile ↗
ORCID search ↗
Multi-frequency Attention Approach for Enhanced Ultrasound Image Segmentation
MICCAI 2024
·
Zicheng Hu
DBLP profile ↗
ORCID search ↗
Near-Interpolators: Rapid Norm Growth and the Trade-Off between Interpolation and Generalization
AISTATS 2024
·
Yutong Wang
DBLP profile ↗
ORCID search ↗
Understanding Surprising Generalization Phenomena in Deep Learning
AAAI 2024
·
Wei Hu