PPaperPicks

Jingyu Wang

Beijing University of Posts and Telecommunications, State Key Laboratory of Networking and Switching Technology, Beijing, China

46 papers at tracked venues · 31 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Balancing Flow and Collaboration: Exploring Visual Noise Cancellation in Mixed Reality Workspace
  2. Bridging the Tokenizer Gap: Semantics and Distribution-aware Knowledge Transfer for Unbiased Cross-Tokenizer Distillation
  3. DIVER: A Robust Text-to-SQL System with Dynamic Interactive Value Linking and Evidence Reasoning
  4. DiffNBR: Spatio-Temporal Diffusion with Information Bottleneck for Next-Basket Recommendation
  5. Erasing Without Remembering: Implicit Knowledge Forgetting in Large Language Models
  6. Example Quality Matters: Multi-Aspects Example Augmentation for Private Library Programming
  7. InkFlow: Connected Handwriting Recognition for Natural Mid-Air Input in Mixed Reality
  8. ModularMoE: Fast LLM Customization with Parameter-Sharing Mixture-of-Experts for Low-Resource Settings
  9. VALU: A Benchmark for Video Anomaly Temporal Localization and Understanding at Multiple Semantic Levels
  10. A Dual-Branch 3D Spatial-Aware Latent Diffusion for Realistic Depth Image Synthesis
  11. A³-Net: Calibration-Free Multi-View 3D Hand Reconstruction for Enhanced Musical Instrument Learning
  12. Beyond Statistical Analysis: Multimodal Framework for Time Series Forecasting with LLM-Driven Temporal Pattern
  13. ChatTime: A Unified Multimodal Time Series Foundation Model Bridging Numerical and Textual Data
  14. ClusterAttn: KV Cache Compression under Intrinsic Attention Clustering
  15. Do LVLMs Truly Understand Video Anomalies? Revealing Hallucination via Co-Occurrence Patterns
  16. Efficient Inter-Operator Scheduling for Concurrent Recommendation Model Inference on GPU
  17. Evaluating and Mitigating Object Hallucination in Large Vision-Language Models: Can They Still See Removed Objects?
  18. Foresail: LLM Sensor Knowledge Empowered Status-guided Network for Multivariate Time-series Classification
  19. Generalizable Hand-Object Modeling from Monocular RGB Images via 3D Gaussians
  20. Perception-R1: Pioneering Perception Policy with Reinforcement Learning
  21. Pose-Guided Temporal Enhancement for Robust Low-Resolution Hand Reconstruction
  22. Prior-Aware Dynamic Temporal Modeling Framework for Sequential 3D Hand Pose Estimation
  23. Rethinking Smoothness for Fast and Adaptable Entity Alignment Decoding
  24. Robustness Verification of Deep Graph Neural Networks Tightened by Linear Approximation
  25. Rule Meets Learning: Confidence-Aware Multi-View Fusion for Self-Supervised 3D Hand Pose Estimation
  26. The Threat of PROMPTS in Large Language Models: A System and User Prompt Perspective
  27. Towards Bare-Hand Interaction for Whiteboard Collaboration in Virtual Reality
  28. Unhackable Temporal Reward for Scalable Video MLLMs
  29. Unified 2D-3D Discrete Priors for Noise-Robust and Calibration-Free Multiview 3D Human Pose Estimation
  30. Unveiling Internal Reasoning Modes in LLMs: A Deep Dive into Latent Reasoning vs. Factual Shortcuts with Attribute Rate Ratio
  31. Watch, Skip, Repeat: Hotspot-Aware Joint Optimization for Video Streaming
  32. Coarse-to-Fine Implicit Representation Learning for 3D Hand-Object Reconstruction from a Single RGB-D Image
  33. Dynamic Support Information Mining for Category-Agnostic Pose Estimation
  34. FM-Delta: Lossless Compression for Storing Massive Fine-tuned Foundation Models
  35. HPipe: Large Language Model Pipeline Parallelism for Long Context on Heterogeneous Cost-effective Devices
  36. Interdependency Matters: Graph Alignment for Multivariate Time Series Anomaly Detection
  37. Keypoint Fusion for RGB-D Based 3D Hand Pose Estimation
  38. MDR: Model-Specific Demonstration Retrieval at Inference Time for In-Context Learning
  39. Multi-Scale Video Anomaly Detection by Multi-Grained Spatio-Temporal Representation Learning
  40. Pre-Tokenization of Numbers for Large Language Models
  41. Rethinking the Power of Timestamps for Robust Time Series Forecasting: A Global-Local Fusion Perspective
  42. SSS: Editing Factual Knowledge in Language Models towards Semantic Sparse Space
  43. STAR-VP: Improving Long-term Viewport Prediction in 360° Videos via Space-aligned and Time-varying Fusion
  44. Safeguarding Sustainable Cities: Unsupervised Video Anomaly Detection through Diffusion-based Latent Pattern Learning
  45. Towards Semantic Consistency: Dirichlet Energy Driven Robust Multi-Modal Entity Alignment
  46. Video Anomaly Detection via Progressive Learning of Multiple Proxy Tasks