PPaperPicks

Fandong Meng

64 papers at tracked venues · 51 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. APB-V: Accelerating Long-Video Understanding via Sequence-Parallelism-aware Approximate Attention
  2. ArrowGEV: Grounding Events in Video via Learning the Arrow of Time
  3. Figure It Out: Improve the Frontier of Reasoning with Executable Visual States
  4. GRAM-R²: Self-Training Generative Foundation Reward Models for Reward Reasoning
  5. Investigating Cross-Modal Skill Injection: Scenarios, Methods, and Hyperparameters
  6. Joint Optimization of Training Data and Policy in RLHF
  7. Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning
  8. ReFreeKV: Towards Threshold-Free KV Cache Compression
  9. SED-SFT: Selectively Encouraging Diversity in Supervised Fine-Tuning
  10. Think Natively: Unlocking Multilingual Reasoning with Consistency-Enhanced Reinforcement Learning
  11. A Law Reasoning Benchmark for LLM with Tree-Organized Structures including Factum Probandum, Evidence and Experiences
  12. A Self-Denoising Model for Robust Few-Shot Relation Extraction
  13. AVG-LLaVA: An Efficient Large Multimodal Model with Adaptive Visual Granularity
  14. Advancing SMoE for Continuous Domain Adaptation of MLLMs: Adaptive Router and Domain-Specific Loss
  15. An Empirical Study of Many-to-Many Summarization with Large Language Models
  16. Beyond Next Token Prediction: Patch-Level Training for Large Language Models
  17. CM-Align: Consistency-based Multilingual Alignment for Large Language Models
  18. ConCISE: Confidence-guided Compression in Step-by-step Efficient Reasoning
  19. Continuous Visual Autoregressive Generation via Score Maximization
  20. DRT: Deep Reasoning Translation via Long Chain-of-Thought
  21. DelTA: An Online Document-Level Translation Agent Based on Multi-Level Memory
  22. Dense Retrievers Can Fail on Simple Queries: Revealing The Granularity Dilemma of Embeddings
  23. Efficient Speech Language Modeling via Energy Distance in Continuous Latent Space
  24. Enhancing Cross-Tokenizer Knowledge Distillation with Contextual Dynamical Mapping
  25. LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning
  26. Less, but Better: Efficient Multilingual Expansion for LLMs via Layer-wise Mixture-of-Experts
  27. LongDPO: Unlock Better Long-form Generation Abilities for LLMs via Critique-augmented Stepwise Information
  28. MiniPLM: Knowledge Distillation for Pre-training Language Models
  29. Personalized Language Model Learning on Text Data Without User Identifiers
  30. PunchBench: Benchmarking MLLMs in Multimodal Punchline Comprehension
  31. Retrieval-Augmented Machine Translation with Unstructured Knowledge
  32. THOR-MoE: Hierarchical Task-Guided and Context-Responsive Routing for Neural Machine Translation
  33. TIU-Bench: A Benchmark for Evaluating Large Multimodal Models on Text-rich Image Understanding
  34. BranchNorm: Robustly Scaling Extremely Deep Transformers
  35. C-LLM: Learn to Check Chinese Spelling Errors Character by Character
  36. CSCD-NS: a Chinese Spelling Check Dataset for Native Speakers
  37. Comments as Natural Logic Pivots: Improve Code Generation via Comment Perspective
  38. Continual Learning with Semi-supervised Contrastive Distillation for Incremental Neural Machine Translation
  39. Cross-Lingual Knowledge Editing in Large Language Models
  40. Enhancing Byzantine-Resistant Aggregations with Client Embedding
  41. Exploring Conditional Variational Mechanism to Pinyin Input Method for Addressing One-to-Many Mappings in Low-Resource Scenarios
  42. Generative Multi-Modal Knowledge Retrieval with Large Language Models
  43. Improving Machine Translation with Large Language Models: A Preliminary Study with Cooperative Decoding
  44. Instruction Position Matters in Sequence Generation with Large Language Models
  45. LCS: A Language Converter Strategy for Zero-Shot Neural Machine Translation
  46. Language Generation with Strictly Proper Scoring Rules
  47. Large Language Models Are Not Robust Multiple Choice Selectors
  48. LexMatcher: Dictionary-centric Data Curation for LLM-based Machine Translation
  49. Multi-Level Cross-Modal Alignment for Speech Relation Extraction
  50. On Large Language Models' Hallucination with Regard to Known Facts
  51. On Prompt-Driven Safeguarding for Large Language Models
  52. On the token distance modeling ability of higher RoPE attention dimension
  53. Outdated Issue Aware Decoding for Factual Knowledge Editing
  54. Plot Retrieval as an Assessment of Abstract Semantic Association
  55. TasTe: Teaching Large Language Models to Translate through Self-Reflection
  56. Teaching Large Language Models to Translate with Comparison
  57. Towards Codable Watermarking for Injecting Multi-Bits Information to LLMs
  58. Towards Multiple References Era - Addressing Data Leakage and Limited Reference Diversity in Machine Translation Evaluation
  59. Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation
  60. Tree-of-Reasoning Question Decomposition for Complex Question Answering with Large Language Models
  61. Trust in Internal or External Knowledge? Generative Multi-Modal Entity Linking with Knowledge Retriever
  62. Understanding and Addressing the Under-Translation Problem from the Perspective of Decoding Objective
  63. Unsupervised Information Refinement Training of Large Language Models for Retrieval-Augmented Generation
  64. XAL: EXplainable Active Learning Makes Classifiers Better Low-resource Learners