PPaperPicks

Thien Huu Nguyen

37 papers at tracked venues · 21 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. CypherSmith: Transforming Text-to-Cypher Generation for LLMs with Synthetic Data
  2. Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI
  3. GloCTM: Cross-Lingual Topic Modeling via a Global Context Space
  4. Lizard: An Efficient Linearization Framework for Large Language Models
  5. Octopus: Gated Selective Attention for Memory-Bounded Long-Context Inference in Large Language Models
  6. TALAS: Teacher-Anchored Layer Alignment with Adaptive Sharpness-Aware Minimization for Embedding Distillation
  7. Towards Fast and Accurate Modeling for Cross-Lingual Label Projection
  8. A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality
  9. Adaptive Prompting for Continual Relation Extraction: A Within-Task Variance Perspective
  10. EMO: Embedding Model Distillation via Intra-Model Relation and Optimal Transport Alignments
  11. Enhancing Discriminative Representation in Similar Relation Clusters for Few-Shot Continual Relation Extraction
  12. Few-Shot, No Problem: Descriptive Continual Relation Extraction
  13. From Selection to Generation: A Survey of LLM-based Active Learning
  14. GUI Agents: A Survey
  15. GloCOM: A Short Text Neural Topic Model via Global Clustering Context
  16. HiCOT: Improving Neural Topic Models via Optimal Transport and Contrastive Learning
  17. Improving Vietnamese-English Cross-Lingual Retrieval for Legal and General Domains
  18. LUSIFER: Language Universal Space Integration for Enhanced Representation in Multilingual Text Embedding Models
  19. MaGiX: A Multi-Granular Adaptive Graph Intelligence Framework for Enhancing Cross-Lingual RAG
  20. Massively Multilingual Instruction-Following Information Extraction
  21. Mitigating Non-Representative Prototypes and Representation Bias in Few-Shot Continual Relation Extraction
  22. Multi-Surrogate-Objective Optimization for Neural Topic Models
  23. Mutual-pairing Data Augmentation for Fewshot Continual Relation Extraction
  24. Sharpness-Aware Minimization for Topic Models with High-Quality Document Representations
  25. ToVo: Toxicity Taxonomy via Voting
  26. Topic Modeling for Short Texts via Optimal Transport-Based Clustering
  27. Continual Relation Extraction via Sequential Multi-Task Learning
  28. Counterfactual Augmentation for Robust Authorship Representation Learning
  29. Identifying Speakers in Dialogue Transcripts: A Text-based Approach Using Pretrained Language Models
  30. Lifelong Event Detection via Optimal Transport
  31. MCECR: A Novel Dataset for Multilingual Cross-Document Event Coreference Resolution
  32. Mastering Context-to-Label Representation Transformation for Event Causality Identification with Diffusion Models
  33. NeuroMax: Enhancing Neural Topic Modeling via Maximizing Mutual Information and Group Topic Regularization
  34. Preserving Generalization of Language models in Few-shot Continual Relation Extraction
  35. Realistic Evaluation of Toxicity in Large Language Models
  36. SharpSeq: Empowering Continual Event Detection through Sharpness-Aware Sequential-task Learning
  37. ULLME: A Unified Framework for Large Language Model Embeddings with Generation-Augmented Learning