P
PaperPicks
Conferences
Siyu Yuan
32 papers at tracked venues · 18 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
ACL
×9
EMNLP
×8
NeurIPS
×6
NAACL
×4
AAAI
×1
CHI
×1
CIKM
×1
ICML
×1
PerCom
×1
Frequent coauthors
Ruihan Yang
DBLP profile ↗
ORCID search ↗
×3
Rui Xu
DBLP profile ↗
ORCID search ↗
×2
Xintao Wang
DBLP profile ↗
ORCID search ↗
×2
Nianqi Li
DBLP profile ↗
ORCID search ↗
×2
Zhijun Xu
DBLP profile ↗
ORCID search ↗
×2
Yikai Zhang
DBLP profile ↗
ORCID search ↗
×2
Qianyu He
DBLP profile ↗
ORCID search ↗
×1
Shiting Huang
DBLP profile ↗
ORCID search ↗
×1
Sizhen Bian
DBLP profile ↗
ORCID search ↗
×1
Weiyuan Li
DBLP profile ↗
ORCID search ↗
×1
Aili Chen
DBLP profile ↗
ORCID search ↗
×1
Jiangjie Chen
DBLP profile ↗
ORCID search ↗
×1
Papers
Metaphor Reasoning is Meta-reasoning
ACL 2026
·
Qianyu He
DBLP profile ↗
ORCID search ↗
ARIA: Training Language Agents with Intention-driven Reward Aggregation
NeurIPS 2025
·
Ruihan Yang
DBLP profile ↗
ORCID search ↗
CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios
EMNLP 2025
·
Shiting Huang
DBLP profile ↗
ORCID search ↗
Character is Destiny: Can Persona-assigned Language Models Make Personal Choices?
EMNLP 2025
·
Rui Xu
DBLP profile ↗
ORCID search ↗
CoSER: Coordinating LLM-Based Persona Simulation of Established Roles
ICML 2025
·
Xintao Wang
DBLP profile ↗
ORCID search ↗
Collaborative Human Activity Recognition with Passive Inter-Body Electrostatic Field
PerCom 2025
·
Sizhen Bian
DBLP profile ↗
ORCID search ↗
Curse of Knowledge: Your Guidance and Provided Knowledge are biasing LLM Judges in Complex Evaluation
EMNLP 2025
·
Weiyuan Li
DBLP profile ↗
ORCID search ↗
DEEPER Insight into Your User: Directed Persona Refinement for Dynamic Persona Modeling
ACL 2025
·
Aili Chen
DBLP profile ↗
ORCID search ↗
EASYTOOL: Enhancing LLM-based Agents with Concise Tool Instruction
NAACL 2025
·
Siyu Yuan
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
NeurIPS 2025
·
Jiangjie Chen
DBLP profile ↗
ORCID search ↗
EvoAgent: Towards Automatic Multi-Agent Generation via Evolutionary Algorithms
NAACL 2025
·
Siyu Yuan
Implicit Reasoning in Transformers is Reasoning through Shortcuts
ACL 2025
·
Tianhe Lin
DBLP profile ↗
ORCID search ↗
KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation
NeurIPS 2025
·
Jiajun Shi
DBLP profile ↗
ORCID search ↗
LLM-Powered Information Extraction for the Dairy Financial Domain: Tackling Data Scarcity and Ambiguity
CIKM 2025
·
Chunyan An
DBLP profile ↗
ORCID search ↗
MultiLingPoT: Boosting Mathematical Reasoning in LLMs through Multilingual Program Integration
EMNLP 2025
·
Nianqi Li
DBLP profile ↗
ORCID search ↗
ORIGAMISPACE: Benchmarking Multimodal LLMs in Multi-Step Spatial Reasoning with Mathematical Constraints
NeurIPS 2025
·
Rui Xu
DBLP profile ↗
ORCID search ↗
Past Meets Present: Creating Historical Analogy with Large Language Models
ACL 2025
·
Nianqi Li
DBLP profile ↗
ORCID search ↗
PunMemeCN: A Benchmark to Explore Vision-Language Models' Understanding of Chinese Pun Memes
EMNLP 2025
·
Zhijun Xu
DBLP profile ↗
ORCID search ↗
Revealing the Barriers of Language Agents in Planning
NAACL 2025
·
Jian Xie
DBLP profile ↗
ORCID search ↗
SELFGOAL: Your Language Agents Already Know How to Achieve High-level Goals
NAACL 2025
·
Ruihan Yang
DBLP profile ↗
ORCID search ↗
The Lighthouse of Language: Enhancing LLM Agents via Critique-Guided Improvement
NeurIPS 2025
·
Ruihan Yang
DBLP profile ↗
ORCID search ↗
ToolHop: A Query-Driven Benchmark for Evaluating Large Language Models in Multi-Hop Tool Use
ACL 2025
·
Junjie Ye
DBLP profile ↗
ORCID search ↗
Unlocking Scientific Concepts: How Effective Are LLM-Generated Analogies for Student Understanding and Classroom Practice?
CHI 2025
·
Zekai Shao
DBLP profile ↗
ORCID search ↗
"A good pun is its own reword": Can Large Language Models Understand Puns?
EMNLP 2024
·
Zhijun Xu
DBLP profile ↗
ORCID search ↗
ANALOGYKB: Unlocking Analogical Reasoning of Language Models with A Million-scale Knowledge Base
ACL 2024
·
Siyu Yuan
Boosting Scientific Concepts Understanding: Can Analogy from Teacher Models Empower Student Models?
EMNLP 2024
·
Siyu Yuan
Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works
EMNLP 2024
·
Xinfeng Yuan
DBLP profile ↗
ORCID search ↗
InCharacter: Evaluating Personality Fidelity in Role-Playing Agents through Psychological Interviews
ACL 2024
·
Xintao Wang
DBLP profile ↗
ORCID search ↗
Light Up the Shadows: Enhance Long-Tailed Entity Grounding with Concept-Guided Vision-Language Models
ACL 2024
·
Yikai Zhang
DBLP profile ↗
ORCID search ↗
TaskBench: Benchmarking Large Language Models for Task Automation
NeurIPS 2024
·
Yongliang Shen
DBLP profile ↗
ORCID search ↗
TimeArena: Shaping Efficient Multitasking Language Agents in a Time-Aware Simulation
ACL 2024
·
Yikai Zhang
DBLP profile ↗
ORCID search ↗
Translate Meanings, Not Just Words: IdiomKB's Role in Optimizing Idiomatic Translation with Language Models
AAAI 2024
·
Shuang Li
DBLP profile ↗
ORCID search ↗