PPaperPicks

Zhouhong Gu

9 papers at tracked venues · 7 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Are Large Language Models Reliable Reviewers? A Benchmark for Error Detection in Financial Documents
  2. PII-Bench: Evaluating Query-Aware Privacy Protection Systems
  3. The "Knowledge-Behavior Gap" in Cultural Taboo Safety of Large Language Models
  4. GAPO: Learning Preferential Prompt through Generative Adversarial Policy Optimization
    ACL 2025 · Zhouhong Gu
  5. MIRAGE: Exploring How Large Language Models Perform in Complex Social Interactive Environments
  6. StrucText-Eval: Evaluating Large Language Model's Reasoning Ability in Structure-Rich Text
    ACL 2025 · Zhouhong Gu
  7. AutoScraper: A Progressive Understanding Web Agent for Web Scraper Generation
  8. DetectBench: Can Large Language Model Detect and Piece Together Implicit Evidence?
    EMNLP 2024 · Zhouhong Gu
  9. Xiezhi: An Ever-Updating Benchmark for Holistic Domain Knowledge Evaluation
    AAAI 2024 · Zhouhong Gu