PPaperPicks

Yiran Hu

14 papers at tracked venues · 7 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. PLAWBENCH: A Rubric-Based Benchmark for Evaluating LLMs in Real-World Legal Practice
  2. Can Language Models Replace Programmers for Coding? REPOCOD Says 'Not Yet'
  3. HCLeK: Hierarchical Compression of Legal Knowledge for Retrieval-Augmented Generation
  4. J&H: Evaluating the Robustness of Large Language Models Under Knowledge-Injection Attacks in Legal Domain
    AAAI 2025 · Yiran Hu
  5. JUREX-4E: Juridical Expert-Annotated Four-Element Knowledge Base for Legal Reasoning
  6. JuDGE: Benchmarking Judgment Document Generation for Chinese Legal System
  7. JustEva: A Toolkit to Evaluate LLM Fairness in Legal Knowledge Inference
  8. Legal Fact Prediction: The Missing Piece in Legal Judgment Prediction
  9. LegalAgentBench: Evaluating LLM Agents in Legal Domain
  10. LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation
  11. SELP: Generating Safe and Efficient Task Plans for Robot Agents with Large Language Models
  12. LEEC for Judicial Fairness: A Legal Element Extraction Dataset with Extensive Extra-Legal Labels
  13. STARD: A Chinese Statute Retrieval Dataset Derived from Real-life Queries by Non-professionals
  14. Unsupervised Real-Time Hallucination Detection based on the Internal States of Large Language Models