PPaperPicks

Ziyang Luo

22 papers at tracked venues · 16 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AURORA: Augmented Understanding via Structured Reasoning and Reinforcement Learning for Reference Audio-Visual Segmentation
    AAAI 2026 · Ziyang Luo
  2. Dialectical Structured Reasoning for Explainable Multimodal Fake News Detection
  3. DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs
  4. From Storage to Experience: A Survey on the Evolution of LLM Agent Memory Mechanisms
  5. AdamMeme: Adaptively Probe the Reasoning Capacity of Multimodal Large Language Models on Harmfulness
  6. Aria-UI: Visual Grounding for GUI Instructions
  7. CodeHalu: Investigating Code Hallucinations in LLMs via Execution-based Verification
  8. MM-CRITIC: A Holistic Evaluation of Large Multimodal Models as Multimodal Critique
  9. MemeArena: Automating Context-Aware Unbiased Evaluation of Harmfulness Understanding for Multimodal Large Language Models
  10. SHARP: Unlocking Interactive Hallucination via Stance Transfer in Role-Playing LLMs
  11. ScratchEval: Are GPT-4o Smarter than My Child? Evaluating Large Multimodal Models with Visual Programming Challenges
  12. ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use
  13. TAViS: Text-bridged Audio-Visual Segmentation with Foundation Models
    ICCV 2025 · Ziyang Luo
  14. Tree-of-Evolution: Tree-Structured Instruction Evolution for Code Generation in Large Language Models
    ACL 2025 · Ziyang Luo
  15. VideoAutoArena: An Automated Arena for Evaluating Large Multimodal Models in Video Analysis through User Simulation
    CVPR 2025 · Ziyang Luo
  16. AMR-Evol: Adaptive Modular Response Evolution Elicits Better Knowledge Distillation for Large Language Models in Code Generation
    EMNLP 2024 · Ziyang Luo
  17. CofiPara: A Coarse-to-fine Paradigm for Multimodal Sarcasm Target Identification with Large Multimodal Models
  18. MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems
  19. Towards Explainable Harmful Meme Detection through Multimodal Debate between Large Language Models
  20. Towards Low-Resource Harmful Meme Detection with LMM Agents
  21. VSCode: General Visual Salient and Camouflaged Object Detection with 2D Prompt Learning
    CVPR 2024 · Ziyang Luo
  22. WizardCoder: Empowering Code Large Language Models with Evol-Instruct
    ICLR 2024 · Ziyang Luo