PPaperPicks

Honggang Zhang

11 papers at tracked venues · 9 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Amadeus: Autoregressive Model with Bidirectional Attribute Modelling for Symbolic Music
  2. Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation
  3. OCR-Critic: Aligning Multimodal Large Language Models' Perception through Critical Feedback
  4. V-Oracle: Making Progressive Reasoning in Deciphering Oracle Bones for You and Me
  5. VersaGen: Unleashing Versatile Visual Control for Text-to-Image Synthesis
  6. We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
  7. Can Textual Semantics Mitigate Sounding Object Segmentation Preference?
  8. Making Visual Sense of Oracle Bones for You and Me
  9. Ref-AVS: Refer and Segment Objects in Audio-Visual Scenes
  10. Unveiling and Mitigating Bias in Audio Visual Segmentation
  11. Wired Perspectives: Multi-View Wire Art Embraces Generative AI