PPaperPicks

Anton van den Hengel

University of Adelaide, Australia

30 papers at tracked venues · 26 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. The Devil is in the Distributions: Explicit Modeling of Scene Content is Key in Zero-Shot Video Captioning
  2. Analytic DAG Constraints for Differentiable DAG Learning
  3. CausalMVC: Causal Content-Style Representation Learning for Deep Multi-View Clustering
  4. EmoDubber: Towards High Quality and Emotion Controllable Movie Dubbing
  5. FlowDubber: Movie Dubbing with LLM-based Semantic-aware Learning and Flow Matching based Voice Enhancing
  6. Interactive Medical Image Analysis with Concept-based Similarity Reasoning
  7. Let Your Video Listen to Your Music! - Beat-Aligned, Content-Preserving Video Editing with Arbitrary Music
  8. Looking in the Mirror: A Faithful Counterfactual Explanation Method for Interpreting Deep Image Classification Models
  9. Medusa: A Multi-Scale High-order Contrastive Dual-Diffusion Approach for Multi-View Clustering
  10. On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
  11. Open-World Objectness Modeling Unifies Novel Object Detection
  12. PedCLIP: A Vision-Language Model for Pediatric X-Rays with Mixture of Body Part Experts
  13. Primitive Vision: Improving Diagram Understanding in MLLMs
  14. Prosody-Enhanced Acoustic Pre-training and Acoustic-Disentangled Prosody Adapting for Movie Dubbing
  15. RandLoRA: Full rank parameter-efficient fine-tuning of large models
  16. Seeing the Trees for the Forest: Rethinking Weakly-Supervised Medical Visual Grounding
  17. Separation of Powers: On Segregating Knowledge from Observation in LLM-enabled Knowledge-based Visual Question Answering
  18. Synergy and Diversity in CLIP: Enhancing Performance Through Adaptive Backbone Ensembling
    ICLR 2025 ·
    Cristian Rodriguez Opazo
  19. Towards Higher Effective Rank in Parameter-Efficient Fine-Tuning Using Khatri-Rao Product
  20. AdaCBM: An Adaptive Concept Bottleneck Model for Explainable and Accurate Diagnosis
  21. BLiRF: Bandlimited Radiance Fields for Dynamic Scene Modeling
  22. CAPE: CAM as a Probabilistic Ensemble for Enhanced DNN Interpretation
  23. From Speaker to Dubber: Movie Dubbing with Prosody and Duration Consistency Learning
  24. Identifiable Latent Polynomial Causal Models through the Lens of Change
  25. Improving the Convergence of Dynamic NeRFs via Optimal Transport
  26. Knowledge Composition using Task Vectors with Learned Anisotropic Scaling
  27. MARec: Metadata Alignment for cold-start Recommendation
  28. StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing
  29. ViewFusion: Towards Multi-View Consistency via Interpolated Denoising
  30. Weakly Supervised Video Individual Counting