PPaperPicks

Volker Tresp

Ludwig Maximilian University of Munich, Germany

30 papers at tracked venues · 17 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AUVIC: Adversarial Unlearning of Visual Concepts for Multi-modal Large Language Models
  2. Memory-R1: Enhancing Large Language Model Agents to Manage and Utilize Memories via Reinforcement Learning
  3. OpenDriveVLA: Towards End-to-end Autonomous Driving with Large Vision Language Action Model
  4. Parameter-Efficient Routed Fine-Tuning: Mixture-of-Experts Demands Mixture of Adaptation Modules
  5. Self-Evolving Multi-Agent Systems via Textual Backpropagation
  6. CL-Cross VQA: A Continual Learning Benchmark for Cross-Domain Visual Question Answering
  7. Can Multimodal Large Language Models Truly Perform Multimodal In-Context Learning?
  8. FedBiP: Heterogeneous One-Shot Federated Learning with Personalized Latent Diffusion Models
  9. FedPop: Federated Population-based Hyperparameter Tuning
  10. FocalPO: Enhancing Preference Optimizing by Focusing on Correct Preference Rankings
  11. Improving Perturbation-based Explanations by Understanding the Role of Uncertainty Calibration
  12. Incremental Uncertainty-aware Performance Monitoring with Active Labeling Intervention
  13. LLaVA Steering: Visual Instruction Tuning with 500x Fewer Parameters through Modality Linear Representation-Steering
  14. Localizing Events in Videos with Multimodal Queries
  15. METok: Multi-Stage Event-based Token Compression for Efficient Long Video Understanding
  16. Multimodal Pragmatic Jailbreak on Text-to-image Models
  17. Perceive. Query & Reason: Enhancing Video QA with Question-Guided Temporal Queries
  18. SwarmAgentic: Towards Fully Automated Agentic System Generation via Swarm Intelligence
  19. WebPilot: A Versatile and Autonomous Multi-Agent System for Web Task Execution with Strategic Exploration
  20. Wiki-TabNER: Integrating Named Entity Recognition into Wikipedia Tables
  21. Explanatory Model Monitoring to Understand the Effects of Feature Shifts on Performance
  22. FedDAT: An Approach for Foundation Model Finetuning in Multi-Modal Heterogeneous Federated Learning
  23. GenTKG: Generative Forecasting on Temporal Knowledge Graph with Large Language Models
  24. LookupViT: Compressing Visual Information to a Limited Number of Tokens
  25. Provably Better Explanations with Optimized Aggregation of Feature Attributions
  26. Self-Discovering Interpretable Diffusion Latent Directions for Responsible Text-to-Image Generation
  27. Temporal Fact Reasoning over Hyper-Relational Knowledge Graphs
  28. VideoINSTA: Zero-shot Long Video Understanding via Informative Spatial-Temporal Reasoning with LLMs
  29. Visual Question Decomposition on Multimodal Large Language Models
  30. zrLLM: Zero-Shot Relational Learning on Temporal Knowledge Graphs with Large Language Models