PPaperPicks

Tom Goldstein

University of Maryland, Department of Computer Science, College Park, MD, USA

33 papers at tracked venues · 28 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. A Technical Report on "Erasing the Invisible": The 2024 NeurIPS Competition on Stress Testing Image Watermarks
  2. ARGUS: Hallucination and Omission Evaluation in Video-LLMs
  3. Can Watermarking Large Language Models Prevent Copyrighted Text Generation and Hide Training Data?
    AAAI 2025 ·
    Michael-Andrei Panaitescu-Liess
  4. Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
  5. Efficient Fine-Tuning and Concept Suppression for Pruned Diffusion Models
  6. Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement
  7. FineGRAIN: Evaluating Failure Modes of Text-to-Image Models with Vision Language Model Judges
  8. Gemstones: A Model Suite for Multi-Faceted Scaling Laws
  9. LLM-Generated Passphrases That Are Secure and Easy to Remember
  10. LiveBench: A Challenging, Contamination-Limited LLM Benchmark
  11. Multimodal Agentic Model Predictive Control
    AAMAS 2025 ·
    Saptarashmi Bandyopadhyay
  12. PUP 3D-GS: Principled Uncertainty Pruning for 3D Gaussian Splatting
  13. Quantifying Cross-Modality Memorization in Vision-Language Models
  14. Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
  15. Speedy-Splat: Fast 3D Gaussian Splatting with Sparse Pixels and Sparse Primitives
  16. The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text
  17. Zero-Shot Vision Encoder Grafting via LLM Surrogates
  18. Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
  19. CALVIN: Improved Contextual Video Captioning via Instruction Tuning
  20. Easy2Hard-Bench: Standardized Difficulty Labels for Profiling LLM Performance and Generalization
  21. Hierarchical Point Attention for Indoor 3D Object Detection
  22. InstructZero: Efficient Instruction Optimization for Black-Box Large Language Models
  23. Investigating Style Similarity in Diffusion Models
  24. NEFTune: Noisy Embeddings Improve Instruction Finetuning
  25. ODIN: Disentangled Reward Mitigates Hacking in RLHF
  26. Object Recognition as Next Token Prediction
  27. On the Reliability of Watermarks for Large Language Models
  28. Privacy Backdoors: Enhancing Membership Inference through Poisoning Pre-trained Models
  29. Shadowcast: Stealthy Data Poisoning Attacks Against Vision-Language Models
  30. Spotting LLMs With Binoculars: Zero-Shot Detection of Machine-Generated Text
  31. Transformers Can Do Arithmetic with the Right Embeddings
  32. Universal Guidance for Diffusion Models
  33. WAVES: Benchmarking the Robustness of Image Watermarks