PPaperPicks

Rima Hazra

11 papers at tracked venues · 3 at CORE A* · active 20242026

Venues

Frequent coauthors

Rachneet Singh Sachdeva
×1

Papers

  1. AURA: Affordance-Understanding and Risk-aware Alignment Technique for Large Language Models
  2. Breaking Boundaries: Investigating the Effects of Model Editing on Cross-linguistic Performance
  3. How (Un)ethical Are Instruction-Centric Responses of LLMs? Unveiling the Vulnerabilities of Safety Guardrails to Harmful Queries
  4. Navigating the Cultural Kaleidoscope: A Hitchhiker's Guide to Sensitivity in Large Language Models
  5. SafeInfer: Context Adaptive Decoding Time Safety Alignment for Large Language Models
  6. Soteria: Language-Specific Functional Parameter Steering for Multilingual Safety Alignment
  7. Turning Logic Against Itself: Probing Model Defenses Through Contrastive Questions
  8. Context Matters: Pushing the Boundaries of Open-Ended Answer Generation with Graph-Structured Knowledge Context
  9. DistALANER: Distantly Supervised Active Learning Augmented Named Entity Recognition in the Open Source Software Ecosystem
  10. Safety Arithmetic: A Framework for Test-time Safety Alignment of Language Models by Steering Parameters and Activations
    EMNLP 2024 · Rima Hazra
  11. Sowing the Wind, Reaping the Whirlwind: The Impact of Editing Language Models
    ACL 2024 · Rima Hazra