PPaperPicks

Rajesh Sharma

Plaksha University, School of AI and Computer Science, Mohali, India

20 papers at tracked venues · 5 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. IndicAG: An Explainable Agentic Framework for Indic-Multilingual Multidimensional Aggression Detection
  2. The Confidence Trap: Gender Bias and Predictive Certainty in LLMs
  3. Crowdsourced Fact-Checking or Biased Commentary? Analyzing Political Bias in Twitter's Community Notes
  4. Decision-making in Urban Trajectories of Bike Users: a Preliminary Study
  5. HYFuse: Aligning Heterogeneous Speech Pre-Trained Representations in Hyperbolic Space for Speech Emotion Recognition
  6. Investigating Prosodic Signatures via Speech Pre-Trained Models for Audio Deepfake Source Attribution
  7. Investigating the Reasonable Effectiveness of Speaker Pre-Trained Models and their Synergistic Power for SingMOS Prediction
  8. PARROT: Synergizing Mamba and Attention-based SSL Pre-Trained Models via Parallel Branch Hadamard Optimal Transport for Speech Emotion Recognition
  9. SNIFR : Boosting Fine-Grained Child Harmful Content Detection Through Audio-Visual Alignment with Cascaded Cross-Transformer
  10. TSGAN: Temporal Social Graph Attention Network for Aggressive Behavior Forecasting
  11. Towards Fusion of Neural Audio Codec-based Representations with Spectral for Heart Murmur Classification via Bandit-based Cross-Attention Mechanism
  12. Towards Machine Unlearning for Paralinguistic Speech Processing
  13. Towards Source Attribution of Singing Voice Deepfake with Multimodal Foundation Models
  14. AVR: synergizing foundation models for audio-visual humor detection
  15. Are Paralinguistic Representations all that is needed for Speech Emotion Recognition?
  16. ComFeAT: combination of neural and spectral features for improved depression detection
  17. Heterogeneity over Homogeneity: Investigating Multilingual Speech Pre-Trained Models for Detecting Audio Deepfake
  18. PERSONA: an application for emotion recognition, gender recognition and age estimation
  19. The reasonable effectiveness of speaker embeddings for violence detection
  20. Towards Multilingual Audio-Visual Question Answering