Publications

Peer-reviewed publications spanning large language model architectures, graph retrieval-augmented generation (GraphRAG), computational audition, multimodal machine learning, and learning from imprecise labels.

Research Topic:
Year:
Venue Tier:
Showing 64 publications
Conference Paper 2026
🧠 LLM

Inference Time Optimization with Confidence Dynamics

International Conference on Machine Learning (ICML 2026)
Paper PDF Details
Conference Paper 2026

Physical AI: The Next Frontier in AI and Robotics to Build Truly Autonomous Machines

Preprints
Paper PDF Details
arXiv 2026
🧠 LLM

Training-Free Agentic AI: Probabilistic Control and Coordination in Multi-Agent LLM Systems

arXiv
Paper PDF Details
Conference Paper 2026
🧠 LLM

Small Language Models: Architecture, Evolution, and the Future of Artificial Intelligence

Preprints
Paper PDF Details
IEEE 2025
🧠 LLM

Improving LLM Retrieval with GraphRAG-SF: Semantic Filtering for Graph-Based RAG Systems

IEEE International Conference on Big Data (IEEE Big Data 2025)
Paper PDF Details
ICLR 2025
🧠 LLM

MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers

The Thirteenth International Conference on Learning Representations (ICLR 2026)
Paper PDF Details
arXiv 2025
πŸ”Š Audio

Deciphering GunType Hierarchy through Acoustic Analysis of Gunshot Recordings

arXiv preprint arXiv:2506.20609
Paper PDF Details
arXiv 2025
🧠 LLM

ProRefine: Inference-time Prompt Refinement with Textual Feedback

arXiv preprint arXiv:2506.05305
Paper PDF Details
ICLR 2025
🧠 LLM

Improving data efficiency via curating LLM-driven rating systems

International Conference on Learning Representations (ICLR 2025)
Paper PDF Details
ICLR 2025
🧠 LLM

LLM Unlearning via Loss Adjustment with Only Forget Data

International Conference on Learning Representations (ICLR 2025)
Paper PDF Details
arXiv 2025
🧠 LLM

Enhancing Retrieval for ESGLLM via ESG-CID--A Disclosure Content Index Finetuning Dataset for Mapping GRI and ESRS

arXiv preprint arXiv:2503.10674
Paper PDF Details
NeurIPS 2024
🏷️ Labels

Imprecise label learning: A unified framework for learning with various imprecise label configurations

Advances in Neural Information Processing Systems 37 (NeurIPS 2024)
Paper PDF Details
CMU Ph.D. Thesis 2024
πŸ”Š Audio 🏷️ Labels

Computational Audition with Imprecise Labels

Carnegie Mellon University PhD Thesis
Details
arXiv 2024
πŸ”Š Audio

Did You Hear That? Introducing AADG: A Framework for Generating Benchmark Data in Audio Anomaly Detection

arXiv preprint arXiv:2410.03904
Paper PDF Details
arXiv 2024

Automatic dataset construction (ADC): Sample collection, data curation, and beyond

arXiv preprint arXiv:2408.11338
Paper PDF Details
arXiv 2024
🧠 LLM

Harnessing business and media insights with large language models

arXiv preprint arXiv:2406.06559
Paper PDF Details
ICLR 2024
🏷️ Labels

Understanding and mitigating the label noise in pre-training on downstream tasks

International Conference on Learning Representations (ICLR 2024)
Paper PDF Details
ICASSP 2024
πŸ”Š Audio 🏷️ Labels

Importance of negative sampling in weak label learning

ICASSP 2024 - IEEE International Conference on Acoustics, Speech and Signal Processing
Paper PDF Details
ICASSP 2024
πŸ”Š Audio πŸŽ₯ Video

Conformer is All You Need for Visual Speech Recognition

ICASSP 2024 - IEEE International Conference on Acoustics, Speech and Signal Processing
Paper PDF Details
ACM 2024
🧠 LLM πŸ”Š Audio πŸŽ₯ Video

Overview of the Tenth Dialog System Technology Challenge: DSTC10

IEEE/ACM Transactions on Audio, Speech, and Language Processing
Paper PDF Details
arXiv 2023
🧠 LLM πŸ”Š Audio πŸŽ₯ Video

Audio-visual fine-tuning of audio-only ASR models

arXiv preprint arXiv:2312.09369
Paper PDF Details
arXiv 2023
πŸ”Š Audio

Psychoacoustic Challenges Of Speech Enhancement On VoIP Platforms

arXiv preprint arXiv:2310.07161
Paper PDF Details
arXiv 2023
🧠 LLM

LoFT: Local proxy fine-tuning for improving transferability of adversarial attacks against large language model

arXiv preprint arXiv:2310.04445
Paper PDF Details
arXiv 2023
πŸ”Š Audio

Online Active Learning For Sound Event Detection

arXiv preprint arXiv:2309.14460
Paper PDF Details
arXiv 2023
πŸ”Š Audio πŸŽ₯ Video

Exploring Domain-Specific Enhancements for a Neural Foley Synthesizer

arXiv preprint arXiv:2309.04641
Paper PDF Details
ICASSP 2023
🧠 LLM πŸ”Š Audio 🏷️ Labels

An approach to ontological learning from weak labels

ICASSP 2023 - IEEE International Conference on Acoustics, Speech and Signal Processing
Paper PDF Details
DCASE 2023
πŸ”Š Audio

DCASE task 7: Foley sound synthesis

DCASE 2023 Challenge Technical Report
Details
arXiv 2023
πŸ”Š Audio

MCCLA: An improved contrastive learning structure for audio representations

arXiv
Details
arXiv 2023
πŸ”Š Audio

Improving Perceptual Quality, Intelligibility, and Acoustics on VoIP Platforms

arXiv preprint arXiv:2303.09048
Paper PDF Details
arXiv 2023
🧠 LLM πŸ”Š Audio

Approach to Learning Generalized Audio Representation Through Batch Embedding Covariance Regularization and Constant-Q Transforms

arXiv preprint arXiv:2303.03591
Paper PDF Details
arXiv 2022
🧠 LLM πŸ”Š Audio

Automated Audio Captioning and Language-Based Audio Retrieval

arXiv preprint arXiv:2207.04156
Paper PDF Details
ICASSP 2022
🧠 LLM πŸ”Š Audio πŸŽ₯ Video

Audio-Visual Scene-Aware Dialog and Reasoning using Audio-Visual Transformers with Joint Student-Teacher Learning

2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
Paper PDF Details
arXiv 2022
🧠 LLM πŸ”Š Audio

On the pragmatism of using binary classifiers over data intensive neural network classifiers for detection of COVID-19 from voice

arXiv preprint arXiv:2204.04802
Paper PDF Details
arXiv 2022
πŸ”Š Audio 🏷️ Labels

Ontological Learning from Weak Labels

arXiv preprint arXiv:2203.02483
Paper PDF Details
Conference Paper 2022
🧠 LLM πŸ”Š Audio πŸŽ₯ Video

DSTC10-AVSD Submission System with Reasoning using Audio-Visual Transformers with Joint Student-Teacher Learning

Proceedings of DSTC10 Workshop at AAAI-2022
Paper PDF Details
Conference Paper 2022
🧠 LLM πŸ”Š Audio πŸŽ₯ Video

Overview of Audio Visual Scene-Aware Dialog with Reasoning Track for Natural Language Generation in DSTC10

Proceedings of DSTC10 Workshop at AAAI-2022
Paper PDF Details
arXiv 2021
🧠 LLM πŸŽ₯ Video

Triple Attention Network architecture for MovieQA

arXiv preprint arXiv:2111.09531
Paper PDF Details
Conference Paper 2021
🧠 LLM πŸ”Š Audio πŸŽ₯ Video

Reasoning for Audio Visual Scene-Aware Dialog Track in DSTC10

Dialog System Technology Challenge 10 (DSTC10)
Details
arXiv 2021
πŸ”Š Audio

An overview of techniques for biomarker discovery in voice signal

arXiv preprint arXiv:2110.04678
Paper PDF Details
arXiv 2021

Feature extraction and evaluation for BioMedical Question Answering

arXiv preprint arXiv:2105.14013
Paper PDF Details
arXiv 2021
🧠 LLM 🏷️ Labels

Training image classifiers using Semi-Weak Label Data

arXiv preprint arXiv:2103.10608
Paper PDF Details
ICASSP 2020
πŸ”Š Audio

Sound event detection in synthetic domestic environments

IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
Paper PDF Details
Conference Paper 2020
πŸ”Š Audio

CMU Sounds for COVID Project

CMU Language Technologies Institute
Details
DCASE 2019
πŸ”Š Audio

Sound event detection in domestic environments with weakly labeled data and soundscape synthesis

Detection and Classification of Acoustic Scenes and Events 2019 Workshop (DCASE2019)
Paper PDF Details
ACM 2019
πŸ”Š Audio πŸŽ₯ Video

Multimodal Behavioral Markers Exploring Suicidal Intent in Social Media Videos

21st ACM International Conference on Multimodal Interaction (ICMI)
Paper PDF Details
IJCAI 2019
πŸ”Š Audio 🏷️ Labels

Learning Sound Events From Webly Labeled Data

28th International Joint Conference on Artificial Intelligence (IJCAI)
Paper PDF Details
Conference Paper 2018
🧠 LLM

Tartan: A retrieval-based socialbot powered by a dynamic finite-state machine architecture

2nd Proceedings of Alexa Prize (Alexa Prize 2018)
Paper PDF Details
DCASE 2018
πŸ”Š Audio πŸŽ₯ Video

Large-Scale Weakly Labeled Semi-Supervised Sound Event Detection in Domestic Environments

Detection and Classification of Acoustic Scenes and Events 2018 Workshop (DCASE2018)
Paper PDF Details
IEEE 2018
🧠 LLM πŸŽ₯ Video

CADP: A Novel Dataset for CCTV Traffic Camera based Accident Analysis

IEEE International Workshop on Traffic and Street Surveillance for Safety and Security (AVSS)
Paper PDF Details
arXiv 2018
πŸŽ₯ Video

Activity Recognition on a Large Scale in Short Videos - Moments in Time Dataset

arXiv preprint arXiv:1809.00241
Paper PDF Details
arXiv 2018
🧠 LLM πŸŽ₯ Video

Natural Language Person Search Using Deep Reinforcement Learning

arXiv preprint arXiv:1809.00365
Paper PDF Details
arXiv 2018
πŸ”Š Audio 🏷️ Labels

A Closer Look at Weak Label Learning for Audio Events

arXiv preprint arXiv:1804.09288
Paper PDF Details
ICASSP 2018
πŸ”Š Audio πŸŽ₯ Video

Framework for evaluation of sound event detection in web videos

IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
Paper PDF Details
ICASSP 2018
πŸ”Š Audio

Content-based Representations of audio using Siamese neural networks

IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
Paper PDF Details
NeurIPS 2017
πŸ”Š Audio πŸŽ₯ Video

NELS - Never-Ending Learner of Sounds

Machine Learning for Audio Signal Processing Workshop, NIPS 2017
Paper PDF Details
DCASE 2017
πŸ”Š Audio

DCASE 2017 challenge setup: Tasks, datasets and baseline system

Detection and Classification of Acoustic Scenes and Events 2017 Workshop (DCASE2017)
Paper PDF Details
Conference Paper 2017
πŸ”Š Audio πŸŽ₯ Video

An Approach for Self Training Audio Event Detectors using Web Data

25th European Signal Processing Conference (EUSIPCO)
Paper PDF Details
Conference Paper 2017
πŸ”Š Audio

A Framework towards Large Scale Learning of Sound Events

Language Technologies Institute Student Research Symposium 2017
Details
arXiv 2016
🧠 LLM πŸ”Š Audio 🏷️ Labels

An approach for self-training audio event detectors using web data

arXiv preprint arXiv:1609.06026
Paper PDF Details
IEEE 2016
πŸ”Š Audio

Experiments on the DCASE Challenge 2016: Acoustic Scene Classification and Sound Event Detection in Real Life Recording

IEEE AASP Challenge: Detection and Classification of Acoustic Scenes and Events
Paper PDF Details
DCASE 2016
πŸ”Š Audio

DCASE challenge task 1

Detection and Classification of Acoustic Scenes and Events (DCASE) 2016
Details
Conference Paper 2016

Repeatability and Scalability of Code at Top level Verification

Regional Engineering Conference 2016
Details
Conference Paper 2015
πŸ”Š Audio

Pipelined implementation of high radix adaptive CORDIC as a coprocessor

2015 International Conference on Computing and Network Communications (CoCoNet)
Paper PDF Details
Conference Paper 2015
πŸ”Š Audio

Hardware Architecture for High Radix Adaptive CORDIC Algorithm

National Institute of Technology Karnataka Surathkal
Details