立体声录音
语音识别
积极倾听
计算机科学
可理解性(哲学)
感知
遮罩(插图)
噪音(视频)
言语感知
声音定位
背景噪声
音质
降噪
心理声学
双耳时差
语音增强
语音处理
稳健性(进化)
助听器
人工智能
听觉场景分析
连贯性(哲学赌博策略)
噪声测量
声学
临界带
听觉感知
听觉掩蔽
作者
Reza Ghanavi,Craig Jin
摘要
A weighted masking method based on the coherent-to-diffuse ratio is presented for robust binaural speech enhancement in real-time hearable devices. The method applies manually tuned weights across custom-defined critical frequency bands to improve the quality and intelligibility of frontal target speech in multi-talker reverberant environments. The algorithm was implemented in real time on a functional hearable prototype and evaluated in a perceptual listening study under realistic binaural hearing conditions. Subjective assessments with normal-hearing participants, including evaluations of audio quality, speech intelligibility, and spatial localization, demonstrated consistent improvements compared to baseline coherence-based filtering methods. Results indicate that the method suppresses diffuse background noise while preserving interaural spatial cues important for listening comfort and spatial awareness in complex acoustic scenes. These findings support the applicability of coherence-weighted masking in real-time binaural enhancement tasks under reverberant, multi-talker conditions, including potential use in hearable and hearing aid technologies. In addition to perceptual listening tests, objective evaluations across multiple reverberant environments demonstrate consistent performance improvements over baseline methods.
科研通智能强力驱动
Strongly Powered by AbleSci AI