JASA express letters最新文献

筛选
英文 中文
Leveraging sound speed dynamics and generative deep learning for ray-based ocean acoustic tomography.
IF 1.2
JASA express letters Pub Date : 2025-04-01 DOI: 10.1121/10.0036312
Priyabrata Saha, Richard X Touret, Etienne Ollivier, Jihui Jin, Matthew McKinley, Justin Romberg, Karim G Sabra
{"title":"Leveraging sound speed dynamics and generative deep learning for ray-based ocean acoustic tomography.","authors":"Priyabrata Saha, Richard X Touret, Etienne Ollivier, Jihui Jin, Matthew McKinley, Justin Romberg, Karim G Sabra","doi":"10.1121/10.0036312","DOIUrl":"https://doi.org/10.1121/10.0036312","url":null,"abstract":"<p><p>A generative deep learning framework is introduced for ray-based ocean acoustic tomography (OAT), an inverse problem for estimating sound speed profiles (SSP) based on arrival-times measurements between multiple acoustic transducers, which is typically ill-posed. This framework relies on a robust low-dimensional parametrization of the expected SSP variations using a variational autoencoder and a linear dynamical model as further regularization. This framework was tested using SSP variations simulated by a regional ocean model with submesoscale permitting horizontal resolution and various transducer configurations spanning the upper ocean over short propagation ranges and was found to outperform conventional linear least squares formulations of OAT.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 4","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-04-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143756277","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
Bayesian matched-field inversion for shear and compressional geoacoustic profiles at the New England Mud Patcha).
IF 1.2
JASA express letters Pub Date : 2025-04-01 DOI: 10.1121/10.0036374
Stan E Dosso, Preston S Wilson, David P Knobles, Julien Bonnel
{"title":"Bayesian matched-field inversion for shear and compressional geoacoustic profiles at the New England Mud Patcha).","authors":"Stan E Dosso, Preston S Wilson, David P Knobles, Julien Bonnel","doi":"10.1121/10.0036374","DOIUrl":"https://doi.org/10.1121/10.0036374","url":null,"abstract":"<p><p>This Letter estimates shear and compressional seabed geoacoustic profiles at the New England Mud Patch through trans-dimensional Bayesian inversion of matched-field acoustic data over a 20-2000 Hz bandwidth. Results indicate low shear-wave speeds (∼35 m/s) with relatively small uncertainties over most of the upper mud layer, increasing in underlying transition and sand layers. Compressional parameters, including attenuation, are also well estimated, but shear-wave attenuation is poorly determined. Comparison of inversions with/without shear parameters and consideration of inter-parameter correlations indicate that estimates of compressional parameters are not substantially influenced by shear effects, with the possible exception of compressional-wave attenuation in the sand layer.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 4","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-04-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143775232","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
Enhancing speech intelligibility in optical microphone systems through physics-informed data augmentation.
IF 1.2
JASA express letters Pub Date : 2025-04-01 DOI: 10.1121/10.0036356
Jia-Wei Chen, Jia-Hui Li, Yi-Hao Jiang, Yi-Chang Wu, Ying-Hui Lai
{"title":"Enhancing speech intelligibility in optical microphone systems through physics-informed data augmentation.","authors":"Jia-Wei Chen, Jia-Hui Li, Yi-Hao Jiang, Yi-Chang Wu, Ying-Hui Lai","doi":"10.1121/10.0036356","DOIUrl":"10.1121/10.0036356","url":null,"abstract":"<p><p>Laser doppler vibrometers (LDVs) facilitate noncontact speech acquisition; however, they are prone to material-dependent spectral distortions and speckle noise, which degrade intelligibility in noisy environments. This study proposes a data augmentation method that incorporates material-specific and impulse noises to simulate LDV-induced distortions. The proposed approach utilizes a gated convolutional neural network with HiFi-GAN to enhance speech intelligibility across various material and low signal-to-noise ratio (SNR) conditions, achieving a short-time objective intelligibility score of 0.76 at 0 dB SNR. These findings provide valuable insights into optimized augmentation and deep-learning techniques for enhancing LDV-based speech recordings in practical applications.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 4","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-04-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143766092","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
Perception-production link mediated by position in the imitation of Korean nasal stops.
IF 1.2
JASA express letters Pub Date : 2025-03-01 DOI: 10.1121/10.0036057
Jiwon Hwang, Yu-An Lu
{"title":"Perception-production link mediated by position in the imitation of Korean nasal stops.","authors":"Jiwon Hwang, Yu-An Lu","doi":"10.1121/10.0036057","DOIUrl":"10.1121/10.0036057","url":null,"abstract":"<p><p>This study explores how perceptual cues in two positions influence imitation of Korean nasal stops. As a result of initial denasalization, nasality cues are secondary in the initial position but primary in the medial position. Categorization and imitation tasks using CV (consonant-vowel) and VCV (vowel-consonant-vowel) items on a continuum from voiced oral to nasal stops were completed by 32 Korean speakers. Results revealed categorical imitation of nasality medially, whereas imitation was gradient or minimal initially. Furthermore, individuals requiring stronger nasality cues to categorize a nasal sound produced greater nasality in imitation. These findings highlight a perception-production link mediated by positional cue reliance.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 3","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-03-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143544861","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
Spatial grouping as a method to improve personalized head-related transfer function prediction.
IF 1.2
JASA express letters Pub Date : 2025-03-01 DOI: 10.1121/10.0036032
Keng-Wei Chang, Yih-Liang Shen, Tai-Shih Chi
{"title":"Spatial grouping as a method to improve personalized head-related transfer function prediction.","authors":"Keng-Wei Chang, Yih-Liang Shen, Tai-Shih Chi","doi":"10.1121/10.0036032","DOIUrl":"10.1121/10.0036032","url":null,"abstract":"<p><p>The head-related transfer function (HRTF) characterizes the frequency response of the sound traveling path between a specific location and the ear. When it comes to estimating HRTFs by neural network models, angle-specific models greatly outperform global models but demand high computational resources. To balance the computational resource and performance, we propose a method by grouping HRTF data spatially to reduce variance within each subspace. HRTF predicting neural network is then trained for each subspace. Results show the proposed method performs better than global models and angle-specific models by using different grouping strategies at the ipsilateral and contralateral sides.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 3","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-03-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143544862","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
Variation in the production of nasal coarticulation by speaker age and speech style.
IF 1.2
JASA express letters Pub Date : 2025-03-01 DOI: 10.1121/10.0036227
Georgia Zellou, Michelle Cohn
{"title":"Variation in the production of nasal coarticulation by speaker age and speech style.","authors":"Georgia Zellou, Michelle Cohn","doi":"10.1121/10.0036227","DOIUrl":"10.1121/10.0036227","url":null,"abstract":"<p><p>This study investigates apparent-time variation in the production of anticipatory nasal coarticulation in California English. Productions of consonant-vowel-nasal words in clear vs casual speech by 58 speakers aged 18-58 (grouped into three generations) were analyzed for degree of coarticulatory vowel nasality. Results reveal an interaction between age and style: the two younger speaker groups produce greater coarticulation (measured as A1-P0) in clear speech, whereas older speakers produce less variable coarticulation across styles. Yet, duration lengthening in clear speech is stable across ages. Thus, age- and style-conditioned changes in produced coarticulation interact as part of change in coarticulation grammars over time.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 3","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-03-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143674970","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
Voice assistant technology continues to underperform on children's speech.
IF 1.2
JASA express letters Pub Date : 2025-03-01 DOI: 10.1121/10.0036052
Holly Bradley, Madeleine E Yu, Elizabeth K Johnson
{"title":"Voice assistant technology continues to underperform on children's speech.","authors":"Holly Bradley, Madeleine E Yu, Elizabeth K Johnson","doi":"10.1121/10.0036052","DOIUrl":"10.1121/10.0036052","url":null,"abstract":"<p><p>Voice assistant (VA) technology is increasingly part of children's everyday lives. But how well do these systems understand children? No study has asked this with children under 5 years old. Here, two versions of Siri, and one of Alexa, were tested on their ability to transcribe utterances produced by 2-, 3-, and 5-year-olds. Human listeners (mothers and undergraduates) were also tested. Results showed that while Siri's performance on children's speech has improved in recent years, even the newest Siri and Alexa models struggle with children's speech. Human listeners far outperformed VA systems with all ages, especially with the youngest children's speech.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 3","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-03-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143544863","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
Bone mineral density and hydroxyapatite alignment in leg cortical bone influence on ultrasound velocity.
IF 1.2
JASA express letters Pub Date : 2025-03-01 DOI: 10.1121/10.0036082
Shuta Kodama, Hiroshi Mita, Norihisa Tamura, Daisuke Koyama, Mami Matsukawa
{"title":"Bone mineral density and hydroxyapatite alignment in leg cortical bone influence on ultrasound velocity.","authors":"Shuta Kodama, Hiroshi Mita, Norihisa Tamura, Daisuke Koyama, Mami Matsukawa","doi":"10.1121/10.0036082","DOIUrl":"10.1121/10.0036082","url":null,"abstract":"<p><p>Bone diagnosis using x-ray techniques, such as computed tomography and dual-energy x-ray absorptiometry, can evaluate bone mineral density (BMD) and microstructure but does not provide elastic properties. This study investigated the ultrasonic properties of racehorse leg cortical bone, focusing on the relationship between wave velocity, BMD, and hydroxyapatite (HAp) crystallite alignment. The results showed a strong correlation between wave velocity and BMD, suggesting that quantitative ultrasound-obtained wave velocity is primarily influenced by BMD, followed by the HAp alignment direction.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 3","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-03-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143588552","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
A method of reference phase velocity selecting for bearing estimation with a horizontal line array in shallow water.
IF 1.2
JASA express letters Pub Date : 2025-03-01 DOI: 10.1121/10.0035934
Dai Liu, Feilong Zhu, Yanjun Zhang, Zhaohui Peng
{"title":"A method of reference phase velocity selecting for bearing estimation with a horizontal line array in shallow water.","authors":"Dai Liu, Feilong Zhu, Yanjun Zhang, Zhaohui Peng","doi":"10.1121/10.0035934","DOIUrl":"https://doi.org/10.1121/10.0035934","url":null,"abstract":"<p><p>In shallow water environments, choosing an appropriate reference phase velocity for direction-of-arrival estimation with a beamformed underwater horizontal line array is very important. The direction of the maximum beamformer output power will deviate from the true source bearing when a mismatched reference phase velocity was used. This Letter analyzed the intrinsic relationship between the reference phase velocity and normal mode amplitude distribution, source bearing, array aperture, and then proposed a multi-parameter weighted reference phase velocity selection method, which has improved the accuracy of source bearing estimation. Numerical simulation and experimental results validated the effectiveness of this method.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 3","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-03-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143607363","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
Effects of spatial asymmetry and voice-gender differences between talkers on spatial release from masking in normal-hearing listeners.
IF 1.2
JASA express letters Pub Date : 2025-03-01 DOI: 10.1121/10.0036249
Yonghee Oh, Josephine Kinder, Phillip Friggle, Caroline Cuthbertson
{"title":"Effects of spatial asymmetry and voice-gender differences between talkers on spatial release from masking in normal-hearing listeners.","authors":"Yonghee Oh, Josephine Kinder, Phillip Friggle, Caroline Cuthbertson","doi":"10.1121/10.0036249","DOIUrl":"10.1121/10.0036249","url":null,"abstract":"<p><p>This study investigated how a listener's spatial release from masking (SRM) performance is affected by spatial asymmetry and voice-gender differences between talkers in multi-talker listening situations. The amounts of SRM were measured with symmetric and asymmetric (toward the right or left) masker configurations in same-gender and different-gender target-masker conditions. The results showed that the SRM was co-varied by talkers' voice-gender differences and spatial asymmetry cues: maximized in the same-gender and asymmetrical target-maskers condition and minimized in the different-gender and symmetrical target-maskers condition. Those findings suggest that the talkers' asymmetry and voice-gender differences could contribute to the variation in SRM independently.</p>","PeriodicalId":73538,"journal":{"name":"JASA express letters","volume":"5 3","pages":""},"PeriodicalIF":1.2,"publicationDate":"2025-03-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":null,"resultStr":null,"platform":"Semanticscholar","paperid":"143660003","PeriodicalName":null,"FirstCategoryId":null,"ListUrlMain":null,"RegionNum":0,"RegionCategory":"","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":"","EPubDate":null,"PubModel":null,"JCR":null,"JCRName":null,"Score":null,"Total":0}
引用次数: 0
0
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
确定
请完成安全验证×
相关产品
×
本文献相关产品
联系我们:info@booksci.cn Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。 Copyright © 2023 布克学术 All rights reserved.
京ICP备2023020795号-1
ghs 京公网安备 11010802042870号
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术官方微信