VOICE AND COUGH AS DIGITAL BIOMARKERS (2025–2026): LINGUISTIC EQUITY, PRIVACY, AND THE PROMISE OF ACOUSTIC DIAGNOSTICS IN GLOBAL HEALTH
DOI:
https://doi.org/10.31435/ijitss.3(51).2026.5732Keywords:
Voice Biomarkers, Cough Analysis, Digital Health, Linguistic Equity, Privacy, Global HealthAbstract
Recent advances in artificial intelligence and mobile sensing have renewed interest in the human voice and cough as non-invasive, low-cost biomarkers of disease. While respiratory, neurological, and psychiatric applications have shown moderate-to-high diagnostic accuracies in classification studies, the translation of these tools from proof-of-concept into routine clinical practice remains incomplete. This narrative review synthesizes the 2025–2026 literature across five domains: (a) the biophysiological foundations that make voice and cough acoustically informative, (b) clinical evidence spanning respiratory, neurodegenerative, mental-health, cardiometabolic, pediatric, and laryngeal use cases, (c) the growing body of work on linguistic equity and structural bias in speech-based systems, (d) privacy and adversarial security concerns specific to voice as biometric data, and (e) deployment prospects in low- and middle-income countries, with tuberculosis screening as the most advanced use case. The analysis reveals that no single disease domain has reached the level of large-scale, independently validated deployment, though respiratory and laryngeal applications are closest. Structural language bias—particularly against speakers of tonal languages—remains a first-order obstacle, and current speaker de-identification systems demonstrably leak identity information. These findings point to three conditions for responsible scale-up: methodological standardization, inclusive multilingual datasets, and governance frameworks that treat voice data as sensitive biometric medical information.
References
Abdelhalim, A. A. I., Osman, H. M. E. M., Sadaka, S. I. H., Mohammed, M. A. Y., Eissa, E., Abdalla, R., & Hassan, D. H. M. (2025). Predictive ability of artificial intelligence algorithms in pediatric respiratory disease diagnosis using cough sounds: A systematic review. Cureus, 17. https://doi.org/10.7759/cureus.88457
Aloradi, A., Gaznepoglu, U. E., Habets, E., & Tenbrinck, D. (2025). VoxATtack: A multimodal attack on voice anonymization systems. In Proceedings of WASPAA 2025. IEEE. https://doi.org/10.1109/WASPAA66052.2025.11230920
Bouguettaya, A., Reeves, K., Tat, P., & Tan, J. Y. (2025). The effect of group singing interventions on health biomarkers in people with stigmatising conditions: A scoping review protocol. Systematic Reviews, 14, Article 171. https://doi.org/10.1186/s13643-025-02971-4
Briganti, G., & Lechien, J. R. (2025). Speech and voice quality as digital biomarkers in depression: A systematic review. Journal of Voice. https://doi.org/10.1016/j.jvoice.2025.05.002
Calvache, C., Soláque, L., Velasco, A., & Peñuela, L. (2021). Biomechanical models to represent vocal physiology: A systematic review. Journal of Voice, 37. https://doi.org/10.1016/j.jvoice.2021.02.014
Cunningham, J. L., Adjagbodjou, A. D., Basoah, J., Jawara, J., Kadoma, K., & Lewis, A. (2025). Toward responsible ASR for African American English speakers: A scoping review of bias and equity in speech technology. arXiv. https://doi.org/10.48550/arXiv.2508.18288
Galvosas, M., & Small, P. M. (2025). The value of continuous cough monitoring: A narrative review. Journal of Thoracic Disease, 17. https://doi.org/10.21037/jtd-2025-876
Guo, C., Chen, L., Li, Z., Lee, K.-A., Ling, Z., & Guo, W. (2024). On the generation and removal of speaker adversarial perturbation for voice-privacy protection. In Proceedings of SLT 2024. IEEE. https://doi.org/10.1109/SLT61566.2024.10832243
Hegde, S., Sreeram, S., Alter, I. L., Shor, C., Valdez, T. A., Meister, K. D., & Rameau, A. (2024). Cough sounds in screening and diagnostics: A scoping review. Laryngoscope, 134(2). https://doi.org/10.1002/lary.31042
Hoekstra, N. E., Chagomerana, M., Smith, Z. M., Kala, A., McLane, I., Verwey, C., Olson, D., Buck, W., Mulindwa, J., Gaudio, A., Kapoor, S., Schuh, H. B., Chiume, M., FitzGerald, E., Elhilali, M., Mvalo, T., Hosseinipour, M. C., & McCollum, E. D. (2025). Performance of an artificial intelligence algorithm for interpreting lung sounds from children hospitalised with pneumonia in Malawi. Journal of Global Health, 15, 04264. https://doi.org/10.7189/jogh.15.04264
Hossain, M. A., Traini, E., & Amenta, F. (2025). Machine learning applications for diagnosing Parkinson’s disease via speech, language, and voice changes: A systematic review. Inventions, 10(4), 48. https://doi.org/10.3390/inventions10040048
Jain, N., Amate, G. A., Alapati, P. R., Bansode, G., Musunoori, S., & B, S. (2025). Real-time accent adaptation in English speech interfaces using federated deep learning across multilingual datasets. In Proceedings of ResgenXAI 2025. IEEE. https://doi.org/10.1109/ResgenXAI64788.2025.11343996
Jayasree, P., Gandhi, K. B., & Sunny, B. (2025). The data privacy protection laws impacting AI medical devices in the US, EU, China, Japan, and India. Asia Pacific Journal of Medical Informatics, 2(3). https://doi.org/10.70818/apjmi.v02i03.030
Jiang, Y., Li, Z., Hu, W., Kong, Y., Wang, X., Zhan, X., Lu, Y., Ye, P., Du, J., He, W., & Tai, J. (2025). Value of novel quantitative acoustic parameters based on mel frequency cepstral coefficients in pediatric vocal cord nodules and laryngopharyngitis. Journal of Voice. https://doi.org/10.1016/j.jvoice.2025.06.017
Kalodanis, K., Feretzakis, G., Rizomiliotis, P., Verykios, V., Papapavlou, C., Koutsikos, I., & Anagnostopoulos, D. (2025). Data governance in healthcare AI: Navigating the EU AI Act’s requirements. Studies in Health Technology and Informatics. https://doi.org/10.3233/SHTI250050
Kan, X., Xiao, Y., Yang, T.-J., Chen, N., & Mathews, R. (2024). Parameter-efficient transfer learning under federated learning for automatic speech recognition. arXiv. https://doi.org/10.48550/arXiv.2408.11873
Kim, J., Yu, J., Kwon, M., & Kim, J. (2025). FairASR: Fair audio contrastive learning for automatic speech recognition. In Proceedings of Interspeech 2025. https://doi.org/10.21437/interspeech.2025-2590
Kostavasili, D., Ganitidis, T., Athanasiou, M., & Nikita, K. S. (2025). Fairness-aware deep learning model for COVID-19 detection from cough audio recordings. In Proceedings of BIBE 2025. IEEE. https://doi.org/10.1109/BIBE66822.2025.00045
Koudounas, A., Pastor, E., Mazzia, V., Giollo, M., Gueudre, T., Reale, E., Cagliero, L., Cumani, S., de Alfaro, L., Baralis, E., & Amberti, D. (2025). Privacy preserving data selection for bias mitigation in speech models. In Proceedings of ACL 2025: Industry Track. https://doi.org/10.18653/v1/2025.acl-industry.52
Kreiman, J. (2023). Why we talk about voices as we do. Journal of the Acoustical Society of America, 153(3). https://doi.org/10.1121/10.0018224
Landry, V., Matschek, J., Pang, R., Munipalle, M., Tan, K., Boruff, J., & Li-Jessen, N. Y. (2025). Audio-based digital biomarkers in diagnosing and managing respiratory diseases: A systematic review and bibliometric analysis. European Respiratory Review, 34. https://doi.org/10.1183/16000617.0246-2024
Latif, L., & Otieno, H. (2025). The colonial legacy of power, profit, and prejudice in global health governance. Modern Africa: Politics, History and Society, 13(2). https://doi.org/10.26806/modafr.v13i2.259
Lee, K. K., Davenport, P., Smith, J., Irwin, R., McGarvey, L. P., Mazzone, S., & Birring, S. (2021). Global physiology and pathophysiology of cough: Part 1. Cough phenomenology: CHEST guideline and expert panel report. CHEST, 159(1), 282–293. https://doi.org/10.1016/j.chest.2020.08.2086
Ma, N., Mirheidari, B., Brown, G. J., Sanjase, N., Maimbolwa, M. M., Chifwamba, S., Muzazu, S. G. Y., Muyoyeta, M., & Kagujje, M. (2025). Deep learning for tuberculosis screening in a high-burden setting using cough analysis and speech foundation models. arXiv. https://doi.org/10.48550/arXiv.2509.09746
MacLean, E., Broger, T., Yerlikaya, S., Fernández-Carballo, B., Pai, M., & Denkinger, C. (2019). A systematic review of biomarkers to detect active tuberculosis. Nature Microbiology, 4, 748–758. https://doi.org/10.1038/s41564-019-0380-2
Mahapatra, I., & Mahapatra, N. R. (2026). Systematic FAIRness assessment of open voice biomarker datasets for mental health and neurodegenerative diseases. In K. Ekštein et al. (Eds.), Text, speech, and dialogue: TSD 2025 (Lecture Notes in Computer Science, Vol. 16029, pp. 356–368). Springer. https://doi.org/10.1007/978-3-032-02548-7_30
Maitra, S. (2025). Voice AI and hermeneutical injustice at the border. In Proceedings of AIES 2025. AAAI/ACM. https://doi.org/10.1609/aies.v8i3.36787
Mootassim-Billah, S., Schoentgen, J., De Bodt, M., Roper, N., Digonnet, A., Le Tensorer, M., Van Nuffelen, G., & Van Gestel, D. (2023). Acoustic analysis of voluntary coughs, throat clearings, and induced reflexive coughs in a healthy population. Dysphagia, 39, 195–209. https://doi.org/10.1007/s00455-023-10574-1
Niemiec, E., Minssen, T., Boland, P., McEntee, P., & Cahill, R. (2025). Legal and regulatory challenges in multi-country data-driven projects developing and validating medical AI systems in the EU. Expert Review of Medical Devices, 22. https://doi.org/10.1080/17434440.2025.2531295
Pande, M., Patange, A. P., Rastogi, S., Yadav, S., Yulduz, U., & Odamova, U. (2025). Cross-domain generative translation of cough audio biomarkers for tuberculosis screening across diverse global populations. Indian Journal of Tuberculosis. https://doi.org/10.1016/j.ijtb.2025.11.008
Pedersen, M. (2025). Minimum standards for voice-related biomarkers in speech and singing. Journal of Clinical Case Reports, Medical Images and Health Sciences, 11(5). https://doi.org/10.55920/JCRMHS.2025.11.001504
Rameau, A., Andreadis, K., German, A., Lachs, M., Rosen, T. E., Pitzrick, M. S., Symes, L., & Klinck, H. (2022). Changes in cough airflow and acoustics after injection laryngoplasty. Laryngoscope, 133(7). https://doi.org/10.1002/lary.30255
Rodrigo, I., & Duñabeitia, J. A. (2025). Listening to the mind: Integrating vocal biomarkers into digital health. Brain Sciences, 15(7), 762. https://doi.org/10.3390/brainsci15070762
Roll, N., & Graham, C. (2025). Scaling conformation bias in automatic speech recognition. In Proceedings of SLaTE 2025. https://doi.org/10.21437/slate.2025-27
Schultz, B., & Vogel, A. (2022). A tutorial review on clinical acoustic markers in speech science. Journal of Speech, Language, and Hearing Research, 65(10). https://doi.org/10.1044/2022_JSLHR-21-00647
Seo, S., Aulov, O., Godil, A., & Mangold, K. (2025). Evaluating identity leakage in speaker de-identification systems. arXiv. https://doi.org/10.48550/arXiv.2508.14012
Storey, E., Harte, N., & Bell, P. (2025). Language bias in self-supervised learning for automatic speech recognition. arXiv. https://doi.org/10.48550/arXiv.2501.19321
Timerbulatov, I. F., Savelieva, E. E., Pestova, R. M., Zagidullina, I. I., & Timerbulatov, R. S. (2025). Assessment of the possibility of acoustic voice analysis. Meditsinskiy Sovet = Medical Council, (7), 185–190. https://doi.org/10.21518/ms2025-029
Wei, S.-L., Liao, Y.-L., Chang, Y.-H., Huang, H.-H., & Chen, H.-H. (2026). Bias in the ear of the listener: Assessing sensitivity in audio language models across linguistic, demographic, and positional variations. arXiv. https://arxiv.org/abs/2602.01030
World Health Organization. (2025). Target product profiles for tuberculosis screening tests (ISBN 9789240113572). WHO Global Programme on Tuberculosis and Lung Health. https://www.who.int/publications/i/item/9789240113572
Yang, M., Liu, X., Du, W., Liu, Y., Zhu, W., Bu, Z., Mao, J., Wang, Q., Chen, S., Zhou, M., & Qu, J.-M. (2026). A device-invariant multi-modal learning framework for respiratory disease classification. npj Digital Medicine. https://doi.org/10.1038/s41746-026-02445-4
Yang, Y., Chen, M., Han, S., Yu, M., Yu, L., Wang, W., Zhang, W., Chen, S., Wang, X., Shan, S., & Wang, Z. (2025). How voice biomarkers have been used in heart failure management: A scoping review. Current Medical Research and Opinion. https://doi.org/10.1080/03007995.2025.2579378
Zhu, Y., Davoust, A., & Falk, T. H. (2025). DeepSick: Deceiving voice-based diagnostic models with synthetic multilingual pathological speech signals. In Proceedings of SMC 2025. IEEE. https://doi.org/10.1109/SMC58881.2025.11343240
Downloads
Published
Issue
Section
License
Copyright (c) 2026 Wiktor Adamiec, Damian Danilczuk, Jerzy Buszko, Michał Dyś, Nicol Baran, Marta Bajkowska-Piterak, Przemysław Piterak, Monika Szlachta-Gubernat, Adrianna Buż

This work is licensed under a Creative Commons Attribution 4.0 International License.
All articles are published in open-access and licensed under a Creative Commons Attribution 4.0 International License (CC BY 4.0). Hence, authors retain copyright to the content of the articles.
CC BY 4.0 License allows content to be copied, adapted, displayed, distributed, re-published or otherwise re-used for any purpose including for adaptation and commercial use provided the content is attributed.

