VOICE AND COUGH AS DIGITAL BIOMARKERS (2025–2026): LINGUISTIC EQUITY, PRIVACY, AND THE PROMISE OF ACOUSTIC DIAGNOSTICS IN GLOBAL HEALTH

Authors

DOI:

https://doi.org/10.31435/ijitss.3(51).2026.5732

Keywords:

Voice Biomarkers, Cough Analysis, Digital Health, Linguistic Equity, Privacy, Global Health

Abstract

Recent advances in artificial intelligence and mobile sensing have renewed interest in the human voice and cough as non-invasive, low-cost biomarkers of disease. While respiratory, neurological, and psychiatric applications have shown moderate-to-high diagnostic accuracies in classification studies, the translation of these tools from proof-of-concept into routine clinical practice remains incomplete. This narrative review synthesizes the 2025–2026 literature across five domains: (a) the biophysiological foundations that make voice and cough acoustically informative, (b) clinical evidence spanning respiratory, neurodegenerative, mental-health, cardiometabolic, pediatric, and laryngeal use cases, (c) the growing body of work on linguistic equity and structural bias in speech-based systems, (d) privacy and adversarial security concerns specific to voice as biometric data, and (e) deployment prospects in low- and middle-income countries, with tuberculosis screening as the most advanced use case. The analysis reveals that no single disease domain has reached the level of large-scale, independently validated deployment, though respiratory and laryngeal applications are closest. Structural language bias—particularly against speakers of tonal languages—remains a first-order obstacle, and current speaker de-identification systems demonstrably leak identity information. These findings point to three conditions for responsible scale-up: methodological standardization, inclusive multilingual datasets, and governance frameworks that treat voice data as sensitive biometric medical information.

References

Abdelhalim, A. A. I., Osman, H. M. E. M., Sadaka, S. I. H., Mohammed, M. A. Y., Eissa, E., Abdalla, R., & Hassan, D. H. M. (2025). Predictive ability of artificial intelligence algorithms in pediatric respiratory disease diagnosis using cough sounds: A systematic review. Cureus, 17. https://doi.org/10.7759/cureus.88457

Aloradi, A., Gaznepoglu, U. E., Habets, E., & Tenbrinck, D. (2025). VoxATtack: A multimodal attack on voice anonymization systems. In Proceedings of WASPAA 2025. IEEE. https://doi.org/10.1109/WASPAA66052.2025.11230920

Bouguettaya, A., Reeves, K., Tat, P., & Tan, J. Y. (2025). The effect of group singing interventions on health biomarkers in people with stigmatising conditions: A scoping review protocol. Systematic Reviews, 14, Article 171. https://doi.org/10.1186/s13643-025-02971-4

Briganti, G., & Lechien, J. R. (2025). Speech and voice quality as digital biomarkers in depression: A systematic review. Journal of Voice. https://doi.org/10.1016/j.jvoice.2025.05.002

Calvache, C., Soláque, L., Velasco, A., & Peñuela, L. (2021). Biomechanical models to represent vocal physiology: A systematic review. Journal of Voice, 37. https://doi.org/10.1016/j.jvoice.2021.02.014

Cunningham, J. L., Adjagbodjou, A. D., Basoah, J., Jawara, J., Kadoma, K., & Lewis, A. (2025). Toward responsible ASR for African American English speakers: A scoping review of bias and equity in speech technology. arXiv. https://doi.org/10.48550/arXiv.2508.18288

Galvosas, M., & Small, P. M. (2025). The value of continuous cough monitoring: A narrative review. Journal of Thoracic Disease, 17. https://doi.org/10.21037/jtd-2025-876

Guo, C., Chen, L., Li, Z., Lee, K.-A., Ling, Z., & Guo, W. (2024). On the generation and removal of speaker adversarial perturbation for voice-privacy protection. In Proceedings of SLT 2024. IEEE. https://doi.org/10.1109/SLT61566.2024.10832243

Hegde, S., Sreeram, S., Alter, I. L., Shor, C., Valdez, T. A., Meister, K. D., & Rameau, A. (2024). Cough sounds in screening and diagnostics: A scoping review. Laryngoscope, 134(2). https://doi.org/10.1002/lary.31042

Hoekstra, N. E., Chagomerana, M., Smith, Z. M., Kala, A., McLane, I., Verwey, C., Olson, D., Buck, W., Mulindwa, J., Gaudio, A., Kapoor, S., Schuh, H. B., Chiume, M., FitzGerald, E., Elhilali, M., Mvalo, T., Hosseinipour, M. C., & McCollum, E. D. (2025). Performance of an artificial intelligence algorithm for interpreting lung sounds from children hospitalised with pneumonia in Malawi. Journal of Global Health, 15, 04264. https://doi.org/10.7189/jogh.15.04264

Hossain, M. A., Traini, E., & Amenta, F. (2025). Machine learning applications for diagnosing Parkinson’s disease via speech, language, and voice changes: A systematic review. Inventions, 10(4), 48. https://doi.org/10.3390/inventions10040048

Jain, N., Amate, G. A., Alapati, P. R., Bansode, G., Musunoori, S., & B, S. (2025). Real-time accent adaptation in English speech interfaces using federated deep learning across multilingual datasets. In Proceedings of ResgenXAI 2025. IEEE. https://doi.org/10.1109/ResgenXAI64788.2025.11343996

Jayasree, P., Gandhi, K. B., & Sunny, B. (2025). The data privacy protection laws impacting AI medical devices in the US, EU, China, Japan, and India. Asia Pacific Journal of Medical Informatics, 2(3). https://doi.org/10.70818/apjmi.v02i03.030

Jiang, Y., Li, Z., Hu, W., Kong, Y., Wang, X., Zhan, X., Lu, Y., Ye, P., Du, J., He, W., & Tai, J. (2025). Value of novel quantitative acoustic parameters based on mel frequency cepstral coefficients in pediatric vocal cord nodules and laryngopharyngitis. Journal of Voice. https://doi.org/10.1016/j.jvoice.2025.06.017

Kalodanis, K., Feretzakis, G., Rizomiliotis, P., Verykios, V., Papapavlou, C., Koutsikos, I., & Anagnostopoulos, D. (2025). Data governance in healthcare AI: Navigating the EU AI Act’s requirements. Studies in Health Technology and Informatics. https://doi.org/10.3233/SHTI250050

Kan, X., Xiao, Y., Yang, T.-J., Chen, N., & Mathews, R. (2024). Parameter-efficient transfer learning under federated learning for automatic speech recognition. arXiv. https://doi.org/10.48550/arXiv.2408.11873

Kim, J., Yu, J., Kwon, M., & Kim, J. (2025). FairASR: Fair audio contrastive learning for automatic speech recognition. In Proceedings of Interspeech 2025. https://doi.org/10.21437/interspeech.2025-2590

Kostavasili, D., Ganitidis, T., Athanasiou, M., & Nikita, K. S. (2025). Fairness-aware deep learning model for COVID-19 detection from cough audio recordings. In Proceedings of BIBE 2025. IEEE. https://doi.org/10.1109/BIBE66822.2025.00045

Koudounas, A., Pastor, E., Mazzia, V., Giollo, M., Gueudre, T., Reale, E., Cagliero, L., Cumani, S., de Alfaro, L., Baralis, E., & Amberti, D. (2025). Privacy preserving data selection for bias mitigation in speech models. In Proceedings of ACL 2025: Industry Track. https://doi.org/10.18653/v1/2025.acl-industry.52

Kreiman, J. (2023). Why we talk about voices as we do. Journal of the Acoustical Society of America, 153(3). https://doi.org/10.1121/10.0018224

Landry, V., Matschek, J., Pang, R., Munipalle, M., Tan, K., Boruff, J., & Li-Jessen, N. Y. (2025). Audio-based digital biomarkers in diagnosing and managing respiratory diseases: A systematic review and bibliometric analysis. European Respiratory Review, 34. https://doi.org/10.1183/16000617.0246-2024

Latif, L., & Otieno, H. (2025). The colonial legacy of power, profit, and prejudice in global health governance. Modern Africa: Politics, History and Society, 13(2). https://doi.org/10.26806/modafr.v13i2.259

Lee, K. K., Davenport, P., Smith, J., Irwin, R., McGarvey, L. P., Mazzone, S., & Birring, S. (2021). Global physiology and pathophysiology of cough: Part 1. Cough phenomenology: CHEST guideline and expert panel report. CHEST, 159(1), 282–293. https://doi.org/10.1016/j.chest.2020.08.2086

Ma, N., Mirheidari, B., Brown, G. J., Sanjase, N., Maimbolwa, M. M., Chifwamba, S., Muzazu, S. G. Y., Muyoyeta, M., & Kagujje, M. (2025). Deep learning for tuberculosis screening in a high-burden setting using cough analysis and speech foundation models. arXiv. https://doi.org/10.48550/arXiv.2509.09746

MacLean, E., Broger, T., Yerlikaya, S., Fernández-Carballo, B., Pai, M., & Denkinger, C. (2019). A systematic review of biomarkers to detect active tuberculosis. Nature Microbiology, 4, 748–758. https://doi.org/10.1038/s41564-019-0380-2

Mahapatra, I., & Mahapatra, N. R. (2026). Systematic FAIRness assessment of open voice biomarker datasets for mental health and neurodegenerative diseases. In K. Ekštein et al. (Eds.), Text, speech, and dialogue: TSD 2025 (Lecture Notes in Computer Science, Vol. 16029, pp. 356–368). Springer. https://doi.org/10.1007/978-3-032-02548-7_30

Maitra, S. (2025). Voice AI and hermeneutical injustice at the border. In Proceedings of AIES 2025. AAAI/ACM. https://doi.org/10.1609/aies.v8i3.36787

Mootassim-Billah, S., Schoentgen, J., De Bodt, M., Roper, N., Digonnet, A., Le Tensorer, M., Van Nuffelen, G., & Van Gestel, D. (2023). Acoustic analysis of voluntary coughs, throat clearings, and induced reflexive coughs in a healthy population. Dysphagia, 39, 195–209. https://doi.org/10.1007/s00455-023-10574-1

Niemiec, E., Minssen, T., Boland, P., McEntee, P., & Cahill, R. (2025). Legal and regulatory challenges in multi-country data-driven projects developing and validating medical AI systems in the EU. Expert Review of Medical Devices, 22. https://doi.org/10.1080/17434440.2025.2531295

Pande, M., Patange, A. P., Rastogi, S., Yadav, S., Yulduz, U., & Odamova, U. (2025). Cross-domain generative translation of cough audio biomarkers for tuberculosis screening across diverse global populations. Indian Journal of Tuberculosis. https://doi.org/10.1016/j.ijtb.2025.11.008

Pedersen, M. (2025). Minimum standards for voice-related biomarkers in speech and singing. Journal of Clinical Case Reports, Medical Images and Health Sciences, 11(5). https://doi.org/10.55920/JCRMHS.2025.11.001504

Rameau, A., Andreadis, K., German, A., Lachs, M., Rosen, T. E., Pitzrick, M. S., Symes, L., & Klinck, H. (2022). Changes in cough airflow and acoustics after injection laryngoplasty. Laryngoscope, 133(7). https://doi.org/10.1002/lary.30255

Rodrigo, I., & Duñabeitia, J. A. (2025). Listening to the mind: Integrating vocal biomarkers into digital health. Brain Sciences, 15(7), 762. https://doi.org/10.3390/brainsci15070762

Roll, N., & Graham, C. (2025). Scaling conformation bias in automatic speech recognition. In Proceedings of SLaTE 2025. https://doi.org/10.21437/slate.2025-27

Schultz, B., & Vogel, A. (2022). A tutorial review on clinical acoustic markers in speech science. Journal of Speech, Language, and Hearing Research, 65(10). https://doi.org/10.1044/2022_JSLHR-21-00647

Seo, S., Aulov, O., Godil, A., & Mangold, K. (2025). Evaluating identity leakage in speaker de-identification systems. arXiv. https://doi.org/10.48550/arXiv.2508.14012

Storey, E., Harte, N., & Bell, P. (2025). Language bias in self-supervised learning for automatic speech recognition. arXiv. https://doi.org/10.48550/arXiv.2501.19321

Timerbulatov, I. F., Savelieva, E. E., Pestova, R. M., Zagidullina, I. I., & Timerbulatov, R. S. (2025). Assessment of the possibility of acoustic voice analysis. Meditsinskiy Sovet = Medical Council, (7), 185–190. https://doi.org/10.21518/ms2025-029

Wei, S.-L., Liao, Y.-L., Chang, Y.-H., Huang, H.-H., & Chen, H.-H. (2026). Bias in the ear of the listener: Assessing sensitivity in audio language models across linguistic, demographic, and positional variations. arXiv. https://arxiv.org/abs/2602.01030

World Health Organization. (2025). Target product profiles for tuberculosis screening tests (ISBN 9789240113572). WHO Global Programme on Tuberculosis and Lung Health. https://www.who.int/publications/i/item/9789240113572

Yang, M., Liu, X., Du, W., Liu, Y., Zhu, W., Bu, Z., Mao, J., Wang, Q., Chen, S., Zhou, M., & Qu, J.-M. (2026). A device-invariant multi-modal learning framework for respiratory disease classification. npj Digital Medicine. https://doi.org/10.1038/s41746-026-02445-4

Yang, Y., Chen, M., Han, S., Yu, M., Yu, L., Wang, W., Zhang, W., Chen, S., Wang, X., Shan, S., & Wang, Z. (2025). How voice biomarkers have been used in heart failure management: A scoping review. Current Medical Research and Opinion. https://doi.org/10.1080/03007995.2025.2579378

Zhu, Y., Davoust, A., & Falk, T. H. (2025). DeepSick: Deceiving voice-based diagnostic models with synthetic multilingual pathological speech signals. In Proceedings of SMC 2025. IEEE. https://doi.org/10.1109/SMC58881.2025.11343240

Downloads

Published

2026-08-10

How to Cite

Adamiec, W., Danilczuk, D., Buszko, J., Dyś, M., Baran, N., Bajkowska-Piterak, M., Piterak, P. ., Szlachta-Gubernat, M., & Buż, A. (2026). VOICE AND COUGH AS DIGITAL BIOMARKERS (2025–2026): LINGUISTIC EQUITY, PRIVACY, AND THE PROMISE OF ACOUSTIC DIAGNOSTICS IN GLOBAL HEALTH. International Journal of Innovative Technologies in Social Science, 1(3(51). https://doi.org/10.31435/ijitss.3(51).2026.5732

Most read articles by the same author(s)

1 2 > >>