A Fair and Transparent Framework for Speech-Based Depression Detection: Balancing Interpretability and Performance

Estevez, Mariel; Ortega, Alfonso; Miguel, Antonio; Lleida, Eduardo

[Submitted on 30 Jun 2026]

Title:A Fair and Transparent Framework for Speech-Based Depression Detection: Balancing Interpretability and Performance

Authors:Mariel Estevez, Alfonso Ortega, Antonio Miguel, Eduardo Lleida

Abstract:While speech provides rich, non-invasive biomarkers for mental-health assessment, clinical adoption is limited by opaque models and potential demographic bias. In this work we propose a methodological framework to evaluate robustness and interpretability for automated depression detection on the extended DAIC-WOZ dataset using low-complexity machine learning baselines (RF, SVM, and MLP) chosen to mitigate overfitting and enhance generalization in combination with human-understandable acoustic features (MFCCs, eGeMAPS). To balance accuracy with clinical trust, we leverage explainability methods (LIME and SHAP) for feature selection, validating our findings with statistical significance tests and demographic fairness analyses to mitigate spurious, artifact-driven correlations. Empirical results demonstrate that an optimized subset of explainable AI (XAI)-selected features combined with an MLP architecture achieves a state-of-the-art test accuracy of 82\%. Ultimately, this work provides a transparent framework for robust and ethical assistive technologies that can be applied to any other binary task.

Comments:	7 pages, 2 figures, 3 tables. This work has been submitted to the IEEE for possible publication
Subjects:	Audio and Speech Processing (eess.AS)
Cite as:	arXiv:2606.31730 [eess.AS]
(or arXiv:2606.31730v1 [eess.AS] for this version)
https://doi.org/10.48550/arXiv.2606.31730

Electrical Engineering and Systems Science> Audio and Speech Processing

Title:A Fair and Transparent Framework for Speech-Based Depression Detection: Balancing Interpretability and Performance

Submission history

Access Paper:

Current browse context:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Electrical Engineering and Systems Science> Audio and Speech Processing

Title:A Fair and Transparent Framework for Speech-Based Depression Detection: Balancing Interpretability and Performance

Submission history

Access Paper:

Current browse context:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators