D3TEC Dataset: a data collection for deep learning research in depression classification featuring voice recordings of Spanish speakers using professional and cellphone microphones
| dc.audience.educationlevel | Investigadores/Researchers | es_MX |
| dc.contributor.advisor | Trejo Rodríguez, Luis Ángel | |
| dc.contributor.author | Brenes García, Luis Felipe | |
| dc.contributor.cataloger | emimmayorquin | |
| dc.contributor.committeemember | Villaseñor Pineda, Luis | |
| dc.contributor.committeemember | Sosa Hernández, Víctor Adrián | |
| dc.contributor.department | School of Engineering and Sciences | es_MX |
| dc.contributor.institution | Campus Monterrey | es_MX |
| dc.contributor.mentor | Cantoral Ceballos, José Antonio | |
| dc.date.accepted | 2024-05-23 | |
| dc.date.accessioned | 2025-05-10T00:34:26Z | |
| dc.date.issued | 2024-05 | |
| dc.description | https://orcid.org/0000-0001-9741-4581 | |
| dc.description.abstract | Depression is a mental health condition that affects millions of people worldwide. Although common, it remains difficult to diagnose due to its heterogeneous symptomatology. Mental health questionnaires are currently the most used assessment method to screen depression; these, however, have a subjective nature due to their dependence on patients' self-assessments. Researchers have been interested in finding an accurate way of identifying depression through an objective biomarker. Recent developments in neural networks and deep learning have enabled the possibility of classifying depression through the computational analysis of voice recordings. However, this approach is heavily dependent on the availability of datasets to train and test deep learning models, and these are scarce. There are also very few languages available. This study proposes a protocol for the collection of a new dataset for deep learning research on voice depression classification, featuring Spanish speakers, professional and smartphone microphones, and a high-quality recording standard. This work aims at creating a high-quality voice depression dataset by recording Spanish speakers with a professional microphone and strict audio quality standards. The data is captured by a smartphone microphone as well for further research in the use of smartphone applications for depression identification. Our methodology involves the strategic collection of depressed and non-depressed voice recordings. Three types of data are collected: voice recordings, depression labels (using the PHQ-9 questionnaire), and additional data that could potentially influence speech. Recordings are captured with professional-grade and smartphone microphones simultaneously to ensure versatility and practical applicability. Several considerations and guidelines are described to ensure high audio quality and avoid potential bias in deep learning research. This data collection effort immediately enables new research topics on depression classification. Some potential uses include deep learning research on Spanish speakers, an evaluation of the impact of audio quality on developing audio classification models, and an evaluation of the applicability of voice depression classification technology on smartphone applications. A preliminary experimentation section is included to showcase the potential research areas that the creation of this dataset enables. This research marks a significant step towards the objective and automated classification of depression in voice recordings. By focusing on the underrepresented demographic of Spanish speakers, the inclusion of smartphone recordings, and addressing the current data limitations in audio quality, this study lays the groundwork for future advancements in deep learning-driven mental health diagnosis. | es_MX |
| dc.description.degree | Master of Science in Computer Science | es_MX |
| dc.format.medium | Texto | es_MX |
| dc.identificator | 120318 | |
| dc.identifier.citation | Brenes, L. F. (2024). D3TEC Dataset: A data collection for deep learning research in depression classification featuring voice recordings of Spanish speakers using professional and cellphone microphones. [Tesis maestría]. Instituto Tecnológico y de Estudios Superiores de Monterrey. Recuperado de: https://hdl.handle.net/11285/703640 | |
| dc.identifier.cvu | 1239618 | es_MX |
| dc.identifier.orcid | https://orcid.org/0009-0003-9901-6140 | |
| dc.identifier.uri | https://hdl.handle.net/11285/703640 | |
| dc.language.iso | eng | es_MX |
| dc.publisher | Instituto Tecnológico y de Estudios Superiores de Monterrey | es_MX |
| dc.relation | Instituto Tecnológico y de Estudios Superiores de Monterrey | |
| dc.relation | CONAHCYT | |
| dc.relation | Organización de los Estados Americanos (OEA) | |
| dc.relation.isFormatOf | acceptedVersion | es_MX |
| dc.rights | openAccess | es_MX |
| dc.rights.uri | http://creativecommons.org/licenses/by-nd/4.0 | es_MX |
| dc.subject.classification | INGENIERÍA Y TECNOLOGÍA::CIENCIAS TECNOLÓGICAS::TECNOLOGÍA DE LOS ORDENADORES::SISTEMAS DE INFORMACIÓN, DISEÑO Y COMPONENTES | |
| dc.subject.keyword | Automated depression diagnosis | |
| dc.subject.keyword | Depression | |
| dc.subject.keyword | Spanish language dataset | |
| dc.subject.keyword | Depression binary classification | |
| dc.subject.keyword | High-Quality Audio Data | |
| dc.subject.keyword | Deep Learning for Audio Analysis | |
| dc.subject.keyword | High-Quality and Smartphone Audio Recordings | |
| dc.subject.lcsh | Technology | es_MX |
| dc.title | D3TEC Dataset: a data collection for deep learning research in depression classification featuring voice recordings of Spanish speakers using professional and cellphone microphones | es_MX |
| dc.type | Tesis de Maestría / master Thesis | es_MX |
Files
Original bundle
1 - 3 of 3
Loading...
- Name:
- BrenesGarcıa_TesisMaestria.pdf
- Size:
- 18.19 MB
- Format:
- Adobe Portable Document Format
- Description:
- Tesis Maestría
Loading...
- Name:
- BrenesGarcia_CartaAutorizacion.pdf
- Size:
- 77.58 KB
- Format:
- Adobe Portable Document Format
- Description:
- Carta Autorización
Loading...
- Name:
- BrenesGarcıa_FirmaActadeGrado.pdf
- Size:
- 320.2 KB
- Format:
- Adobe Portable Document Format
- Description:
- Firma Acta de Grado
License bundle
1 - 1 of 1
Loading...
- Name:
- license.txt
- Size:
- 1.3 KB
- Format:
- Item-specific license agreed upon to submission
- Description:

