Skip to main content

hupa-voicedb: Príncipe de Asturias Hospital Voice Disorders Database Reader module

PyPI PyPI - Status PyPI - Python Version GitHub

This Python module provides functions to retrieve data and information easily from Príncipe de Asturias Hospital Voice Disorders Database.

This module does not contain the database itself. The database belongs to Prof. Juan I. Godino-Llorente (email: ignacio.godino@upm.es) at Universidad Politécnica de Madrid, and he kindly makes it available for free to non-commercial research use. Users must contact him to obtain the license and to download the database.

Install

pip install hupa-voicedb

Use

from hupa import HUPA

# to initialize (must call this once in every Python session)
db = HUPA('<path to the root directory of the extracted database>')

# to get a copy of the full database as a Pandas dataframe
df = db.query() # default columns: "edad", "sexo", "Codigo"

# to get the patholgy code-name lookup table
# (note: not all pathologies are included in the database)
lut = db.pathologies

# to get age, gender, and R scores
df = db.query(["edad", "sexo", "R"])

# use Pandas' itertuples to read audio data iteratively
for id, *info in df.itertuples():
  # read audio data
  # (normalize to [0,1] unless given additional argument: normlize=False)
  fs, x = db.read_data(id)

  # run the acoustic data through your analysis function, get measurements
  params = my_analysis_function(fs, x)

  # log the measurements along with the age and GRBAS info
  my_logger.log_outcome(id, *auxdata, *params)

# alternately, use database's `iter_data` method to process acoustic data
# iteratively over queried data (all female speakers along with age and G score)
for id, fs, x, auxdata in db.iter_data(auxdata_fields=["edad", "G"],
                                       sexo="M"):
  # run the acoustic data through your analysis function, get measurements
  params = my_analysis_function(fs, x)

  # log the measurements along with the age and GRBAS info
  my_logger.log_outcome(id, *auxdata, *params)

# Finally, to get a dataframe of all the WAV files with their full paths
df = db.get_files(auxdata_fields=['Codigo'])

Metadata

Release files for hupa-voicedb 0.1.1

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for hupa-voicedb 0.1.1
File Size Uploaded
hupa-voicedb-0.1.1.tar.gz 13.4 kB Details

Built distribution (wheel)

Table of built distributions (wheels) for hupa-voicedb 0.1.1
File Interpreter ABI Platform
hupa_voicedb-0.1.1-py3-none-any.whl Python 3 none any Details

Total release size: 26.6 kB

Release files / hupa-voicedb-0.1.1.tar.gz

Download URL hupa-voicedb-0.1.1.tar.gz
Size 13.4 kB
Tags Source
SHA-256 checksum
How to use checksums
48b5f6a77e00e0cf595d51499e0a2ac8651f2a362af41ef0753efcdc0a43b650
BLAKE2b-256 checksum
How to use checksums
eee4060df5c55f457403d12d3bd60f232d98735e0c391010c6a98a5f7f52d6ea
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via twine/4.0.1 CPython/3.11.2

Release files / hupa_voicedb-0.1.1-py3-none-any.whl

Download URL hupa_voicedb-0.1.1-py3-none-any.whl
Size 13.2 kB
Tags Python 3
SHA-256 checksum
How to use checksums
f388b206fd3a536c3b1f4797e53fdcb8b645b055e96d24f890f38784224fbb2c
BLAKE2b-256 checksum
How to use checksums
9d7db4861e685cc8bab15de28e5153db5141a693121eb4cb588e6187a3f30fad
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via twine/4.0.1 CPython/3.11.2

Release history Release notifications | RSS feed

This release

0.1.1 This release

2 release files

0.1.0

2 release files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page