DiVR (Disordered Voice Recognition) - Benchmark
This repository contains the work that enables working with various disordered voice databases using the divr-diagnosis label standardization toolkit.
Installation
pip install divr-benchmark
How to use
While you can generate your own tasks, we provide a battery of tasks that we have used across a wide range of experiments. You can read more about them in Tasks.
Generating tasks
You can generate new tasks from the databases (AVFAD, MEEI, SVD, Torgo, UASpeech, UncommonVoice, VOICED). Of these SVD, Torgo and VOICED as publicly accessible and scripts can download the data automatically provided the database is still available on the expected URLs.
from divr_diagnosis import diagnosis_maps
from divr_benchmark import Benchmark, Diagnosis
benchmark = Benchmark(
storage_path="/home/user/divr_benchmark/storage",
version="v1",
sample_rate=16000,
)
diag_map = diagnosis_maps.CaRLab_2025()
async def filter_func(database_func: DatabaseFunc):
# You can filter the data by min_tasks, so thate every speaker has at least N audios
# this is called 'task' because in most datasets the audios represent different vocal tasks
db = await database_func(name="svd", min_tasks=None)
diag_level = diag_map.max_diag_level
def filter_unclassified(tasks): # example of filtering tasks by label
# You can also get task.speaker_id which can be used to count
# number of diag/speaker and restrict which diags are used for the dataset
return [task for task in tasks if not task.label.incompletely_classified]
return Dataset(
train=filter_unclassified(db.all_train(level=diag_level)),
val=filter_unclassified(db.all_val(level=diag_level)),
test=filter_unclassified(db.all_test(level=diag_level)),
)
benchmark.generate_task(
filter_func=filter_func,
task_path="/home/user/divr_benchmark/tasks/all",
diagnosis_map=diag_level,
allow_incomplete_classification=False,
)
Using existing tasks
Almost all functions of the library accept a level parameter which decides which level of diagnosis is the operation performed on. These parameters default to the maximum diagnostic level if left as None, i.e. the narrowest diagnosis furthest away from the binary detection.
from divr_diagnosis import diagnosis_maps
from divr_benchmark import Benchmark, Diagnosis
benchmark = Benchmark(
storage_path="/home/user/divr_benchmark/storage",
version="v1",
sample_rate=16000,
)
# The diagnosis map here can be different from the one used for generating the tasks
# the library will automatically map diagnosis which can be mapped to the new map
# automatically, and unmapped items will be left as unclassified
diag_map = diagnosis_maps.CaRLab_2025()
task = benchmark.load_task(
task_path="/home/user/divr_benchmark/tasks/all",
diag_level=None,
diagnosis_map=diag_map,
load_audios=True,
)
# Training at default level of diagnosis
for train_point in task.train:
point_id = train_point.id
audio = train_point.audio
label = task.diag_to_index(
diag=train_point.label,
level=None,
)
# Training at root/0th level of diagnosis. Equivalent to binary detection
for train_point in task.train:
point_id = train_point.id
audio = train_point.audio
label = task.diag_to_index(
diag=train_point.label,
level=0,
)
# Validating
for val_point in task.val:
point_id = val_point.id
audio = val_point.audio
label = task.diag_to_index(
diag=val_point.label,
level=None,
)
# Testing
for test_point in task.test:
point_id = test_point.id
audio = test_point.audio
label = task.diag_to_index(
diag=test_point.label,
level=None,
)
# Class weights for cross entropy loss
class_weights = task.train_class_weights(level=None) # level defaults to max level of label
loss_fn = nn.CrossEntropyLoss(weight=torch.tensor(class_weights))
# Convert predicted index to diagnosis
diagnosis = task.index_to_diag(
index=index,
level=None,
)
print(diagnosis.name)
# Get all unique diagnosis in the data
diagnosis_names = task.unique_diagnosis(level=None)
How to cite
Coming soon
Metadata
Release files for divr-benchmark 0.1.4
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| divr_benchmark-0.1.4.tar.gz | 1.3 MB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| divr_benchmark-0.1.4-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 2.8 MB
Release files / divr_benchmark-0.1.4.tar.gz
| Download URL | divr_benchmark-0.1.4.tar.gz |
|---|---|
| Size | 1.3 MB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
9626db04fdbba5257271540ab6a34e883a29f1e5e52b0b620e9520cd503007cc
|
|
BLAKE2b-256 checksum How to use checksums |
9c6df526abd6fba9cbd76e7e4efb8af2eb5afb9439460db5b81cff0cf92ba58c
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/6.1.0 CPython/3.12.9
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Mar 16, 2025.
Transparency logRelease files / divr_benchmark-0.1.4-py3-none-any.whl
| Download URL | divr_benchmark-0.1.4-py3-none-any.whl |
|---|---|
| Size | 1.5 MB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
e2919484f25fad3fa8d08606a056ce93fc7164277663b422bdff38294ae8bd69
|
|
BLAKE2b-256 checksum How to use checksums |
f13a2e0a971d9dccec1f4d21a03706e97e691b6c4135662d8fd8c0155bbaf8e9
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/6.1.0 CPython/3.12.9
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Mar 16, 2025.
Transparency log