persephone

A tool for developing automatic phoneme transcription models

These details have not been verified by PyPI

Project links

Homepage

Project description

Persephone v0.3.0 (beta version)

Persephone (/pərˈsɛfəni/) is an automatic phoneme transcription tool. Traditional speech recognition tools require a large pronunciation lexicon (describing how words are pronounced) and much training data so that the system can learn to output orthographic transcriptions. In contrast, Persephone is designed for situations where training data is limited, perhaps as little as an hour of transcribed speech. Such limitations on data are common in the documentation of low-resource languages. It is possible to use such small amounts of data to train a transcription model that can help aid transcription, yet such technology has not been widely adopted.

The speech recognition tool presented here is named after the goddess who was abducted by Hades and must spend one half of each year in the Underworld. Which of linguistics or computer science is Hell, and which the joyful world of spring and light? For each it’s the other, of course. — Alexis Michaud

The goal of Persephone is to make state-of-the-art phonemic transcription accessible to people involved in language documentation. Creating an easy-to-use user interface is central to this. The user interface and APIs are a work in progress and currently Persephone must be run via a command line.

The tool is implemented in Python/Tensorflow with extensibility in mind. Currently just one model is implemented, which uses bidirectional long short-term memory (LSTMs) and the connectionist temporal classification (CTC) loss function.

We are happy to offer direct help to anyone who wants to use it. If you’re having trouble, contact Oliver Adams at oliver.adams@gmail.com. We are also very welcome to thoughts, constructive criticism, help with design, development and documentation, along with any bug reports or pull requests you may have.

Documentation

Documentation can be found here.

Contributors

Persephone has been built based on the code contributions of:

Oliver Adams
Janis Lesinskis
Ben Foley
Nay San

Citation

If you use this code in a publication, please cite Evaluating Phonemic Transcription of Low-Resource Tonal Languages for Language Documentation:

@inproceedings{adams18evaluating,
title = {Evaluating phonemic transcription of low-resource tonal languages for language documentation},
author = {Adams, Oliver and Cohn, Trevor and Neubig, Graham and Cruz, Hilaria and Bird, Steven and Michaud, Alexis},
booktitle = {Proceedings of LREC 2018},
year = {2018}
}

Project details

These details have not been verified by PyPI

Project links

Homepage

Release history Release notifications | RSS feed

0.4.2

Apr 29, 2020

0.4.1

Nov 13, 2019

0.4.0

Aug 31, 2019

0.3.2

Aug 5, 2018

0.3.1

Jul 15, 2018

This version

0.3.0

Jul 14, 2018

0.2.0

Mar 20, 2018

0.1.9

Feb 19, 2018

0.1.8

Feb 10, 2018

0.1.7

Feb 9, 2018

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

persephone-0.3.0.tar.gz (45.8 kB view details)

Uploaded Jul 14, 2018 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

persephone-0.3.0-py3-none-any.whl (51.4 kB view details)

Uploaded Jul 14, 2018 Python 3

File details

Details for the file persephone-0.3.0.tar.gz.

File metadata

Download URL: persephone-0.3.0.tar.gz
Upload date: Jul 14, 2018
Size: 45.8 kB
Tags: Source
Uploaded using Trusted Publishing? No

File hashes

Hashes for persephone-0.3.0.tar.gz
Algorithm	Hash digest
SHA256	`594e5bfb5260ffac65408e009dee638df67709648c497edc00c471f288f2b705`
MD5	`402bfb7fde6f29c8f6054c471231a8d7`
BLAKE2b-256	`f27f1d0f177177e65a65c1a6ab59646c264107d5d85e748e003d438f253f814e`

See more details on using hashes here.

File details

Details for the file persephone-0.3.0-py3-none-any.whl.

File metadata

Download URL: persephone-0.3.0-py3-none-any.whl
Upload date: Jul 14, 2018
Size: 51.4 kB
Tags: Python 3
Uploaded using Trusted Publishing? No

File hashes

Hashes for persephone-0.3.0-py3-none-any.whl
Algorithm	Hash digest
SHA256	`816f6eaa9eba4819a0c2745ba86ff0f01e7a84055b4589f9d632f7429177b613`
MD5	`a81d8e0bdf58e93eb3c9f1c64e00d917`
BLAKE2b-256	`78ab34db08ddb86991a59dbaeb9bb2e16de4aa7188d6e0758f78fc825481bbfd`

See more details on using hashes here.

persephone 0.3.0

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Persephone v0.3.0 (beta version)

Documentation

Contributors

Citation

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes