Creates audio supercuts
Project description
Audiogrep transcribes audio files and then creates “audio supercuts” based on search phrases. It uses [CMU Pocketsphinx](http://cmusphinx.sourceforge.net/) for speech-to-text and [pydub](http://pydub.com/) to stitch things together.
Here’s some [sample output](http://lav.io/2015/02/audiogrep-automatic-audio-supercuts/).
##Requirements Install using pip ` pip install audiogrep ` Install [ffmpeg](http://ffmpeg.org/) with Ogg/Vorbis support. If you’re on a mac with [homebrew](http://brew.sh/) you can install ffmpeg with: ` brew install ffmpeg --with-libvpx --with-libvorbis ` Finally, install [CMU Pocketsphinx](http://cmusphinx.sourceforge.net/). For mac users I followed [these instructions](https://github.com/watsonbox/homebrew-cmu-sphinx) to get it working: ` brew tap watsonbox/cmu-sphinx brew install --HEAD watsonbox/cmu-sphinx/cmu-sphinxbase brew install --HEAD watsonbox/cmu-sphinx/cmu-sphinxtrain # optional brew install --HEAD watsonbox/cmu-sphinx/cmu-pocketsphinx `
##How to use it First, transcribe the audio (you’ll only need to do this once per audio track, but it can take some time) ` # transcribes all mp3s in the selected folder audiogrep --input path/to/*.mp3 --transcribe ` Then, basic use: ` # returns all phrases with the word 'word' in them audiogrep --input path/to/*.mp3 --search 'word' ` The previous example will extract phrase chunks containing the search term, but you can also just get individual words: ` audiogrep --input path/to/*.mp3 --search 'word' --output-mode word ` If you add the ‘–regex’ flag you can use regular expressions. For example: ` # creates a supercut of every instance of the words "spectre", "haunting" and "europe" audiogrep --input path/to/*.mp3 --search 'spectre|haunting|europe' --output-mode word ` You can also construct ‘frankenstein’ sentences (mileage may vary): ` # stupid joke audiogrep --input path/to/*.mp3 --search 'my voice is my passport' --output-mode franken `
###Options
audiogrep can take a number of options:
####–input / -i mp3 file or pattern for input
####–output / -o Name of the file to generate. By default this is “supercut.mp3”
####–search / -s Search term
####–output-mode / -m Splice together phrases, single words, fragments with wildcards, or “frankenstein” sentences. Options are: * sentence: (this is the default) * word * fragment * franken
####–padding / -p Time in milliseconds to add between audio segments. Default is 0.
####–crossfade / -c Time in milliseconds to crossfade audio segments. Default is 0.
####–demo / -d Show the results of the search without outputing a file
Project details
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
Hashes for audiogrep-0.1.2-py2-none-any.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | 52e98bb0c3e375f3b55a0e14bd07c916c9354f8d173de2d4b43ef76371563dad |
|
MD5 | 48f2d321ad51172828059a4d10cb0110 |
|
BLAKE2b-256 | bd36732ac22bb17a0057c871ad37daf66a228ad3b88b5bf07441fa52f315ee80 |