A Python library for audio feature extraction, classification, segmentation and applications

This doc contains general info. Click here for the complete wiki

News

Latest pyAudioAnalysis update [2018-08-12] now compatible with Python 3
Check out pyVisualizeMp3Tags a python script for visualization of mp3 tags and lyrics
Check out paura a python script for realtime recording and analysis of audio data
PLOS-One Paper regarding pyAudioAnalysis (please cite!)
Checkout the tutorial library for the course "Multimodal Information Processing & Analysis" of the MSc in Data Science in NCSR Demokritos

General

pyAudioAnalysis is a Python library covering a wide range of audio analysis tasks. Through pyAudioAnalysis you can:

Extract audio features and representations (e.g. mfccs, spectrogram, chromagram)
Classify unknown sounds
Train, parameter tune and evaluate classifiers of audio segments
Detect audio events and exclude silence periods from long recordings
Perform supervised segmentation (joint segmentation - classification)
Perform unsupervised segmentation (e.g. speaker diarization)
Extract audio thumbnails
Train and use audio regression models (example application: emotion recognition)
Apply dimensionality reduction to visualize audio data and content similarities

Installation

Install dependencies:

pip install numpy matplotlib scipy sklearn hmmlearn simplejson eyed3 pydub

Clone the source of this library:

git clone https://github.com/tyiannak/pyAudioAnalysis.git

Install using pip:

pip install -e .

(also works with pip3 now)

An audio classification example

More examples and detailed tutorials can be found at the wiki

pyAudioAnalysis provides easy-to-call wrappers to execute audio analysis tasks. Eg, this code first trains an audio segment classifier, given a set of WAV files stored in folders (each folder representing a different class) and then the trained classifier is used to classify an unknown audio WAV file

from pyAudioAnalysis import audioTrainTest as aT
aT.featureAndTrain(["classifierData/music","classifierData/speech"], 1.0, 1.0, aT.shortTermWindow, aT.shortTermStep, "svm", "svmSMtemp", False)
aT.fileClassification("data/doremi.wav", "svmSMtemp","svm")
Result:
(0.0, array([ 0.90156761,  0.09843239]), ['music', 'speech'])

In addition, command-line support is provided for all functionalities. E.g. the following command extracts the spectrogram of an audio signal stored in a WAV file: python audioAnalysis.py fileSpectrogram -i data/doremi.wav

Author

Theodoros Giannakopoulos, Director of Machine Learning at Behavioral Signals

Name		Name	Last commit message	Last commit date
Latest commit History 487 Commits
pyAudioAnalysis		pyAudioAnalysis
tests		tests
.gitignore		.gitignore
LICENSE.md		LICENSE.md
README.md		README.md
icon.png		icon.png
setup.py		setup.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

A Python library for audio feature extraction, classification, segmentation and applications

News

General

Installation

An audio classification example

Further reading

Author

About

Releases

Packages

Languages

License

waagnermann/pyAudioAnalysis

Folders and files

Latest commit

History

Repository files navigation

A Python library for audio feature extraction, classification, segmentation and applications

News

General

Installation

An audio classification example

Further reading

Author

About

Resources

License

Stars

Watchers

Forks

Releases

Packages 0

Languages

Packages