Skip to content

Repository files navigation

pymbrola

GitHub Actions Workflow Status PyPI - Version PyPI - Python Version GitHub License PyPI - Status Docker Image Size (tag) GitHub Release Codecov


A Python interface for the MBROLA speech synthesizer, enabling programmatic creation of MBROLA-compatible phoneme files and automated audio synthesis. This module validates phoneme, duration, and pitch sequences, generates .pho files, and can call the MBROLA executable to synthesize speech audio from text-like inputs.

References: Dutoit, T., Pagel, V., Pierret, N., Bataille, F., & Van der Vrecken, O. (1996, October). The MBROLA project: Towards a set of high quality speech synthesizers free of use for non commercial purposes. In Proceeding of Fourth International Conference on Spoken Language Processing. ICSLP'96 (Vol. 3, pp. 1393-1396). IEEE. https://doi.org/10.1109/ICSLP.1996.607874

Installation

You can install pymbrola from the PyPi repository using pip or uv:

pip install mbrola # pip installation
uv add mbrola      # uv installation

In either case, you will need Python>=3.10. To synthesise audios via MBROLA, you will need to download it and compile it. The pymbrola package has functions for this. This will download MBROLa from numediat/MBROLA to you home folder ~/.mbrola and compile it.

import mbrola

mbrola.install_mbrola()

Important

MBROLA is currently available only on Linux-based systems like Ubuntu, or on Windows via the Windows Subsystem for Linux (WSL). Native Windows and macOS are not yet compatible with the pymbrola package.

Finally, you will need to download some MBROLA voices from numediart/MBROLA-voices. These voices will be automatically downloaded and found by pymbrola at ~/.mbrola/Voices:

mbrola.install_voice("it4")  # install it4 voice
mbrola.install_voice(["it4", "fr4"])  # install several voices
mbrola.install_voice()  # install all voices (~534M)

Tip

A Docker image of Ubuntu 22.04 with a ready-to-go installation of MBROLA is available, for convenience.

Usage

import mbrola

# Create an MBROLA object
caffe = MBROLA(
    phon=["k", "a", "f", "f", "E1"],
    durations=100,  # or [100, 120, 100, 110]
    pitch=[100, [200, 50, 200], 100, 100, 200],
)

# Display phoneme sequence
print(caffe)

# Export PHO file
caffe.export_pho("caffe.pho")

# Synthesize and save audio (WAV file)
caffe.make_sound("caffe.wav", voice="it4")

The mbrola module uses the MBROLA command line tool under the hood. Ensure MBROLA is installed and available in your system path, or WSL if on Windows.

License

pymbrola is distributed under the terms of the MIT license.

Supported by

Funded by the European Union. Views and opinions expressed are however those of the author(s) only and do not necessarily reflect those of the European Union or the European Research Council Executive Agency (ERCEA). Neither the European Union nor the granting authority can be held responsible for them. This work is supported by the ERC StG 101115991 (GALA) awarded to Chiara Santolin

https://erc.europa.eu/homepage

About

A Python front-end for the MBROLA speech synthesizer.

Topics

Resources

Code of conduct

Contributing

Security policy

Stars

1 star

Watchers

1 watching

Forks

Releases

Packages

Used by

Contributors

Languages