Speech

this is an organically evolving personal wiki-form knowledge base with on-the-fly, copy-edited n otherwise curated patchworks of folksnomies n headings containing trails n spirals of topics, descriptions, notes, breadcrumbs n stubs, links to sites, systems, software, manuals, organisations, people, articles, guides, slides, papers, books, comments, videos, screencasts, webcasts, scratchpads, etc | content is orientated towards mostly free/libre/open, mostly Linux | quality and age varies drastically | sometimes old things are first, sometimes last | Ctrl + mouse wheel to zoom in if text is too small | use the Table of Contents menu to navigate long pages | use the header -ToC links to shrink n expand the menu | dead link? Wayback Machine! | probably need to fix the theme CSS after an update | Chat to msg me (this I am not checking atm) | e

Resources

Weather

Edinburgh

Scotland

Smiley / Lorem

About / ToDo

Meta / misc

Maths

Breath

Being

Grounding

Living

Camping

Mapping

Organising

Media

Digital lit.

Design

Politics / p

Free/open

Volly Guide

Fire brand

Radio / TV

Signal

Data / Open

Type / Emoji

Documents

Markdown

Semantic

Database

VCS / Git

Backup

Compression

Cryptography

Regex

Languages

C/C++ / Lisp

Perl / PHP

Python / Ruby

JavaScript / Lua

Creative coding

Visual / Pd

ML / AI

Storage / Files

Editors/ IDE

Vim / Emacs

Dotfiles / Box

Logging / Search

Notebooks

Computing

Computer / CA

OSs / *nix / CLI

Distros / Packages

Android / Apps

Apple / Windows

Amiga / Emulation

Web dev

Web systems

Wiki / Forums

Feeds

Open social

Scraping

Net/web media

E-mail

Chat / IRC

VoIP / Comms

File sharing

Link / Wi-Fi

Internet / Mesh

Transport / DNS

HTTP(S) / SSH

Stack

MediaWiki

Web Audio

GFX / Colours

UI / X11 / GUI

Terminals / TUI

WM/DE / Wayland

AwesomeWM / i3

File managers

Clipboard

Demoscene

Gaming / AR

Photos / Images

Lighting / Laser

CAD / 3D

Video / Vision

Visuals

Audio / s / AV

Softsynths

Speech / vox

Speaker / s

Sampling

Sound banks

Notation

MIDI / OSC

Tracker

DAW

Generative

Styles

Playback / MPD

Net AV/media

Rip / Tag / t

DJing

Stations

WP: Speech_synthesis

to sort/categorise

https://forums.homeseer.com/showthread.php?t=175012

https://archive.org/details/flexibleformants00lalw - Flexible formant synthesizer : a tool for improving speech production quality
https://archive.org/details/flexiblehighqual00hsie - A flexible and high quality articulatory speech synthesizer

SAM

WP: Software_Automatic_Mouth - or SAM, is a speech synthesis program developed and sold by Don’t Ask Software. The program was released for the Apple II, Lisa, Atari 8-bit family, and Commodore 64.

https://github.com/s-macke/SAM

SAM: Software Automatic Mouth - WASM

rsynth

rsynth - Text-to-Speech.
- https://github.com/rhdunn/rsynth - fork
- YouTube: Virtualized dragging brake equipment detector using rsynth

Festival

Festival - or The Festival Speech Synthesis System, offers a general framework for building speech synthesis systems as well as including examples of various modules. As a whole it offers full text to speech through a number APIs: from shell level, though a Scheme command interpreter, as a C++ library, from Java, and an Emacs interface. Festival is multi-lingual (currently English (British and American), and Spanish) though English is the most advanced. Other groups release new languages for the system.

https://vocaloid.fandom.com/wiki/Festival_Speech_Synthesis_System

Festvox - aims to make the building of new synthetic voices more systemic and better documented, making it possible for anyone to build a new voice.

Rocaloid

Rocaloid - a free, open-source singing voice synthesis system. Its ultimate goal is to fast synthesize natural, flexible and multi-lingual vocal parts. Like other vocal synthesizing software, after installing the vocal database, inputting lyrics and pitch, you can synthesize attractive vocal parts. What’s more, Rocaloid highlights on providing you more controllable parameters which enabling to take control of exquisite dimensions of the synthesized voice and export with better quality. By using a fully constructed Rocaloid Database, you can synthesize singing voice in any phonetic-based languages.
- https://github.com/Rocaloid - dead?

Festvox

Festvox - aims to make the building of new synthetic voices more systemic and better documented, making it possible for anyone to build a new voice. Specifically we offer: Documentation, including scripts explaining the background and specifics for building new voices for speech synthesis in new and supported languages. Example speech databases to help building new voices. Links, demos and a repository for new voices. This work is firmly grounded within Edinburgh University's Festival Speech Synthesis System and Carnegie Mellon University's small footprint Flite synthesis engine.

MaryTTS

MaryTTS is an open-source, multilingual Text-to-Speech Synthesis platform written in Java. It was originally developed as a collaborative project of DFKI’s Language Technology Lab and the Institute of Phonetics at Saarland University. It is now maintained by the Multimodal Speech Processing Group in the Cluster of Excellence MMCI and DFKI.

eSpeak

eSpeak - a compact open source software speech synthesizer for English and other languages, for Linux and Windows. eSpeak uses a "formant synthesis" method. This allows many languages to be provided in a small size. The speech is clear, and can be used at high speeds, but is not as natural or smooth as larger synthesizers which are based on human speech recordings.

https://github.com/divVerent/ecantorix - a singing synthesis frontend for espeak. It works by using espeak to generate raw speech samples, then adjusting their pitch and length and finally creating a LMMS project file referencing the samples in sync to the input file.

OpenSource SpeechSynth

http://web.media.mit.edu/~stefanm/osss/

MBROLA

http://tcts.fpms.ac.be/synthesis/ - The MBROLA Project
- WP: MBROLA

Assistive Context-Aware Toolkit

Assistive Context-Aware Toolkit (ACAT) - an open source platform developed at Intel Labs to enable people with motor neuron diseases and other disabilities to have full access to the capabilities and applications of their computers through very constrained interfaces suitable for their condition. More specifically, ACAT enables users to easily communicate with others through keyboard simulation, word prediction and speech synthesis. Users can perform a range of tasks such as editing, managing documents, navigating the Web and accessing emails. ACAT was originally developed by researchers at Intel Labs for Professor Stephen Hawking, through a very iterative design process over the course of three years.
- http://blogs.msdn.com/b/cdndevs/archive/2015/08/14/intel-just-open-sourced-stephen-hawking-s-speech-system-and-it-s-a-net-4-5-winforms-app.aspx [1]

Praat

Praat - doing phonetics by computer

Gnuspeech

gnuspeech - makes it easy to produce high quality computer speech output, design new language databases, and create controlled speech stimuli for psychophysical experiments. gnuspeechsa is a cross-platform module of gnuspeech that allows command line, or application-based speech output. The software has been released as two tarballs that are available in the project Downloads area of http://savannah.gnu.org/projects/gnuspeech. [2]

Project Merlin

Project Merlin - A truly free virtual singer, no matter how you want to send all kinds of ideas. [3]
- https://github.com/ProjectMeilin - not fully open yet?
- http://www.cstr.ed.ac.uk/projects/merlin/
- http://ml.cs.yamanashi.ac.jp/world/english
- YouTube: UTAU】Honeyworks ママ ver.acoustic 【徵音梅林cover】
- YouTube: 【徴音梅林】Umbrella カバー

UTAU

WP: Utau - a Japanese singing synthesizer application created by Ameya/Ayame. This program is similar to the Vocaloid software, with the difference that it is shareware instead of being released under third party licensing

https://github.com/stakira/OpenUtau - Open source UTAU editing environment.

Sinsy

Sinsy - HMM-based Singing Voice Synthesis System
- http://sinsy.sourceforge.net
- https://github.com/zyamusic/sinsy

Sinsy - Singing Voice Synthesizer - how to

https://github.com/hyperzlib/Sinsy-Remix - The HMM-Based Singing Voice Syntheis System Remix "Sinsy-r"

Mozilla TTS

https://github.com/mozilla/TTS - Deep learning for Text to Speech

CMU Flite

CMU Flite - a small, fast run-time open source text to speech synthesis engine developed at CMU and primarily designed for small embedded machines and/or large servers. Flite is designed as an alternative text to speech synthesis engine to Festival for voices built using the FestVox suite of voice building tools.
- https://github.com/festvox/flite

mesing

https://github.com/usdivad/mesing

Adobe VoCo

IPOX

IPOX - an experimental, all-prosodic speech synthesizer, developed many years ago by Arthur Dirksen and John Coleman. It is still available for downloading, and was designed to run on a 486 PC running Windows 3.1 or higher, with a 16-bit Windows-compatible sound card, such as the Soundblaster 16. It still seems to run on e.g. XP, but I haven't tried it on Vista.

NPSS

Neural Parametric Singing Synthesizer

https://github.com/seaniezhao/torch_npss - pytorch implementation of Neural Parametric Singing Synthesizer 歌声合成

Pink Trombone

Pink Trombone - Bare-handed procedural speech synthesis, version 1.1, March 2017, by Neil Thapen
- https://github.com/giuliomoro/pink-trombone

Klatter

https://github.com/fundamental/klatter - a bare bones formant synthesizer based upon the description given in the 1979 paper "Software For a Cascade/Parallel Formant Synthesizer" by Dennis Klatt. This program was not designed for interactive use, though there is code for some minimal midi control. In it's current state, it is enough of a curiosity that it will be preserved, though it may not see much if any use.

Tacotron 2

https://github.com/Rayhane-mamah/Tacotron-2 - DeepMind's Tacotron-2 Tensorflow implementation

https://github.com/NVIDIA/tacotron2 - Tacotron 2 - PyTorch implementation with faster-than-realtime inference

Real-Time-Voice-Cloning

https://github.com/CorentinJ/Real-Time-Voice-Cloning - Clone a voice in 5 seconds to generate arbitrary speech in real-time

leesampler

https://github.com/GloomyGhost-MosquitoSeal/lessampler - a Singing Voice Synthesizer

Neural Parametric Singing Synthesizer

A Neural Parametric Singing Synthesizer

VoiceOfFaust

https://github.com/magnetophon/VoiceOfFaust - Turn your voice into a synthesizer!