Showing posts with label text to speech. Show all posts
Showing posts with label text to speech. Show all posts

Friday, October 16, 2015

Version 1.3 of "Text to Speech for Video" is now available

I've released version 1.3 of my open source text-to-speech program, Text to Speech for Video. Several enhancements have been added to make working with the program easier and more productive, and the "Huckleberry" voice has been improved and expanded (word count now over 2300).

Below is a short video demonstrating what the current three voices (Huckleberry, Donna, and Eve) sound like:
 


The visual elements were created in Muvizu Play+ I've just started working with Play+, but I already like it a lot. I hope to post a more extensive review sometime later this year. *
Share/Save/Bookmark

Wednesday, November 20, 2013

"Text To Speech for Video", new open source program

I am pleased to announce that Text to Speech for Video is now available. Having been rather disappointed in the quality, price, and licensing conditions of many of the text to speech products in existence, I decided to create my own free and open source alternative specifically intended for creators of animated videos who want an easy to use way of generating speech for their characters.  Below is a short video demonstrating what the output sounds like:




Right now there are two voices available - one male with a southern U.S. accent (demonstrated in the video above) and one female with an asian accent (I'd describe it as Indian subcontinent).  Each voice has about 1500 words now (including some duplicates with different emotiveness - questioning, emphatic, etc.).  I hope to add more words to the existing voices and more voices as well in the coming months/years.

It's also quite easy (if you have the patience) to record your own "voice". Since each word is stored as a separate wav file (22 khz, 2 channel stereo), the voices take up quite a lot of room (program plus the two voices currently about 80mb which will grow over time as more words/voices are added).  The upside is that, in contrast to many synthetically generated voices, the output can sound a lot like real speech (since it is real speech, at least on a word for word basis). *
Share/Save/Bookmark