Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.
Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
Start Free
$300 Free Credits to Build on Google Cloud
New customers can spin up VMs, build with AI, and query data at no cost.
Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
Text-to-Speech conversor for Basque and Spanish. It includes linguistic processing and built voices for the languages aforementioned. Its acoustic engine is based on hts_engine and it uses a high quality vocoder called AhoCoder.
Developed by Aholab Signal Processing Laboratory: https://aholab.ehu.es/aholab/
http://aholab.ehu.es/ahocoder/
DeepSpeech is an open source embedded (offline, on-device) speech-to-text engine which can run in real time on devices ranging from a Raspberry Pi 4 to high power GPU servers. DeepSpeech is an open-source Speech-To-Text engine, using a model trained by machine learning techniques based on Baidu's Deep Speech research paper. Project DeepSpeech uses Google's TensorFlow to make the implementation easier. A pre-trained English model is available for use and can be downloaded following the instructions in the usage docs. ...
Text-to-Speech TTS for Basque, Spanish, Catalan, Galician and English
Text-to-Speech conversor for Basque, Spanish, Catalan, Galician and English.
It includes linguistic processing and built voices for all the languages aforementioned. Its acoustic engine is based on hts_engine and it uses a high quality vocoder called AhoCoder.
Developed by Aholab Signal Processing Laboratory: https://aholab.ehu.es/aholab/
http://aholab.ehu.es/ahocoder/
Text to Speech engine for English and many other languages. Compact size with clear but artificial pronunciation. Available as a command-line program with many options, a shared library for Linux, and a Windows SAPI5 version.
The project provides a ready-to-use interface for the julius CSR engine for a handicapped child which is not able to use the keyboard well. It integrates into X11 and Windows.
Find out how you can help: http://simon-listens.org/index.php?support
Menestrel-Graphics system designed conveniently say (listen) texts, with the format txt, doc, odt, html, and fb2, which may be in zip-form. Festival TTS engine is used with a Russian voice
Easily add interactive animated characters to any application. Can be directly used as a C++ library or through COM\ActiveX wrapper. It makes it easy to trigger animations, set emotions, and speak using a SAPI 5 speech synthesiszer. Built on OGRE3D.
This is a Java Wrapper for Cepstral.com text-to-speech engine. Cepstral makes very affordable realistic synthetic voices and provides the developers with C++ API's. We have developed a JSAPI compliant Java-to-JNI-to-C++ Wrapper to use with Cepstral TTS.