High accuracy RAG for answering questions from scientific documents
A Unified Framework for Text-to-3D and Image-to-3D Generation
Free, high-quality text-to-speech API endpoint to replace OpenAI
GLM-4-Voice | End-to-End Chinese-English Conversational Model
A Powerful Native Multimodal Model for Image Generation
A julia code generator for regular expressions
AI-powered tool for generating, optimizing, and translating subtitles
Generate audiobooks from e-books, voice cloning & 1107+ languages
Reading book source
Label Studio is a multi-type data labeling and annotation tool
A Model Context Protocol (MCP) server
TextWorld is a sandbox learning environment for the training
Python library for building agents that leverages Google Antigravity
A Family of Open Sourced Music Foundation Models
Make bilingual epub books Using AI translate
Qwen-Image is a powerful image generation foundation model
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
Unifying 3D Mesh Generation with Language Models
Create videos with Stable Diffusion
Javascript Canvas Library and SVG-to-Canvas Parser
A modular voice assistant application for experimenting
Open source healthcare AI
An Open Source text-to-speech system built by inverting Whisper
Tools like web browser, computer access and code runner for LLMs
Interface for OuteTTS models