From Images to High-Fidelity 3D Assets
Multi-lingual large voice generation model, providing inference
Instant voice cloning by MIT and MyShell. Audio foundation model
Spark-TTS Inference Code
Open-source infrastructure for Computer-Use Agents. Sandboxes
Offline Text To Speech synthesis for python
Toolkit to help you get started with Spec-Driven Development
Asynchronous multi-platform robot framework written in Python
Python inference and LoRA trainer package for the LTX-2 audio–video
Making ALL Software Agent-Native
SOTA Open Source TTS
Vibe-Trading: Your Personal Trading Agent
Any model. Any hardware. Zero compromise
Tools like web browser, computer access and code runner for LLMs
Definitions for AI/ML tasks like dataset creation
Context management for Claude Code. Hooks maintain state via ledgers
Universal LLM Deployment Engine with ML Compilation
Efficient Triton Kernels for LLM Training
A set of ready to use Agent Skills for research, science, engineering
Long-form streaming TTS system for multi-speaker dialogue generation
Collection of Kaggle Solutions and Ideas
Containerized automation engine for programmable CI/CD workflows
On-device Speech-to-Intent engine powered by deep learning
A multi-function Discord bot
Taming Stable Diffusion for Lip Sync