Unifying 3D Mesh Generation with Language Models
Persian NLP Toolkit
Official MiniMax Model Context Protocol (MCP) server
Converts text to speech in realtime
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
Apache-2.0 open-source image generation and editing model family
An Open Source implementation of Notebook LM with more flexibility
Python library and CLI tool to interface with Google Translate
A community-supported supercharged version of paperless
Implementation of Imagen, Google's Text-to-Image Neural Network
Designed for text embedding and ranking tasks
ASCII art library for Python
An AI-agent skill that turns Markdown into paste-ready WeChat article
Claude Code skill implementing Manus-style persistent planning
Offline inference engine for art, real-time voice conversations
A Unified Framework for Text-to-3D and Image-to-3D Generation
Multimodal-Driven Architecture for Customized Video Generation
Framework for building realtime multimodal voice AI agents apps
AI-powered code assistant for Vim. OpenAI and ChatGPT plugin for Vim
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
Speech-AI-Forge is a project developed around TTS generation model
A Multi-Modal World Model for Reconstructing, Generating, Simulation
Tools to ease the creation of snippets, syntax definitions, etc.
Python & command-line tool to gather text on the Web
Clone a voice in 5 seconds to generate arbitrary speech in real-time