The behavior guidance framework for customer-facing LLM agents
Accurate × Fast × Comprehensive
A fast TTS architecture with conditional flow matching
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
Speech recognition module for Python
ComfyUI wrapper nodes for HunyuanVideo
A Sublime Text 2/3 plugin to see git diff in gutter
Simple LaTeX parser providing latex-to-unicode and unicode-to-latex
Capable of understanding text, audio, vision, video
Open source OSINT tool for gathering data on emails, phones, and IPs
An open-source toolkit for monitoring Language Learning Models (LLMs)
Automated translation solution for visual novels
Chat with it via text and voice
Unified web UI for training and running open models locally
Spark-TTS Inference Code
A Unified Framework for Text-to-3D and Image-to-3D Generation
Python module for parsing semi-structured text into python tables
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Easy-to-use and powerful NLP library with Awesome model zoo
Designed for text embedding and ranking tasks
Foundational video generation model with 13.6B parameters
Persian NLP Toolkit
Easily compute clip embeddings and build a clip retrieval system
Collection of Gemma 3 variants that are trained for performance
MOSS-TTS-Nano is an open-source multilingual tiny speech generation