A speech-text foundation model for real time dialogue
A python library that makes AMR parsing, generation and visualization
Transforming Multimodal Content into Captivating Multilingual Audio
Temporal-Consistent Diffusion Model for Real-World Video
AI-friendly PPT builder skill: 17 hand-polished Chinese PPTX templates
A straightforward method for training your LLM
A Multi-Modal World Model for Reconstructing, Generating, Simulation
CineCLI is a cross-platform command-line movie browser
World's first open-source, agentic video production system
Multi-tool for semantic search
Stable Diffusion web UI
Tensor search for humans
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
Fast stable diffusion on CPU and AI PC
Build Vision Agents quickly with any model or video provider
Flowly is 100x faster than OpenClaw
MOSS‑TTS Family open‑source speech and sound generation model
Fast multimodal LLM for real-time voice interaction and AI apps
General-purpose image editing model that delivers high-fidelity
Real-time voice interactive digital human
Terminal-based CPU stress and monitoring utility
Paste Markdown and AI responses into Word Excel instantly fast
A python tool that uses GPT-4, FFmpeg, and OpenCV
Benchmark LLMs by fighting in Street Fighter 3
The ultimate RAG for your monorepo