An Open Source text-to-speech system built by inverting Whisper
ChatGPT extension for scientific research work
Agentic, Reasoning, and Coding (ARC) foundation models
Persistent context and multi-instance coordination
Distill high-value content like books, long videos, podcasts, and more
Long-form streaming TTS system for multi-speaker dialogue generation
A specialized Claude Code workspace for creating long-form
Structured RAG: ingest, index, query
Ultimate meta-skill for generating best-in-class Claude Code skills
Claude Code blog skill suite
MOSS‑TTS Family open‑source speech and sound generation model
Fully Local Manus AI. No APIs, No $200 monthly bills
OCR expert VLM powered by Hunyuan's native multimodal architecture
Claude Code skill for generating production-quality SVG+PNG technical
Generate high-definition story short videos with one click using AI
Large Multimodal Models for Video Understanding and Editing
Your Personal Research Multi-Tool
Machine Learning Pipelines for Kubeflow
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Open Multilingual Multimodal Chat LMs
Guiding Instruction-based Image Editing via Multimodal Large Language
No-code tool for creating a neural search solution in minutes
WaveRNN Vocoder + TTS
A PyTorch implementation of "Capsule Graph Neural Network"
Code base for the precision, recall, density, and coverage metrics