ChatGPT extension for scientific research work
An Open Source text-to-speech system built by inverting Whisper
Distill high-value content like books, long videos, podcasts, and more
Agentic, Reasoning, and Coding (ARC) foundation models
Persistent context and multi-instance coordination
Ultimate meta-skill for generating best-in-class Claude Code skills
MOSS‑TTS Family open‑source speech and sound generation model
A specialized Claude Code workspace for creating long-form
Structured RAG: ingest, index, query
Claude Code blog skill suite
Long-form streaming TTS system for multi-speaker dialogue generation
Guiding Instruction-based Image Editing via Multimodal Large Language
Codex plugin that turns attached object images into code-only
Generate high-definition story short videos with one click using AI
OCR expert VLM powered by Hunyuan's native multimodal architecture
Your Personal Research Multi-Tool
Claude Code skill for generating production-quality SVG+PNG technical
Fully Local Manus AI. No APIs, No $200 monthly bills
Large Multimodal Models for Video Understanding and Editing
Machine Learning Pipelines for Kubeflow
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Open Multilingual Multimodal Chat LMs
No-code tool for creating a neural search solution in minutes
WaveRNN Vocoder + TTS
A PyTorch implementation of "Capsule Graph Neural Network"