An Open Real-time Video-Language Interaction System
MOSS-TTS-Nano is an open-source multilingual tiny speech generation
Distill your ex into an AI Skill
SpikingJelly is an open-source deep learning framework
A general fine-tuning kit geared toward image/video/audio diffusion
Llama Chinese community, real-time aggregation
Memory Management Kit for Agents
AI-Powered Wiki Generator for GitHub/Gitlab/Bitbucket Repositories
High-resolution models for human tasks
Video understanding codebase from FAIR for reproducing video models
Multimodal-Driven Architecture for Customized Video Generation
Codex plugin that turns attached object images into code-only
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
SDK for building interactive UI components over MCP for AI tools
Machine learning image inpainting task that removes watermarks
Foundation Model for Tabular Data
Personal notes from Wu Enda's machine learning course
AI-Powered Data Processing: Use LOTUS to process all of your datasets
AI-Powered Personalized Learning Assistant
Tiny vision language model
Open-sourced unified customization model
The python library for real-time communication
MCP integration platforms for AI agents to use tools at any scale
Speech-AI-Forge is a project developed around TTS generation model
Mixture-of-Experts Vision-Language Models for Advanced Multimodal