A Customizable Image-to-Video Model based on HunyuanVideo
Clean Jupyter notebooks of outputs, metadata, and empty cells
A step-by-step guide to build your own AI agent
Unifying 3D Mesh Generation with Language Models
A personal context-agent that learns how you work
Controllable & emotion-expressive zero-shot TTS
Open-sourced unified customization model
Controllable and fast Text-to-Speech for over 7000 languages
Self hosted & open source anonymous 360 review software
Open source codebase for Scale Agentex
Python Serverless Microframework for AWS
Unified Multimodal Understanding and Generation Models
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
PyTorch code and models for VJEPA2 self-supervised learning from video
kaldi-asr/kaldi is the official location of the Kaldi project
Educational framework exploring multi-agent orchestration
Framework for managing and maintaining multi-language pre-commit hooks
This repo contains the code for 1D tokenizer and generator
Flexible Photo Recrafting While Preserving Your Identity
Multi-Agent daTa geneRation Infra and eXperimentation framework
Build cross-modal and multimodal applications on the cloud
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Powering Amazon custom machine learning chips