Create videos with Stable Diffusion
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
ImageBind One Embedding Space to Bind Them All
Topic Modelling for Humans
PyTorch code and models for VJEPA2 self-supervised learning from video
Training Large Language Model to Reason in a Continuous Latent Space
CLIP, Predict the most relevant text snippet given an image
AI tool that removes hardcoded subtitles and text from videos locally
Physical Symbolic Optimization
State-of-the-art (SoTA) text-to-video pre-trained model
Implementation of Video Diffusion Models
An open-source toolkit for monitoring Language Learning Models (LLMs)
A Hyperparameter Tuning Library for Keras
Multi-Modal Neural Networks for Semantic Search, based on Mid-Fusion
Medical imaging toolkit for deep learning
A Family of Open Sourced Music Foundation Models
Recovering the Visual Space from Any Views
Synchronized Translation for Videos
PyTorch code and models for V-JEPA self-supervised learning from video
A fast library for AutoML and tuning
AI-Driven Exploration in the Space of Code
PyTorch version of Stable Baselines
Generate Any 3D Scene in Seconds
Superfast AI decision making and processing of multi-modal data
Mixture-of-Experts Vision-Language Models for Advanced Multimodal