AI cognitive-enhancement Skills based on Anthropic's J-space
CLIP, Predict the most relevant text snippet given an image
ImageBind One Embedding Space to Bind Them All
PyTorch code and models for VJEPA2 self-supervised learning from video
Training Large Language Model to Reason in a Continuous Latent Space
Create videos with Stable Diffusion
AI tool that removes hardcoded subtitles and text from videos locally
AI-Driven Exploration in the Space of Code
State-of-the-art (SoTA) text-to-video pre-trained model
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
Topic Modelling for Humans
Text-space optimizer that trains reusable natural-language skills
A Hyperparameter Tuning Library for Keras
Based on AI Agent + MCP toolchain + penetration Skill orchestration
Multi-Modal Neural Networks for Semantic Search, based on Mid-Fusion
Foundational video generation model with 13.6B parameters
PyTorch code and models for V-JEPA self-supervised learning from video
A Family of Open Sourced Music Foundation Models
An open-source toolkit for monitoring Language Learning Models (LLMs)
Medical imaging toolkit for deep learning
Recovering the Visual Space from Any Views
Multi-modal large language model designed for audio understanding
Implementation of the Surya Foundation Model for Heliophysics
Open Source Differentiable Computer Vision Library
A toolkit to optimize ML models for deployment for Keras & TensorFlow