Generate blog articles from video or audio
SOTA discrete acoustic codec models with 40/75 tokens per second
Controllable and fast Text-to-Speech for over 7000 languages
Unified Multimodal Understanding and Generation Models
Python examples of popular machine learning algorithms
DeepMind model for tracking arbitrary points across videos & robotics
Global weather forecasting model using graph neural networks and JAX
Expose your FastAPI endpoints as Model Context Protocol (MCP) tools
code for Mesh R-CNN, ICCV 2019
Uncommon Objects in 3D dataset
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
PyTorch code and models for VJEPA2 self-supervised learning from video
Language modeling in a sentence representation space
An AI-powered security review GitHub Action using Claude
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
Renderer for the harmony response format to be used with gpt-oss
Integrate, train and manage any AI models and APIs with your database
Detecting silent model failure. NannyML estimates performance
A library for scientific machine learning & physics-informed learning
A cross-platform Python library for differentiable programming
Data science on data without acquiring a copy
Chatbot daemon that connects to your favorite chat services
AsrTools: Smart Voice-to-Text Tool
Structured Outputs
E2M converts various file types (doc, docx, epub, html, htm, url