Generate blog articles from video or audio
Browser MCP is a Model Context Provider (MCP) server
FAIR Sequence Modeling Toolkit 2
PyTorch code and models for VJEPA2 self-supervised learning from video
Language modeling in a sentence representation space
Demo of a customer service use case implemented with the OpenAI Agents
Python framework for AI workflows and pipelines with chain of thought
Library for serving Transformers models on Amazon SageMaker
Marrying Grounding DINO with Segment Anything & Stable Diffusion
A Universal Customization Method for Single and Multi Conditioning
Safety reasoning models built-upon gpt-oss
Scalable generative AI framework built for researchers and developers
Code for Cicero, an AI agent that plays the game of Diplomacy
In-App assistant SDK to build a multimodal conversational UX websites
A fast library for AutoML and tuning
Hub of ready-to-use datasets for ML models
Enabling web apps to get accessed by AI agents
Python SDK for the Computer Use model Lux, developed by OpenAGI
Multi-modal large language model designed for audio understanding
Large Multimodal Models for Video Understanding and Editing
MiniMax-M2, a model built for Max coding & agentic workflows
RGBD video generation model conditioned on camera input
GUI Exploration Lab. One of the best GUI agent solutions
Solve puzzles. Learn CUDA
Rust async runtime based on io-uring