Code for running inference and finetuning with SAM 3 model
Containerized automation engine for programmable CI/CD workflows
Pruna is a model optimization framework built for developers
Lightweight demo to build a conversational AI search engine quickly
Based on AI Agent + MCP toolchain + penetration Skill orchestration
Effortless data labeling with AI support from Segment Anything
Python framework for building scalable multi-agent systems
local-first semantic code search engine
Implementation of the Surya Foundation Model for Heliophysics
Universal LLM Deployment Engine with ML Compilation
High-performance inference framework for large language models
Cloud-native open source data warehouse for analytics and AI queries
A modular graph-based Retrieval-Augmented Generation (RAG) system
A tension reasoning engine over 131 S-class problems
Video Object and Interaction Deletion
Improve human sleep through scientifically
Video understanding codebase from FAIR for reproducing video models
GPU accelerated decision optimization
On-device Speech-to-Intent engine powered by deep learning
Local long-term memory engine for AI apps with persistent storage
TokenSpeed is a speed-of-light LLM inference engine
A tiny scalar-valued autograd engine and a neural net library
A community-supported supercharged version of paperless
InvokeAI is a leading creative engine for Stable Diffusion models
An Open Source package that allows video game creators