Open source demo platform where you can easily showcase your AI models
From Paper to Presentation in One Click
New family of code large language models (LLMs)
A Systematic Framework for Interactive World Modeling
Generate blog articles from video or audio
Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
SOTA discrete acoustic codec models with 40/75 tokens per second
Controllable and fast Text-to-Speech for over 7000 languages
DeepMind model for tracking arbitrary points across videos & robotics
Tooling for the Common Objects In 3D dataset
code for Mesh R-CNN, ICCV 2019
Uncommon Objects in 3D dataset
PyTorch code and models for VJEPA2 self-supervised learning from video
Language modeling in a sentence representation space
An AI-powered security review GitHub Action using Claude
Proofs, cases, concept supplements, and reference explanations
A library for scientific machine learning & physics-informed learning
Best practices on recommendation systems
Welcome to GR00T Whole-Body Control (WBC)
Automated translation solution for visual novels
Open platform connecting AI agents to tools via unified MCP server
WhatsApp MCP server enabling AI access to chats and messaging
Containerized automation engine for programmable CI/CD workflows
Supercharge Your LLM with the Fastest KV Cache Layer
AI logo animation skill: turn raster logos into smooth SVG animation