scikit-learn compatible tabular foundation model
1B text generation model based on the HRM architecture
Robust Speech Recognition Across Languages, Dialects
Tiny vision language model
Repo for SeedVR2 & SeedVR
The official PyTorch implementation of Google's Gemma models
Inference code for scalable emulation of protein equilibrium ensembles
Diversity-driven optimization and large-model reasoning ability
New family of code large language models (LLMs)
Qwen-Image-Layered: Layered Decomposition for Inherent Editablity
DeepMind model for tracking arbitrary points across videos & robotics
code for Mesh R-CNN, ICCV 2019
DeepSeek Coder: Let the Code Write Itself
Open-source framework for intelligent speech interaction
Large Multimodal Models for Video Understanding and Editing
Generates original ARC-AGI-1-style tasks distribution-matched
AI cognitive-enhancement Skills based on Anthropic's J-space
Codex plugin that turns attached object images into code-only
Audio Language Models are Few-Shot Learners
A 0.1B Omni model trained from scratch
Foundation model for image generation
Fast-stable-diffusion + DreamBooth
A Pragmatic VLA Foundation Model
OpenTinker is an RL-as-a-Service infrastructure for foundation models