A Next-Generation Training Engine Built for Ultra-Large MoE Models
Llama Chinese community, real-time aggregation
A simple, performant and scalable Jax LLM
scikit-learn compatible tabular foundation model
Repo for SeedVR2 & SeedVR
Bidirectional token-classification model for identifiable info
The Cradle framework is a first attempt at General Computer Control
General-purpose image editing model that delivers high-fidelity
New family of code large language models (LLMs)
Solve puzzles. Learn CUDA
Implement CPU from scratch and play with large model deployments
FAIR Sequence Modeling Toolkit 2
Experimental, AI/ML-powered and open sourced Marketing Mix Modeling
End-to-end speech processing toolkit
Data and tools for generating and inspecting OLMo pre-training data
The open source post-building layer for agents
Collaborative & Open-Source Quality Assurance for all AI models
This repository contains code released by Google Research
Definitions for AI/ML tasks like dataset creation
Multimodal embedding and reranking models built on Qwen3-VL
"Big Model" trains a visual multimodal VLM with 26M parameters
Open-weight, large-scale hybrid-attention reasoning model
Robust recipes to align language models with human and AI preferences
Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
Running large language models on a single GPU