Genome modeling and design across all domains of life
Infinite Worlds with Versatile Interactions
Audio Language Models are Few-Shot Learners
A 0.1B Omni model trained from scratch
Official inference repo for FLUX.2 models
Native and Compact Structured Latents for 3D Generation
Lets make video diffusion practical
Open image model at the forefront of design
Fast stable diffusion on CPU and AI PC
Advanced language and coding AI model
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
A trainable PyTorch reproduction of AlphaFold 3
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Official repository for LTX-Video
Recovering the Visual Space from Any Views
Z80-μLM is a 2-bit quantized language model
GLM-4 series: Open Multilingual Multimodal Chat LMs
Python bindings for llama.cpp
Industrial-level controllable zero-shot text-to-speech system
Code for running inference and finetuning with SAM 3 model
An experimental version of DeepSeek model
Python inference and LoRA trainer package for the LTX-2 audio–video
Open-source multi-speaker long-form text-to-speech model
High-Resolution 3D Assets Generation with Large Scale Diffusion Models
Agentic, Reasoning, and Coding (ARC) foundation models