A Customizable Image-to-Video Model based on HunyuanVideo
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Flux 2 image generation model pure C inference
Accurate × Fast × Comprehensive
PyTorch code and models for the DINOv2 self-supervised learning
Implementation of the Surya Foundation Model for Heliophysics
Repo for SeedVR2 & SeedVR
Codex plugin that turns attached object images into code-only
Ling is a MoE LLM provided and open-sourced by InclusionAI
Z80-μLM is a 2-bit quantized language model
Repo of Qwen2-Audio chat & pretrained large audio language model
High-Fidelity and Controllable Generation of Textured 3D Assets
An experimental version of DeepSeek model
CodeGeeX: An Open Multilingual Code Generation Model (KDD 2023)
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
The official PyTorch implementation of Google's Gemma models
4M: Massively Multimodal Masked Modeling
Block Diffusion for Ultra-Fast Speculative Decoding
Access to Anthropic's safety-first language model APIs
OCR expert VLM powered by Hunyuan's native multimodal architecture
Models for object and human mesh reconstruction
Recovering the Visual Space from Any Views
Instructions on how to use the Realtime API on Microcontrollers
Clean and efficient FP8 GEMM kernels with fine-grained scaling
Global weather forecasting model using graph neural networks and JAX