Official implementation of DreamCraft3D
An Efficient Agentic Model for Computer Use
Unified Multimodal Understanding and Generation Models
Generating Immersive, Explorable, and Interactive 3D Worlds
Official implementation of Watermark Anything with Localized Messages
Video understanding codebase from FAIR for reproducing video models
Personalize Any Characters with a Scalable Diffusion Transformer
Achieving 3+ generation speedup on reasoning tasks
Open-Source Financial Large Language Models
ICLR2024 Spotlight: curation/training code, metadata, distribution
PyTorch code and models for the DINOv2 self-supervised learning
A Customizable Image-to-Video Model based on HunyuanVideo
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Large-language-model & vision-language-model based on Linear Attention
Chat & pretrained large audio language model proposed by Alibaba Cloud
Real-time behaviour synthesis with MuJoCo, using Predictive Control
Example Discord bot written in Python that uses the completions API
Code for the paper Hybrid Spectrogram and Waveform Source Separation
GLM-130B: An Open Bilingual Pre-Trained Model (ICLR 2023)
Repo for external large-scale work
llama.go is like llama.cpp in pure Golang
A method to increase the speed and lower the memory footprint
Implementation of model parallel autoregressive transformers on GPUs
Code release for ConvNeXt V2 model
A minimal PyTorch re-implementation of the OpenAI GPT