Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Qwen2.5-Coder is the code version of Qwen2.5, the large language model
Open-source, high-performance Mixture-of-Experts large language model
Video+code lecture on building nanoGPT from scratch
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Powerful open source image generation model
Open Multilingual Multimodal Chat LMs
The ChatGPT Retrieval Plugin lets you easily find personal documents
Towards Ultimate Expert Specialization in Mixture-of-Experts Language
Official code for Style Aligned Image Generation via Shared Attention
Official PyTorch Implementation of "Scalable Diffusion Models"
Code release for ConvNeXt V2 model
Learning to Act by Watching Unlabeled Online Videos
Code release for "Masked-attention Mask Transformer
Simple Text-Generator with OpenAI gpt-2 Pytorch Implementation
A library for Multilingual Unsupervised or Supervised word Embeddings
Code for reproducing key results in the paper
Code for "Image Generation from Scene Graphs", Johnson et al, CVPR 201
Efficient Image Captioning code in Torch, runs on GPU
Open-source code agent designed for Lean 4
JetBrains’ 4B parameter code model for completions
OpenAI’s compact 20B open model for fast, agentic, and local use
Vision-language-action model for robot control via images and text