Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
OCR expert VLM powered by Hunyuan's native multimodal architecture
High-Resolution Image Synthesis with Latent Diffusion Models
Memory-efficient and performant finetuning of Mistral's models
Easy Docker setup for Stable Diffusion with user-friendly UI
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Video+code lecture on building nanoGPT from scratch
Official DeiT repository
Chinese LLaMA-2 & Alpaca-2 Large Model Phase II Project
Example Discord bot written in Python that uses the completions API
Let us control diffusion models
Fine-tuning ChatGLM-6B with PEFT
Chinese LLaMA & Alpaca large language model + local CPU/GPU training
A GUI tool for generating subtitle from videos, generating srt files
800,000 step-level correctness labels on LLM solutions to MATH problem
Text-to-3D & Image-to-3D & Mesh Exportation with NeRF + Diffusion
Learning to Act by Watching Unlabeled Online Videos
RoBERTa Chinese pre-training model: RoBERTa for Chinese
Code release for "Masked-attention Mask Transformer
GLIDE: a diffusion-based text-conditional image synthesis model
PyTorch implementation of YOLOv4
Real Time Speech Enhancement in the Waveform Domain (Interspeech 2020)
Large-scale autoregressive pixel model for image generation by OpenAI
A mix of GAN implementations including progressive growing