Generates original ARC-AGI-1-style tasks distribution-matched
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
A Multi-Modal World Model for Reconstructing, Generating, Simulation
Unified Multimodal Understanding and Generation Models
Use ChatGPT to summarize the arXiv papers
Open image model at the forefront of design
INT4/INT5/INT8 and FP16 inference on CPU for RWKV language model
Inference script for Oasis 500M
Fast and Universal 3D reconstruction model for versatile tasks
Open Source Speech Language Model
Implementation of "MobileCLIP" CVPR 2024
High-resolution models for human tasks
Ling is a MoE LLM provided and open-sourced by InclusionAI
MOSS‑TTS Family open‑source speech and sound generation model
Advancing Open-source World Models
A Systematic Framework for Interactive World Modeling
Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
High-Resolution Image Synthesis with Latent Diffusion Models
Memory-efficient and performant finetuning of Mistral's models
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Easy Docker setup for Stable Diffusion with user-friendly UI
Video+code lecture on building nanoGPT from scratch
Official DeiT repository