Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Repo of Qwen2-Audio chat & pretrained large audio language model
A Powerful Native Multimodal Model for Image Generation
Inference script for Oasis 500M
Official implementation of DreamCraft3D
code for Mesh R-CNN, ICCV 2019
The official PyTorch implementation of Google's Gemma models
A 0.1B Omni model trained from scratch
MOSS‑TTS Family open‑source speech and sound generation model
Long-form streaming TTS system for multi-speaker dialogue generation
Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
LLM-based Reinforcement Learning audio edit model
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
CodeGeeX: An Open Multilingual Code Generation Model (KDD 2023)
Qwen2.5-Coder is the code version of Qwen2.5, the large language model
CodeGeeX2: A More Powerful Multilingual Code Generation Model
Open-source, high-performance Mixture-of-Experts large language model
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Powerful open source image generation model
Open Multilingual Multimodal Chat LMs
ChatGPT interface with better UI
The ChatGPT Retrieval Plugin lets you easily find personal documents
Release for Improved Denoising Diffusion Probabilistic Models
Towards Ultimate Expert Specialization in Mixture-of-Experts Language
Official code for Style Aligned Image Generation via Shared Attention