Fast and Universal 3D reconstruction model for versatile tasks
Official implementation of DreamCraft3D
Open-source large language model family from Tencent Hunyuan
Community plugin marketplace for Claude Cowork and Claude Code
Netease Youdao's open-source embedding and reranker models
Audio foundation model excelling in audio understanding
Tongyi Deep Research, the Leading Open-source Deep Research Agent
Phi-3.5 for Mac: Locally-run Vision and Language Models
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training
1B text generation model based on the HRM architecture
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
Tiny vision language model
Repo for SeedVR2 & SeedVR
This repository contains the official implementation of research
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Reproduction of Poetiq's record-breaking submission to the ARC-AGI-1
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Renderer for the harmony response format to be used with gpt-oss
High-Resolution Image Synthesis with Latent Diffusion Models
Repo of Qwen2-Audio chat & pretrained large audio language model
Generates original ARC-AGI-1-style tasks distribution-matched
Codex plugin that turns attached object images into code-only
An Open Real-time Video-Language Interaction System