A Family of Open Sourced Music Foundation Models
Generate Any 3D Scene in Seconds
Research code artifacts for Code World Model (CWM)
Inference code for scalable emulation of protein equilibrium ensembles
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
Foundation Models for Time Series
Hackable and optimized Transformers building blocks
tiktoken is a fast BPE tokeniser for use with OpenAI's models
RGBD video generation model conditioned on camera input
Recovering the Visual Space from Any Views
OpenTinker is an RL-as-a-Service infrastructure for foundation models
Unified Multimodal Understanding and Generation Models
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Infinite Worlds with Versatile Interactions
Tiny vision language model
Repo for SeedVR2 & SeedVR
Open-source large language model family from Tencent Hunyuan
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Implementation of the Surya Foundation Model for Heliophysics
Miso TTS is an 8 billion, highly emotive text-to-speech model
scikit-learn compatible tabular foundation model
Audio foundation model excelling in audio understanding
An experimental version of DeepSeek model
Accurate × Fast × Comprehensive
An Open Real-time Video-Language Interaction System