FAIR Sequence Modeling Toolkit 2
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
Audio foundation model excelling in audio understanding
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training
Chinese and English multimodal conversational language model
The Clay Foundation Model - An open source AI model and interface
Ling-V2 is a MoE LLM provided and open-sourced by InclusionAI
NVIDIA Isaac GR00T N1.5 is the world's first open foundation model
Foundation model for image generation
A Pragmatic VLA Foundation Model
Collection of Gemma 3 variants that are trained for performance
Open-weight, large-scale hybrid-attention reasoning model
Diversity-driven optimization and large-model reasoning ability
Open-source deep-learning framework
Open-Source Financial Large Language Models
Open-source framework for intelligent speech interaction
CogView4, CogView3-Plus and CogView3(ECCV 2024)
Phi-3.5 for Mac: Locally-run Vision and Language Models
Pokee Deep Research Model Open Source Repo
The official PyTorch implementation of Google's Gemma models
Implementation of the Surya Foundation Model for Heliophysics
OpenTinker is an RL-as-a-Service infrastructure for foundation models
Multimodal embedding and reranking models built on Qwen3-VL
Stable Virtual Camera: Generative View Synthesis with Diffusion Models
Multi-modal large language model designed for audio understanding