Phi-3.5 for Mac: Locally-run Vision and Language Models
Tiny vision language model
Repo for SeedVR2 & SeedVR
AI cognitive-enhancement Skills based on Anthropic's J-space
A theoretical reconstruction of the Claude Mythos architecture
Video Object and Interaction Deletion
Z80-μLM is a 2-bit quantized language model
High-resolution models for human tasks
Pretrained time-series foundation model developed by Google Research
Open-Source Financial Large Language Models
PyTorch code and models for the DINOv2 self-supervised learning
Open-source large language model family from Tencent Hunyuan
A Multi-Modal World Model for Reconstructing, Generating, Simulation
Unified Multimodal Understanding and Generation Models
GLM-4-Voice | End-to-End Chinese-English Conversational Model
DeepSeek Coder: Let the Code Write Itself
A series of math-specific large language models of our Qwen2 series
Tongyi Deep Research, the Leading Open-source Deep Research Agent
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
scikit-learn compatible tabular foundation model
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
Audio Language Models are Few-Shot Learners
Open image model at the forefront of design
Collection of Gemma 3 variants that are trained for performance
Video understanding codebase from FAIR for reproducing video models