Language modeling in a sentence representation space
Qwen3-Coder is the code version of Qwen3
26m function call model that runs on incredibly small devices
gpt-oss-120b and gpt-oss-20b are two open-weight language models
ICLR2024 Spotlight: curation/training code, metadata, distribution
Convert Google Gemini web into OpenAI-compatible API
PyTorch code and models for the DINOv2 self-supervised learning
4M: Massively Multimodal Masked Modeling
Claude Code image, a one-stop open source transit service
Qwen-Image is a powerful image generation foundation model
Reference PyTorch implementation and models for DINOv3
Memory-efficient and performant finetuning of Mistral's models
Open Multilingual Multimodal Chat LMs
Chat & pretrained large audio language model proposed by Alibaba Cloud
Compact 3B-param multimodal model for efficient on-device reasoning
Compact 8B multimodal instruct model optimized for edge deployment
Efficient MoE model for reasoning, coding, and AI agent workflows
Flexible text-to-text transformer model for multilingual NLP tasks
Efficient 8B multimodal model tuned for advanced reasoning tasks.
High-precision 14B multimodal model built for advanced reasoning tasks
Ultra-efficient 3B multimodal instruct model built for edge deployment
NVFP4 DiffusionGemma model for fast multimodal text generation
Unified multimodal Gemma model for local coding and reasoning
Google’s flagship dense multimodal model for coding and reasoning
Hermes 4 FP8: hybrid reasoning Llama-3.1-405B model by Nous Research