A series of math-specific large language models of our Qwen2 series
Towards self-verifiable mathematical reasoning
Python inference and LoRA trainer package for the LTX-2 audio–video
Strong, Economical, and Efficient Mixture-of-Experts Language Model
Kimi K2 is the large language model series developed by Moonshot AI
Open-source large language model family from Tencent Hunyuan
Qwen3-VL, the multimodal large language model series by Alibaba Cloud
Scaling Reinforcement Learning with LLMs
Pushing the Limits of Mathematical Reasoning in Open Language Models
Advancing Formal Mathematical Reasoning via Reinforcement Learning
Open-weight, large-scale hybrid-attention reasoning model
DeepSeek LLM: Let there be answers
800,000 step-level correctness labels on LLM solutions to MATH problem
llama.go is like llama.cpp in pure Golang
Compact 3B-param multimodal model for efficient on-device reasoning
Efficient 8B multimodal model tuned for advanced reasoning tasks.
High-precision 14B multimodal model built for advanced reasoning tasks
Instruction-tuned 7B language model for chat and complex tasks
QwQ-32B is a reasoning-focused language model for complex tasks
Efficient MoE reasoning model for coding and math workloads
Hermes 4 FP8: hybrid reasoning Llama-3.1-405B model by Nous Research
Powerful 14B LLM with strong instruction and long-text handling
VaultGemma: 1B DP-trained Gemma variant for private NLP tasks