OCR expert VLM powered by Hunyuan's native multimodal architecture
Qwen2.5-VL-3B-Instruct: Multimodal model for chat, vision & video
Powerful 14B LLM with strong instruction and long-text handling
Efficient 8B multimodal model tuned for advanced reasoning tasks.
High-precision 14B multimodal model built for advanced reasoning tasks
Efficient 14B multimodal instruct model with edge deployment and FP8
Compact 3B-param multimodal model for efficient on-device reasoning
QwQ-32B is a reasoning-focused language model for complex tasks
Multimodal 7B model for image, video, and text understanding tasks