The easy-to-use Vue low-code visual AI form designer
AI framework for automated short video creation and editing tools
Pushing the Frontier of Long Audio-Visual Generation
OCR expert VLM powered by Hunyuan's native multimodal architecture
Beyond the Imitation Game collaborative benchmark for measuring
The ultimate tool to automate custom telegram message forwarding
Node.js module for rendering pdf pages to images, svgs and HTML files
A PyTorch implementation of "Capsule Graph Neural Network"
Qwen2.5-VL-3B-Instruct: Multimodal model for chat, vision & video
Powerful 14B LLM with strong instruction and long-text handling
Efficient 8B multimodal model tuned for advanced reasoning tasks.
High-precision 14B multimodal model built for advanced reasoning tasks
Efficient 14B multimodal instruct model with edge deployment and FP8
Compact 3B-param multimodal model for efficient on-device reasoning
QwQ-32B is a reasoning-focused language model for complex tasks
Multimodal 7B model for image, video, and text understanding tasks