Lightweight multimodal translation model for 55 languages
Rule-based information extraction.
Multimodal Transformer for document image understanding and layout
Layout-aware OCR model for multilingual document understanding
Qwen2.5-VL-3B-Instruct: Multimodal model for chat, vision & video
PMC browser
Document Management System
ClinicalBERT model trained on MIMIC notes for clinical NLP tasks
Small 3B-base multimodal model ideal for custom AI on edge hardware