Lightweight multimodal translation model for 55 languages
Instruction-tuned 7B language model for chat and complex tasks
Large-scale xAI model for local inference with SGLang, Grok-2.5
BGE-Large v1.5: High-accuracy English embedding model for retrieval
Portuguese ASR model fine-tuned on XLSR-53 for 16kHz audio input
Efficient English embedding model for semantic search and retrieval
Vision-language-action model for robot control via images and text
Russian ASR model fine-tuned on Common Voice and CSS10 datasets
4-engine reverse image search that works on Instagram & Pinterest
Custom BLEURT model for evaluating text similarity using PyTorch
Multimodal Transformer for document image understanding and layout
Compact English sentence embedding model for semantic search tasks
Flexible text-to-text transformer model for multilingual NLP tasks
Qwen2.5-VL-3B-Instruct: Multimodal model for chat, vision & video
Summarization model fine-tuned on CNN/DailyMail articles
CTC-based forced aligner for audio-text in 158 languages
Multimodal 7B model for image, video, and text understanding tasks
Powerful 14B LLM with strong instruction and long-text handling
ClinicalBERT model trained on MIMIC notes for clinical NLP tasks