High-performance inference server for text embeddings models API layer
Wan2.2: Open and Advanced Large-Scale Video Generative Model
MTEB: Massive Text Embedding Benchmark
Large-language-model & vision-language-model based on Linear Attention
Toolkit for conversational AI
Generate audiobooks from e-books
Open source libraries and APIs to build custom preprocessing pipelines
Create prompt-friendly codebase digests from any Git repository URL
Build AI-powered semantic search applications
Build cross-modal and multimodal applications on the cloud
Chinese and English multimodal conversational language model
Run the Stable Diffusion releases in a Docker container
A list of accessible speech corpora for ASR, TTS
Generates random text based on context-free grammars defined in BNF
fastNLP: A Modularized and Extensible NLP Framework
An NLP library for building bots
Easy-OCR solution and Tesseract trainer for GNU/Linux
FEM allows users to create fuzzy functional groups for use in ecology.
Edit the OCR text layer of DjVu documents in a web browser
The IRC's Talking Robot
Powerful, native multimodal AI agentic model
NVFP4 DiffusionGemma model for fast multimodal text generation
Efficient multimodal MoE model for coding, tools, and reasoning
4-bit Command A+ model for enterprise agents and multilingual tasks
Open multimodal model for coding, agents, and long-context tasks