A mcp server for vikingdb store and search
The Triton Inference Server provides an optimized cloud
Control Gmail, Google Calendar, Docs, Sheets, Slides, Chat, Forms
Lemonade helps users run local LLMs with the highest performance
High-performance inference server for text embeddings models API layer
Large Language Model Text Generation Inference
Fast stable diffusion on CPU and AI PC
Personal AI, On Personal Devices
Fantasy Premier League MCP Server
TokenSpeed is a speed-of-light LLM inference engine
GPU accelerated decision optimization
OCR model for complex documents with layout-aware structured outputs
TensorRT LLM provides users with an easy-to-use Python API
Arcade Tool Development Kit (TDK), Worker, Evals, and CLI
An MCP server for interacting with Google Colab
TensorFlow is an open source library for machine learning
Voice Recognition to Text Tool
A lightweight text-to-speech model with zero-shot voice cloning
Official inference framework for 1-bit LLMs
Supercharge Your LLM with the Fastest KV Cache Layer
A text-to-speech, speech-to-text and speech-to-speech library
OCR expert VLM powered by Hunyuan's native multimodal architecture
Build cross-modal and multimodal applications on the cloud
GPT4V-level open-source multi-modal model based on Llama3-8B
Hunyuan Translation Model Version 1.5