Official MiniMax Model Context Protocol (MCP) server
Speech-AI-Forge is a project developed around TTS generation model
OCR expert VLM powered by Hunyuan's native multimodal architecture
Miso TTS is an 8 billion, highly emotive text-to-speech model
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
An advanced paper search agent powered by large language models
GUI Exploration Lab. One of the best GUI agent solutions
Open-weight, large-scale hybrid-attention reasoning model
Large-language-model & vision-language-model based on Linear Attention
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training
LLM-based Reinforcement Learning audio edit model
Capable of understanding text, audio, vision, video
Open-source MCP server that gives your coding agent
MCP integration platforms for AI agents to use tools at any scale
Question and Answer based on Anything
AI agents running research on single-GPU nanochat training
Python package built to ease deep learning on graph
E2M converts various file types (doc, docx, epub, html, htm, url
Evaluate your LLM's response with Prometheus and GPT4
Open source framework for deep learning satellite and aerial imagery
Stable Diffusion built-in to Blender
Marrying Grounding DINO with Segment Anything & Stable Diffusion
A lightweight audio-to-MIDI converter with pitch bend detection
Solve puzzles. Learn CUDA
A Collection of Cheatsheets, Books, Questions, and Portfolio