Run Local LLMs on Any Device. Open-source
Oobabooga - The definitive Web UI for local AI, with powerful features
Redundancy-aware KV Cache Compression for Reasoning Models
Open-source model for program synthesis
Multimodal embedding and reranking models built on Qwen3-VL
A lightning fast audio upsampler
Z80-μLM is a 2-bit quantized language model
Fast and accurate AI powered file content types detection
Tool for exploring and debugging transformer model behaviors
Personalize Any Characters with a Scalable Diffusion Transformer
State-of-the-art Parameter-Efficient Fine-Tuning
A course of learning LLM inference serving on Apple Silicon
DeepSeek Coder: Let the Code Write Itself
Designed for text embedding and ranking tasks
Python SDK for the Computer Use model Lux, developed by OpenAGI
Jazzy theme for Django
Google Flights MCP and Python Library
An MCP server for interacting with Google Colab
Trainable, memory-efficient, and GPU-friendly PyTorch reproduction
The official repo of Qwen chat & pretrained large language model
Convert Google Gemini web into OpenAI-compatible API
Self-host the powerful Chatterbox TTS model
Supercharge Your Model Training
A simple, secure MCP-to-OpenAPI proxy server
Low-code framework for building custom LLMs, neural networks