GPT4V-level open-source multi-modal model based on Llama3-8B
The only cheat sheet you need
Personal mini-web in text
Speech-AI-Forge is a project developed around TTS generation model
Fully open source deep research agent
Open source RAG framework for building scalable modular AI apps
Open-Source Dual-Arm Mobile Robot with Motorized Lift
Enables the best performance on NVIDIA RTX Graphics Cards
Collaborative & Open-Source Quality Assurance for all AI models
Chatbot daemon that connects to your favorite chat services
Static site generator that supports Markdown and reST syntax
dj-stripe automatically syncs your Stripe Data to your local database
MCP integration platforms for AI agents to use tools at any scale
A multi-function Discord bot
Official inference framework for 1-bit LLMs
Stable Diffusion web UI
Ansible for DevOps examples
Automatic SSRF fuzzer and exploitation tool
tensorboard for pytorch (and chainer, mxnet, numpy, etc.)
OCR expert VLM powered by Hunyuan's native multimodal architecture
An SSH/Telnet/Serial client in your browser
Request recommended movies, TV shows and anime to Jellyseer/Overseer
A distributed and persistent archive replay system using IPFS
Build cross-modal and multimodal applications on the cloud
GUI shell for running local LLM on desktop