PyTorch extensions for fast R&D prototyping and Kaggle farming
A lightweight vision library for performing large object detection
The Triton Inference Server provides an optimized cloud
Library for serving Transformers models on Amazon SageMaker
Private AI platform for agents, enterprise search and RAG pipelines
RAGFlow is an open-source RAG (Retrieval-Augmented Generation) engine
SGLang is a fast serving framework for large language models
Automate native Android apps with AI using accessibility APIs
ID-based RAG FastAPI: Integration with Langchain and PostgreSQL
Letta (formerly MemGPT) is a framework for creating LLM services
Build and run agents you can see, understand and trust
GitLab automatic code review tool based on large models
Diversity-driven optimization and large-model reasoning ability
This repository provides an advanced RAG
An MCP server that autonomously evaluates web applications
Get started w/ building Fullstack Agents using Gemini 2.5 & LangGraph
Repo of Qwen2-Audio chat & pretrained large audio language model
Lightweight Python library for adding real-time multi-object tracking
MII makes low-latency and high-throughput inference possible
Multi-Modal Neural Networks for Semantic Search, based on Mid-Fusion
A unified framework for scalable computing
The library to build & auto-optimize LLM applications
Capable of understanding text, audio, vision, video
HexStrike AI MCP Agents is an advanced MCP server
A Powerful web scraper powered by LLM | OpenAI, Gemini & Ollama