Minimal Python framework for scalable AI inference servers fast
Data Infrastructure providing an approach to multimodal AI workloads
Connect any LLM to your internal knowledge sources
High accuracy RAG for answering questions from scientific documents
An Easy-to-use, Scalable and High-performance RLHF Framework
AI-Powered Wiki Generator for GitHub/Gitlab/Bitbucket Repositories
Multi-modal large language model designed for audio understanding
Chat with your documents on your local device using GPT models
The beginning of scalable pixel-native search
The ultimate RAG for your monorepo
Context database designed specifically for AI Agents
RAG-Anything: All-in-One RAG Framework
Julia Devito inversion
High-performance inference server for text embeddings models API layer
A simple, easy-to-hack GraphRAG implementation
SimpleMem: Efficient Lifelong Memory for LLM Agents
High-Resolution 3D Assets Generation with Large Scale Diffusion Models
Central interface to connect your LLM's with external data
Fast State-of-the-Art Static Embeddings
Build production-ready AI agents in both Python and Typescript
Python library and shell utilities to monitor filesystem events
Low-latency AI inference engine optimized for mobile devices
Recovering the Visual Space from Any Views
Document content and metadata extraction microservice
"VideoRAG: Chat with Your Videos