LocalGPT is a private, on-premises document intelligence platform for questioning, summarizing, and analyzing files with locally hosted language models. Its data remains on the user’s machine, making it suitable for confidential or offline workflows. The retrieval system combines semantic similarity, keyword matching, late chunking, contextual enrichment, and sentence-level pruning. A smart router chooses between retrieval-augmented generation and direct model responses for each query. An independent verification pass is intended to improve answer reliability before results are returned. The platform works with Ollama models, offers a browser interface and API, and can run through local or Docker-based setups. It supports CPU and several hardware acceleration environments while retaining conversation history within a session.

Features

  • Private on-device document processing
  • Hybrid semantic and keyword retrieval
  • Automatic RAG or direct-answer routing
  • Context enrichment and pruning
  • Local browser interface and API
  • CPU, GPU, HPU, and MPS support

Project Samples

Project Activity

See All Activity >

Categories

Libraries

License

MIT License

Follow LocalGPT

LocalGPT Web Site

Other Useful Business Software
Host LLMs in Production With On-Demand GPUs Icon
Host LLMs in Production With On-Demand GPUs

NVIDIA L4 GPUs. 5-second cold starts. Scale to zero when idle.

Deploy your model, get an endpoint, pay only for compute time. No GPU provisioning or infrastructure management required.
Try Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of LocalGPT!

Additional Project Details

Programming Language

Python

Related Categories

Python Libraries

Registered

6 days ago