private-gpt

Interact with your documents using the power of GPT, 100% privately, no data leaks

24 articles 57.2k View on GitHub ↗
24 articles
Troubleshooting Poor RAG Retrieval Results in PrivateGPT: A Complete Guide

Troubleshoot poor RAG retrieval in PrivateGPT. Learn to fix mismatched embedding models, optimize similarity_top_k, manage doc_id metadata, and re-ingest vector indices for better results.

how-to-guide
Mar 6, 2026
Handling Large Document Ingestion Using Parallel Processing Modes in PrivateGPT

Discover how to efficiently process massive document collections with PrivateGPT"s parallel processing modes. Optimize your ingestion scaling for seamless large document handling.

how-to-guide
Mar 6, 2026
How to Configure Similarity Thresholds for Filtering RAG Retrieval Results in PrivateGPT

Configure PrivateGPT similarity thresholds to filter RAG retrieval results. Learn how similarity_top_k and similarity_value enhance relevancy and optimize your chatbot's performance.

how-to-guide
Mar 6, 2026
How to Use the PrivateGPT Low-Level Embeddings API for Custom RAG Pipelines

Learn to use PrivateGPT's low-level embeddings API to create custom RAG pipelines. Integrate dense vectors into your bespoke generation workflows via the OpenAI-compatible endpoint.

how-to-guide
Mar 6, 2026
PrivateGPT Dependency Injection Architecture: How Components Are Wired Together

Explore PrivateGPTs dependency injection architecture. Learn how its three-layer system using the injector library wires components together for efficient runtime object management.

architecture
Mar 6, 2026
How PrivateGPT's File Watcher Automates Document Ingestion in Real-Time

Discover how PrivateGPT's file watcher automates real-time document ingestion by automatically triggering Llama-Index to embed and store new files in the vector database.

internals
Mar 6, 2026
How to Configure Node Storage Backends in PrivateGPT: Simple File vs PostgreSQL

Configure PrivateGPT node storage backends easily. Learn to set up simple file storage or PostgreSQL for your Zylon AI private-gpt repository. Maximize data control.

how-to-guide
Mar 6, 2026
How to Configure Prompt Styles in Private-GPT: Llama 2, Llama 3, Mistral, ChatML, and Tag

Learn to configure prompt styles in PrivateGPT, including Llama 2, Llama 3, Mistral, ChatML, and Tag. Format your chat messages effectively for LLM compatibility.

how-to-guide
Mar 6, 2026
How to Enable and Implement Streaming Responses for Chat and Completion Endpoints in PrivateGPT

Enable streaming responses in PrivateGPT for chat and completion endpoints. Learn how to use SSE for incremental token delivery instead of full JSON payloads. Optimize your applications today.

how-to-guide
Mar 6, 2026
How the Health Check Endpoint Functions in PrivateGPT for System Monitoring

Learn how the PrivateGPT health check endpoint at /health monitors your system. This simple GET route confirms API server responsiveness with a quick JSON status check.

how-to-guide
Mar 6, 2026
How to Implement Context Filtering in Private-GPT to Restrict RAG Responses to Specific Documents

Secure your RAG responses with Private-GPT context filtering. Learn how to restrict retrieval to specific documents using the ContextFilter for enhanced privacy and control.

how-to-guide
Mar 6, 2026
How to Tune LLM Generation Parameters in PrivateGPT: Temperature, Top-k, Top-p, and Repeat Penalty

Master LLM generation parameters in PrivateGPT. Learn to tune temperature, max_tokens, top_k, top_p, and repeat_penalty for optimal AI output by editing settings models.

how-to-guide
Mar 6, 2026

Have a question about this repo?

These articles cover the highlights, but your codebase questions are specific. Give your agent direct access to the source. Share this with your agent to get started:

Share the following with your agent to get started:
curl -s "https://instagit.com/install.md"

Works with
Claude Codex Cursor VS Code OpenClaw Any MCP Client

Maintain an open-source project? Get it listed too →