private-gpt
Interact with your documents using the power of GPT, 100% privately, no data leaks
Troubleshoot poor RAG retrieval in PrivateGPT. Learn to fix mismatched embedding models, optimize similarity_top_k, manage doc_id metadata, and re-ingest vector indices for better results.
Handling Large Document Ingestion Using Parallel Processing Modes in PrivateGPTDiscover how to efficiently process massive document collections with PrivateGPT"s parallel processing modes. Optimize your ingestion scaling for seamless large document handling.
How to Configure Similarity Thresholds for Filtering RAG Retrieval Results in PrivateGPTConfigure PrivateGPT similarity thresholds to filter RAG retrieval results. Learn how similarity_top_k and similarity_value enhance relevancy and optimize your chatbot's performance.
How to Use the PrivateGPT Low-Level Embeddings API for Custom RAG PipelinesLearn to use PrivateGPT's low-level embeddings API to create custom RAG pipelines. Integrate dense vectors into your bespoke generation workflows via the OpenAI-compatible endpoint.
PrivateGPT Dependency Injection Architecture: How Components Are Wired TogetherExplore PrivateGPTs dependency injection architecture. Learn how its three-layer system using the injector library wires components together for efficient runtime object management.
How PrivateGPT's File Watcher Automates Document Ingestion in Real-TimeDiscover how PrivateGPT's file watcher automates real-time document ingestion by automatically triggering Llama-Index to embed and store new files in the vector database.
How to Configure Node Storage Backends in PrivateGPT: Simple File vs PostgreSQLConfigure PrivateGPT node storage backends easily. Learn to set up simple file storage or PostgreSQL for your Zylon AI private-gpt repository. Maximize data control.
How to Configure Prompt Styles in Private-GPT: Llama 2, Llama 3, Mistral, ChatML, and TagLearn to configure prompt styles in PrivateGPT, including Llama 2, Llama 3, Mistral, ChatML, and Tag. Format your chat messages effectively for LLM compatibility.
How to Enable and Implement Streaming Responses for Chat and Completion Endpoints in PrivateGPTEnable streaming responses in PrivateGPT for chat and completion endpoints. Learn how to use SSE for incremental token delivery instead of full JSON payloads. Optimize your applications today.
How the Health Check Endpoint Functions in PrivateGPT for System MonitoringLearn how the PrivateGPT health check endpoint at /health monitors your system. This simple GET route confirms API server responsiveness with a quick JSON status check.
How to Implement Context Filtering in Private-GPT to Restrict RAG Responses to Specific DocumentsSecure your RAG responses with Private-GPT context filtering. Learn how to restrict retrieval to specific documents using the ContextFilter for enhanced privacy and control.
How to Tune LLM Generation Parameters in PrivateGPT: Temperature, Top-k, Top-p, and Repeat PenaltyMaster LLM generation parameters in PrivateGPT. Learn to tune temperature, max_tokens, top_k, top_p, and repeat_penalty for optimal AI output by editing settings models.
Have a question about this repo?
These articles cover the highlights, but your codebase questions are specific. Give your agent direct access to the source. Share this with your agent to get started:
curl -s "https://instagit.com/install.md" Maintain an open-source project? Get it listed too →