MinerU
Transforms complex documents like PDFs into LLM-ready markdown/JSON for your Agentic workflows.
Deploy MinerU for production integration. Set up using Docker, FastAPI, OpenAI-compatible server, or Gradio to expose REST endpoints for document processing and retrieval.
MinerU Backend Source Files: Complete Guide to VLM, Pipeline, and Hybrid LocationsDiscover the location of MinerU backend source files within the mineru/backend/ directory. Explore VLM, Pipeline, and Hybrid sub-packages including utils.py.
How to Set the Language for OCR in MinerU: CLI, Python API, and HTTP MethodsEasily set the OCR language in MinerU using CLI, Python API, or HTTP requests. Learn how to specify languages for accurate text recognition with Pytorch-Paddle-OCR.
How to Specify the Parsing Method (Auto, TXT, OCR) in MinerUControl MinerU parsing with auto txt or ocr methods via CLI API or FastAPI. Learn how to specify your desired parsing method for efficient document analysis.
MinerU OmniDocBench Accuracy: Benchmark Results for Pipeline and VLM BackendsDiscover MinerU OmniDocBench accuracy. See how MinerU's pipeline and VLM backends outperform GPT-4o and Gemini 2.5 Pro with just 1.2B parameters. Get benchmark results now.
How MinerU Detects and Handles Hallucinations in PDF ParsingMinerU prevents PDF parsing hallucinations by classifying documents. Learn how it avoids false text generation for corrupted or image-heavy files. Read more.
How to Use MinerU in HTTP Client Mode with OpenAI-Compatible ServersLearn to use MinerU in HTTP client mode with OpenAI-compatible servers. Offload VLM inference to remote servers for efficient processing.
VRAM Requirements for MinerU's VLM Backend: Complete Hardware GuideDiscover MinerU VLM backend VRAM requirements. Learn the minimum 8 GB VRAM needed and the recommended 10 GB for full table acceleration. Get your hardware guide now.
How to Utilize MinerU's VLM Backend for High-Accuracy ParsingMaster PDF parsing with MinerU's VLM backend. Configure environment variables, choose an inference engine, and use MagicModel for accurate formula and table extraction, delivering structured Content-List V2 output.
How to Convert PDF to JSON Using MinerU: A Complete Technical GuideLearn to convert PDF to JSON with MinerU. This technical guide details the six-stage pipeline for structured data extraction from PDFs using MagicModel inference. Explore the opendatalab MinerU repository.
How to Convert PDF to Markdown Using MinerU: A Complete Technical GuideEasily convert PDF to Markdown with MinerU. This technical guide details MinerU's multi-stage pipeline for accurate document conversion to clean Markdown.
Complete Guide to Output Formats Supported by MinerUExplore MinerU's seven output formats: Multimodal Markdown, NLP Markdown, JSON, raw output, images, and ZIP. Discover the best format for your data analysis.
Have a question about this repo?
These articles cover the highlights, but your codebase questions are specific. Give your agent direct access to the source. Share this with your agent to get started:
curl -s "https://instagit.com/install.md" Maintain an open-source project? Get it listed too →