Paperless-ngx
Scan, index, and archive all your paper documents. OCR, full-text search.
Open the official app on docs.paperless-ngx.com
This tool is hosted by its maintainers. Click below to open docs.paperless-ngx.com in a new tab — it's their official demo.
Browse text & tools tools →What's next with Paperless-ngx?
Choose how you want to get started.
Use it free
Open the official tool or demo — no account needed.
Self-host it
Run the open-source version on your own infrastructure.
What is Paperless-ngx?
Paperless-ngx is an open-source document management system designed to digitize, index, and archive physical documents. It enables users to scan paper documents, extract text using OCR, and organize files through metadata tagging. Targeted at small businesses, legal professionals, and researchers, the tool addresses the challenge of managing physical paperwork by creating a centralized digital repository. By automating document processing, it reduces manual data entry and improves accessibility for long-term record-keeping.
How it works
Paperless-ngx is a community-supported platform that transforms physical documents into searchable digital archives. It integrates scanning, optical character recognition (OCR), and metadata tagging to streamline document management. The tool is particularly useful for organizations needing to comply with record-keeping regulations while reducing reliance on paper. Its modular design allows customization for diverse workflows. Core features include batch scanning with OCR support for PDFs, manual metadata tagging for documents, and integration with Docker for deployment. Users can search across indexed content and export data in structured formats.
How to use it
- 1Install the Docker container or run the Python backend. 2. Use the built-in scanner or import scanned PDFs. 3. Apply OCR to extract text and assign metadata tags. 4. Archive documents in a structured database for retrieval. Practical tips: Regularly back up the database, use consistent tagging conventions, and leverage the Ansible playbooks for automated setup.
What it can do
- Document Management
- Search
Use cases
Assumptions and limitations
Assumptions
- source: https://github.com/paperless-ngx
- license: GPL-3.0 — free to use
- privacy: Opens an external demo
Limitations
- Requires technical expertise for Docker setup and configuration
- Limited native support for non-PDF scanned documents
- Manual metadata tagging can be time-consuming for large archives
- No built-in collaboration features for team editing
- Depends on external tools for advanced search capabilities
Understanding the result
Scan, index, and archive all your paper documents. OCR, full-text search.
Tool details
- Clearly flagged when a network request is needed.
- No account, no sign-up, and no tracking of your content.
- Powered by paperless-ngx (GPL-3.0).
- Built with
- paperless-ngx (https://github.com/paperless-ngx)
- License
- GPL-3.0
- Runs locally
- No — requires a network request
- Verification
- Not yet verified
- Input
- Document, Text
- Output
- Document
Built with https://github.com/paperless-ngx. OpenToolVault provides the discovery and browser interface while crediting the original project maintainers.
- Built with
- paperless-ngx
- License
- GPL-3.0
Open-source project
OpenToolVault is an independent directory. We are not affiliated with or endorsed by this project.
References
- /paperless-ngx — GitHub Repository
Upstream project · GitHub
- GPL-3.0 License
Upstream project
Frequently asked
How does Paperless-ngx handle scanned documents?
The tool uses OCR to convert scanned PDFs into searchable text. Users can import scanned files, and the system automatically extracts text for indexing. Manual correction of OCR results is supported through the web interface.
How does the document indexing process work?
After scanning, documents are processed through OCR to create searchable text. Users then assign metadata tags (e.g., date, subject, author) to categorize files. The system builds a database of indexed content, allowing full-text searches across all documents.
How do I set up Paperless-ngx on a server?
Install Docker and run the Paperless-ngx container using the provided Dockerfile. Alternatively, deploy the Python backend with PostgreSQL. Configure the database settings, set up email notifications, and import initial documents via the web interface or API.
How does Paperless-ngx compare to Mendeley or Zotero?
Unlike Mendeley/Zotero, which focus on academic research management, Paperless-ngx is designed for general document archiving. It lacks citation management features but offers superior OCR and full-text search capabilities for scanned documents. It also supports broader file formats and custom metadata fields.
What should I do if the OCR fails to recognize text?
First, check that the scanned document is clear and properly formatted. If OCR errors persist, manually edit the extracted text through the web interface. For complex layouts, consider using PDFs with embedded text instead of scanned images. Retrying with a different OCR engine may also help.