Paperless-ngx
Scan, index, and archive physical documents.
Open the official app on docs.paperless-ngx.com
This tool is hosted by its maintainers. Click below to open docs.paperless-ngx.com in a new tab — it's their official demo.
Browse pdf & tools tools →What's next with Paperless-ngx?
Choose how you want to get started.
Use it free
Open the official tool or demo — no account needed.
Self-host it
Run the open-source version on your own infrastructure.
What is Paperless-ngx?
Paperless-ngx is an open-source document management system designed to digitize, index, and archive physical documents. It enables users to scan paper documents, extract text via OCR, and organize files using metadata tags for efficient retrieval. The tool is primarily used by individuals, small businesses, and organizations needing to transition from paper-based workflows to digital storage. It solves problems related to document disorganization, manual data entry, and the physical storage of paper records. By automating indexing and providing a centralized repository, Paperless-ngx streamlines document management for users who require long-term archival solutions or frequent access to historical records.
How it works
Paperless-ngx is a community-maintained project that transforms physical documents into searchable digital files. It leverages optical character recognition (OCR) to convert scanned images into editable text, enabling full-text search across documents. The tool's primary purpose is to replace traditional paper storage with a structured digital archive. It supports multi-user access, version control, and integration with external tools like email clients and document editors. Paperless-ngx excels at automating document processing tasks. It can scan and index PDFs, images, and other file formats, assigning metadata such as document type, date, and tags. Users can search across all indexed content using natural language queries.
How to use it
- 1Install the Docker container via the official repository or package manager. 2. Configure the OCR engine and connect external tools like email clients. 3. Scan physical documents using a connected scanner or upload files manually. 4. Use the web interface to tag, search, and organize indexed documents. Practical tips: Regularly back up the database, use consistent naming conventions for files, and leverage the built-in search filters to refine results.
What it can do
- document archive
Use cases
Assumptions and limitations
Assumptions
- source: https://github.com/paperless-ngx/paperless-ngx
- license: GPL-3.0 — free to use
- privacy: Self-hosted — you control your data
Limitations
- Requires technical expertise for Docker setup and configuration
- OCR accuracy may vary with low-quality scans or non-standard fonts
- Lacks built-in collaboration features for real-time document editing
- Limited native support for large-scale enterprise integration
- Depends on external services like Tesseract for OCR processing
Understanding the result
Scan, index, and archive physical documents.
Tool details
- Clearly flagged when a network request is needed.
- No account, no sign-up, and no tracking of your content.
- Powered by (GPL-3.0).
- Built with
- (paperless-ngx/paperless-ngx)
- License
- GPL-3.0
- Runs locally
- No — requires a network request
- Verification
- Not yet verified
- Input
- Query
- Output
- Text
Built with paperless-ngx/paperless-ngx. OpenToolVault provides the discovery and browser interface while crediting the original project maintainers.
- Built with
- License
- GPL-3.0
Open-source project
OpenToolVault is an independent directory. We are not affiliated with or endorsed by this project.
References
- / — GitHub Repository
Upstream project · GitHub
- GPL-3.0 License
Upstream project
Frequently asked
What licensing model does Paperless-ngx use?
Paperless-ngx is distributed under the GNU General Public License version 3.0 (GPL-3.0), allowing users to freely use, modify, and distribute the software while requiring derivative works to maintain open-source compliance.
How does Paperless-ngx handle document indexing?
The tool uses OCR to convert scanned documents into searchable text. Users can manually tag files with metadata, or automate tagging via rules. Full-text search is enabled through integrated search engines, allowing queries across all indexed content.
How do I scan and index documents?
Connect a scanner to your system, then use the web interface to upload files or trigger scans. Paperless-ngx processes each document, extracts text, and prompts users to assign tags. Alternatively, automate the process by integrating with email clients to auto-index incoming documents.
How does Paperless-ngx compare to alternatives like Mendeley or Evernote?
Unlike Mendeley (focused on academic research) or Evernote (general note-taking), Paperless-ngx specializes in document archiving with advanced OCR and metadata tagging. It lacks collaborative editing features but excels in long-term archival and compliance-focused workflows.
What should I do if OCR fails to recognize text?
Retry the scan with higher resolution settings, use a different OCR engine, or manually edit the extracted text. For recurring issues, verify that the document format is supported and adjust OCR configuration parameters in the settings.