Skip to content

Paperless-ngx

Scan, index, and archive physical documents.

Self-hostedNot yet verified
Report issueDemo online
GPL-3.0

Open the official app on docs.paperless-ngx.com

This tool is hosted by its maintainers. Click below to open docs.paperless-ngx.com in a new tab — it's their official demo.

Browse pdf & tools tools →

What's next with Paperless-ngx?

Choose how you want to get started.

Use it free

Open the official tool or demo — no account needed.

Free

Self-host it

Run the open-source version on your own infrastructure.

Open

What is Paperless-ngx?

Paperless-ngx is an open-source document management system designed to digitize, index, and archive physical documents. It enables users to scan paper documents, extract text via OCR, and organize files using metadata tags for efficient retrieval. The tool is primarily used by individuals, small businesses, and organizations needing to transition from paper-based workflows to digital storage. It solves problems related to document disorganization, manual data entry, and the physical storage of paper records. By automating indexing and providing a centralized repository, Paperless-ngx streamlines document management for users who require long-term archival solutions or frequent access to historical records.

How it works

Paperless-ngx is a community-maintained project that transforms physical documents into searchable digital files. It leverages optical character recognition (OCR) to convert scanned images into editable text, enabling full-text search across documents. The tool's primary purpose is to replace traditional paper storage with a structured digital archive. It supports multi-user access, version control, and integration with external tools like email clients and document editors. Paperless-ngx excels at automating document processing tasks. It can scan and index PDFs, images, and other file formats, assigning metadata such as document type, date, and tags. Users can search across all indexed content using natural language queries.

How to use it

  1. 1Install the Docker container via the official repository or package manager. 2. Configure the OCR engine and connect external tools like email clients. 3. Scan physical documents using a connected scanner or upload files manually. 4. Use the web interface to tag, search, and organize indexed documents. Practical tips: Regularly back up the database, use consistent naming conventions for files, and leverage the built-in search filters to refine results.

What it can do

  • document archive

Use cases

Assumptions and limitations

Assumptions

  • source: https://github.com/paperless-ngx/paperless-ngx
  • license: GPL-3.0 — free to use
  • privacy: Self-hosted — you control your data

Limitations

  • Requires technical expertise for Docker setup and configuration
  • OCR accuracy may vary with low-quality scans or non-standard fonts
  • Lacks built-in collaboration features for real-time document editing
  • Limited native support for large-scale enterprise integration
  • Depends on external services like Tesseract for OCR processing

Understanding the result

Scan, index, and archive physical documents.

Tool details

  • Clearly flagged when a network request is needed.
  • No account, no sign-up, and no tracking of your content.
  • Powered by (GPL-3.0).
Built with
(paperless-ngx/paperless-ngx)
License
GPL-3.0
Runs locally
No — requires a network request
Verification
Not yet verified
Input
Query
Output
Text
Open-source source & license

Built with paperless-ngx/paperless-ngx. OpenToolVault provides the discovery and browser interface while crediting the original project maintainers.

Built with
License
GPL-3.0
View source on GitHub

Open-source project

OpenToolVault is an independent directory. We are not affiliated with or endorsed by this project.

References

Frequently asked

What licensing model does Paperless-ngx use?

Paperless-ngx is distributed under the GNU General Public License version 3.0 (GPL-3.0), allowing users to freely use, modify, and distribute the software while requiring derivative works to maintain open-source compliance.

How does Paperless-ngx handle document indexing?

The tool uses OCR to convert scanned documents into searchable text. Users can manually tag files with metadata, or automate tagging via rules. Full-text search is enabled through integrated search engines, allowing queries across all indexed content.

How do I scan and index documents?

Connect a scanner to your system, then use the web interface to upload files or trigger scans. Paperless-ngx processes each document, extracts text, and prompts users to assign tags. Alternatively, automate the process by integrating with email clients to auto-index incoming documents.

How does Paperless-ngx compare to alternatives like Mendeley or Evernote?

Unlike Mendeley (focused on academic research) or Evernote (general note-taking), Paperless-ngx specializes in document archiving with advanced OCR and metadata tagging. It lacks collaborative editing features but excels in long-term archival and compliance-focused workflows.

What should I do if OCR fails to recognize text?

Retry the scan with higher resolution settings, use a different OCR engine, or manually edit the extracted text. For recurring issues, verify that the document format is supported and adjust OCR configuration parameters in the settings.

Spotted something wrong with Paperless-ngx, or want to maintain it? See how to help.