Skip to content

Paperless-ngx

Scan, index, and archive all your paper documents. OCR, full-text search.

Self-hostedNot yet verified
Demo online
GPL-3.0★ 44184

Open the official app on docs.paperless-ngx.com

This tool is hosted by its maintainers. Click below to open docs.paperless-ngx.com in a new tab — it's their official demo.

Browse text & tools tools →

What's next with Paperless-ngx?

Choose how you want to get started.

Use it free

Open the official tool or demo — no account needed.

Free

Self-host it

Run the open-source version on your own infrastructure.

Open

What is Paperless-ngx?

Paperless-ngx is an open-source document management system designed to digitize, index, and archive physical documents. It enables users to scan paper documents, extract text using OCR, and organize files through metadata tagging. Targeted at small businesses, legal professionals, and researchers, the tool addresses the challenge of managing physical paperwork by creating a centralized digital repository. By automating document processing, it reduces manual data entry and improves accessibility for long-term record-keeping.

How it works

Paperless-ngx is a community-supported platform that transforms physical documents into searchable digital archives. It integrates scanning, optical character recognition (OCR), and metadata tagging to streamline document management. The tool is particularly useful for organizations needing to comply with record-keeping regulations while reducing reliance on paper. Its modular design allows customization for diverse workflows. Core features include batch scanning with OCR support for PDFs, manual metadata tagging for documents, and integration with Docker for deployment. Users can search across indexed content and export data in structured formats.

How to use it

  1. 1Install the Docker container or run the Python backend. 2. Use the built-in scanner or import scanned PDFs. 3. Apply OCR to extract text and assign metadata tags. 4. Archive documents in a structured database for retrieval. Practical tips: Regularly back up the database, use consistent tagging conventions, and leverage the Ansible playbooks for automated setup.

What it can do

  • Document Management
  • Search

Use cases

Assumptions and limitations

Assumptions

  • source: https://github.com/paperless-ngx
  • license: GPL-3.0 — free to use
  • privacy: Opens an external demo

Limitations

  • Requires technical expertise for Docker setup and configuration
  • Limited native support for non-PDF scanned documents
  • Manual metadata tagging can be time-consuming for large archives
  • No built-in collaboration features for team editing
  • Depends on external tools for advanced search capabilities

Understanding the result

Scan, index, and archive all your paper documents. OCR, full-text search.

Tool details

  • Clearly flagged when a network request is needed.
  • No account, no sign-up, and no tracking of your content.
  • Powered by paperless-ngx (GPL-3.0).
Built with
paperless-ngx (https://github.com/paperless-ngx)
License
GPL-3.0
Runs locally
No — requires a network request
Verification
Not yet verified
Input
Document, Text
Output
Document
Open-source source & license

Built with https://github.com/paperless-ngx. OpenToolVault provides the discovery and browser interface while crediting the original project maintainers.

Built with
paperless-ngx
License
GPL-3.0
View source on GitHub

Open-source project

OpenToolVault is an independent directory. We are not affiliated with or endorsed by this project.

References

Frequently asked

How does Paperless-ngx handle scanned documents?

The tool uses OCR to convert scanned PDFs into searchable text. Users can import scanned files, and the system automatically extracts text for indexing. Manual correction of OCR results is supported through the web interface.

How does the document indexing process work?

After scanning, documents are processed through OCR to create searchable text. Users then assign metadata tags (e.g., date, subject, author) to categorize files. The system builds a database of indexed content, allowing full-text searches across all documents.

How do I set up Paperless-ngx on a server?

Install Docker and run the Paperless-ngx container using the provided Dockerfile. Alternatively, deploy the Python backend with PostgreSQL. Configure the database settings, set up email notifications, and import initial documents via the web interface or API.

How does Paperless-ngx compare to Mendeley or Zotero?

Unlike Mendeley/Zotero, which focus on academic research management, Paperless-ngx is designed for general document archiving. It lacks citation management features but offers superior OCR and full-text search capabilities for scanned documents. It also supports broader file formats and custom metadata fields.

What should I do if the OCR fails to recognize text?

First, check that the scanned document is clear and properly formatted. If OCR errors persist, manually edit the extracted text through the web interface. For complex layouts, consider using PDFs with embedded text instead of scanned images. Retrying with a different OCR engine may also help.

Spotted something wrong with Paperless-ngx, or want to maintain it? See how to help.