ScanMole

Quelltext

Probleme (Issues)

Lizenzierung

Sonstiges

Dieses Projekt nutzt eine andere Arbeitssprache als Deutsch. Der folgende Abschnitt wurde automatisch aus der README-Datei generiert und liegt daher nur in der Originalsprache vor.

README

Easy, scriptable document scanning for Linux: ADF duplex batches in, searchable (OCRed) PDFs out.

It consists of two components, shipped as two Python packages, so servers and scripts can install the CLI alone while desktops get the whole experience:

  1. scanmole: CLI scanning engine.
  2. scanmole-gui: GTK4/libadwaita frontend (depends on scanmole). A thin subprocess wrapper around the CLI using its --json event protocol; it contains no scanning logic itself.


ScanMole logo: a mole with glasses holding a scanned document

⭐ Found this useful? Support open-source and star this project:

GitHub repository


Features

Main features:

  • Scan a stack of paper into one searchable PDF with a single command: duplex batch, blank backsides dropped, OCR text layer, archival PDF/A output by default.
  • Automatic page size detection crops every page to the paper’s real edges, so receipts come out receipt-sized and mixed stacks need no set-up.
  • Small files by default: 1-bit black-and-white at 300 dpi lands at roughly 100 KB per A4 text page, and ocrmypdf shrinks that further where jbig2enc is installed.
  • Works with anything SANE can drive, including driverless eSCL devices via sane-airscan. Device capabilities are probed and mapped instead of hardcoded, and devices without a native 1-bit mode get software binarization automatically.
  • Automation-grade CLI with defined exit codes, filename templates and a versioned JSON event protocol; interrupted batches can be rebuilt from the preserved page images without rescanning the paper.
  • Easy-to-use GTK4/libadwaita GUI on top of the same engine, with a live filename preview and translations (German included).

Demo

Screenshots

Screenshot: The ScanMole GUI with a connected ScanSnap iX500, ready to scan
Screenshot: The ScanMole GUI after a finished scan on a Brother ADS-4550W, with saved pages and a skipped blank in the result bar
Screenshot: The ScanMole GUI in its two-column layout with two scanners connected
Screenshot: The ScanMole GUI's settings dialog with color scheme, language and desktop integration
Screenshot: The ScanMole CLI listing devices and scanning a duplex batch to a searchable PDF

Installation

PyPI package version: scanmole
PyPI package version: scanmole-gui

ScanMole needs Python ≄ 3.12. Its two packages are available on PyPI: scanmole (the CLI) and scanmole-gui (the desktop frontend, pulls the CLI automatically).

Desktop (CLI + GUI), using uv (recommended): the GUI uses the distribution’s PyGObject/GTK (see the packages below), so its virtualenv must see the system site packages:

uv venv --system-site-packages ~/.venvs/scanmole
source ~/.venvs/scanmole/bin/activate
uv pip install scanmole-gui

Tip: after the first scanmole-gui start, the settings dialog can install a menu entry, so later starts come straight from the desktop’s application grid without any venv activation.

Server or scripting (CLI only): the CLI is pure stdlib and installs into any isolated environment:

uv tool install scanmole

Using pip or pipx instead of uv:

pipx install --system-site-packages scanmole-gui   # desktop
pip install scanmole                               # CLI only

For development installs from a repository checkout, see DEVELOPMENT.md.

ScanMole’s runtime shells out to external tools, which come from distribution packages. Install them as follows (on a CLI-only machine, the GTK/PyGObject packages at the end of each list can be skipped):

Debian/Ubuntu

Only Debian 13+ and Ubuntu 24.04+ are supported (older releases lack the required Python ≄ 3.12):

sudo apt install sane-utils sane-airscan img2pdf ocrmypdf \
                 tesseract-ocr tesseract-ocr-deu jbig2enc \
                 python3-gi gir1.2-gtk-4.0 gir1.2-adw-1

Fedora

sudo dnf install sane-backends sane-airscan img2pdf ocrmypdf \
                 tesseract tesseract-langpack-deu \
                 python3-gobject gtk4 libadwaita

For smaller PDFs, it is highly recommended to also install jbig2enc (ocrmypdf picks it up automatically). Fedora does not package it (last checked: Fedora 44, 2026-Q3; a leftover of the long-expired JBIG2 encoding patents), so build it from source:

sudo dnf install gcc-c++ automake libtool leptonica-devel zlib-devel
git clone https://github.com/agl/jbig2enc.git /tmp/jbig2enc
cd /tmp/jbig2enc
./autogen.sh && ./configure && make
sudo make install   # installs the jbig2 binary under /usr/local/bin

Updating works the same way; the leading line reuses the clone when it still exists and starts fresh otherwise (/tmp does not survive a reboot):

git -C /tmp/jbig2enc pull || git clone https://github.com/agl/jbig2enc.git /tmp/jbig2enc
cd /tmp/jbig2enc
./autogen.sh && ./configure && make
sudo make install

Device-specific notes

Brother

Modern Brother devices (e.g. the Brother ADS-4550W) work driverless via sane-airscan (eSCL) and need no additional packages or configuration beyond the dependencies above.

Older devices without eSCL support (e.g. the Brother ADS-2600W) need Brother’s proprietary brscan4/brscan5 driver packages from the Brother support site; network devices must additionally be registered with brsaneconfig4/brsaneconfig5.

ScanSnap

ScanSnap devices (e.g. the ScanSnap iX500; formerly sold under the Fujitsu brand, Ricoh products today) use the stock SANE fujitsu backend over USB. They need no additional packages or configuration beyond the dependencies above.

Usage

Command Line Interface (CLI)

scanmole --list-devices        # what SANE sees (webcams/v4l are ignored)
scanmole                       # ADF duplex, lineart, 300 dpi, auto size, deu+eng OCR
                               #   -> ./2026-08-15_scan_001.pdf (auto-numbered)
scanmole '{YYYY}-{MM}_scan_{NN}'       # template -> ./2026-08_scan_01.pdf
scanmole -o invoice.pdf --mode gray -r 300 -l deu+eng
scanmole --source flatbed --no-ocr --keep-blanks draft
scanmole --from-images pages/*.png -o rebuild.pdf   # pipeline without a scanner

Output names may contain placeholders, in the CLI and the GUI alike: {YYYY}, {MM}, {DD} (date), {hh}, {mm}, {ss} (time), {N}/{NN}/… (zero-padded auto-increment, bumped until the name is free) and {device}; the default is {YYYY}-{MM}-{DD}_scan_{NNN}.pdf.

Run scanmole --help for the full list of options. Common examples:

scanmole -r 300 'contract_{YYYY}-{MM}-{DD}_{NN}'  # higher dpi for small print
scanmole --source adf --keep-blanks               # single-sided stack, keep every page
scanmole --mode gray -l deu --no-pdfa notes       # grayscale, German-only OCR, plain PDF
scanmole --keep-images /tmp/pages -v receipts     # keep page images, verbose log

What if my scanner acts up, for example wrong page sizes in auto mode, surviving blank pages, or a badly mapped mode? Every device behaves a little differently at the edges of a scan, and we can usually fix it from a few captured files alone: see reporting scanner problems and device quirks for exactly what to include.

The GUI

Start the GUI from the environment set up in Installation:

uv run scanmole-gui

The settings dialog can install a menu entry (.desktop file) for your user, so later starts work straight from the desktop’s application grid.

scanmole-gui is a form over the same engine with the same defaults. It covers and presents the CLI features in an easy-to-use way. The Scan button turns into Cancel while a batch runs, a collapsible log shows the underlying CLI output, and a result bar opens the finished PDF or its folder. The GUI remembers the last used form values and the window size in ~/.config/scanmole/gui.json and restores them on the next start.

Exit codes

CodeMeaning
0Success: PDF written.
1Unexpected internal error.
2Usage or input error: bad arguments, invalid page size, conflicting options. No PDF was produced.
3Acquisition failure: scanimage failed, no usable device, device vanished mid-batch, or a device probe timed out.
4Missing external tool: scanimage, img2pdf or ocrmypdf is not installed.
5Processing failure: img2pdf or ocrmypdf failed after successful acquisition. Scanned pages are preserved in the work directory (path in the error message), so the batch can be rebuilt with --from-images instead of rescanning the paper.
6Nothing to scan: feeder empty, or every page was blank. Not a malfunction; no PDF was produced.
130Interrupted (SIGINT).
143Terminated (SIGTERM), e.g. a GUI cancel.

The --json protocol

One JSON object per line on stdout; human-readable log on stderr:

{"event":"hello","version":"1.0.0"}
{"event":"devices","devices":[{"device":"...","vendor":"...","model":"...","type":"..."}]}
{"event":"start","device":"...","source":"adf-duplex","mode":"lineart","resolution":300,"page_size":"a4","output":"..."}
{"event":"settings","device":"...","source":"ADF Duplex","mode":"Lineart","resolution":300}
{"event":"page","n":1,"file":"...","blank":false,"mean":0.87}
{"event":"scan_done","total":5,"kept":4,"blanks":1}
{"event":"ocr_start","lang":"deu"}
{"event":"done","output":"out.pdf","pages":4,"bytes":812345,"seconds":41.2}
{"event":"error","message":"...","code":3}

hello opens every --json run and carries the CLI version, which is also the API version (SemVer): a frontend and the CLI are compatible as long as their major versions match (from 1.0.0 on). start carries the requested settings; settings (scanner runs only) reports the values actually negotiated with the SANE backend. error.code mirrors the process exit code. This protocol, the option names and the exit codes are the compatibility boundary for any frontend or reimplementation; the authoritative definition is the CLI contract in ARCHITECTURE.md.

FAQ

What to do if my scanner is not listed?

If a USB scanner does not show up in scanimage -L:

  1. Check that the needed packages are installed (see Installation): sane-backends provides scanimage and sane-find-scanner, sane-airscan provides the driverless eSCL route and airscan-discover, usbutils provides lsusb.
  2. Check lsusb. If the scanner is missing there too, the problem is cabling, power or the USB port, not software.
  3. Run sane-find-scanner -q. It talks raw USB without any backend; if it finds the device while scanimage -L stays empty, the cause is permissions or a disabled backend.
  4. Permissions: SANE grants access to locally logged-in desktop users. After the first plug-in, replug the device and log out and in once so the udev ACLs apply. Over ssh or headless there is no desktop session (“works locally, fails over ssh”); that needs a udev rule granting access to a scanner group.
  5. Check /etc/sane.d/dll.conf: the line for your vendor’s backend must not be commented out (Canon pixma/canon_dr, Epson epsonds/epson2, ScanSnap fujitsu; the backend keeps its historic name).
  6. Network-capable devices from roughly 2015 on usually speak eSCL and work driverless via sane-airscan: make sure avahi-daemon is running and check what airscan-discover finds. Worth trying even when a device’s USB route fails.
  7. Vendor drivers (Canon scangearmp2, Epson epsonscan2, Brother brscan4/brscan5) are the last resort for devices without an in-tree backend or eSCL support.

What to do if my scanner isn’t working as expected?

See CONTRIBUTING.md: Report scanner problems and device quirks.

How can I optimize the PDF file size?

The defaults are already tuned for small files: 1-bit black-and-white (lineart) at 300 dpi compresses losslessly to roughly 100 KB per A4 text page. Devices without a native 1-bit mode need no special handling; ScanMole converts their gray output in software automatically.

To keep files small:

  1. Stay with the 300 dpi black-and-white default for usual documents. Use --mode gray or --mode color only when a document really needs it (photos, stamps, faint or colored originals): they store 8 or 24 bits per pixel instead of 1, and sizes explode. The same goes for resolutions above 300 dpi, since data grows quadratically with dpi.
  2. Use -r 200 for documents where quality matters less; it roughly halves the data. Where no text layer is needed either, --no-ocr skips OCR entirely.
  3. Highly recommended: make sure jbig2enc is installed. ocrmypdf detects it automatically during its optimization pass and recodes 1-bit pages losslessly to a fraction of their size. command -v jbig2 shows whether it is present; if it prints nothing, follow the installation instructions.

Why is a page missing from my PDF, or a blank page kept?

Duplex scanning reads both sides of every sheet, and ScanMole drops a page as blank when its mean brightness is above 0.995, i.e. when less than 0.5% of it is “ink”. That is what removes the empty backsides of single-sided documents. Both failure directions have knobs: if a page with faint content was dropped, raise --blank-threshold towards 1, or use --keep-blanks to keep every page (--blank-threshold 0 disables the detection entirely). If a truly blank page survives, something dark is pulling its mean down, typically punch holes, staple shadows or a skewed scan showing the scan-bed edge; if tuning the threshold does not fix it, report the device quirk.

Contributing

See CONTRIBUTING.md for how to report issues and submit changes. Translations are welcome.

This project’s functionality is mature, so there might be little activity on the repository in the future. Don’t get fooled by this, the project is under active maintenance and used on a daily basis by the maintainers.

Copyright (c) 2026 foundata GmbH (https://foundata.com)

This project is licensed under the GNU General Public License v3.0 or later (SPDX-License-Identifier: GPL-3.0-or-later), see LICENSES/GPL-3.0-or-later.txt for the full text.

The REUSE.toml file provides detailed licensing and copyright information in a human- and machine-readable format. This includes parts that may be subject to different licensing or usage terms, such as third-party components. The repository conforms to the REUSE specification. You can use reuse spdx to create a SPDX software bill of materials (SBOM).

REUSE status

Trademarks

  • ScanSnap is a trademark of PFU Limited, a Ricoh Group company (ScanSnap scanners were sold under the Fujitsu brand until 2023)
  • Fujitsu is a trademark of Fujitsu Limited
  • Brother is a trademark of Brother Industries, Ltd

Their use here is purely descriptive and does not imply any affiliation with or endorsement by the trademark holders.

Author information

This project was created and is maintained by foundata GmbH.