Book Annotation Extractor
Normalize exported Kindle, Kobo, and CSV highlights with provenance, filters, dedup review, and portable reports.
The problem it solves
Book Annotation Extractor [](https://github.com/loganpendragonmultiverse/book-annotation-extractor/actions/workflows/ci.yml)
Who Book Annotation Extractor is for
- People who need a small, purpose-built utility instead of a broad platform.
- Users working with books, annotations, highlights who want the documented v1.1.0 behavior.
- People who prefer an open-source release with visible limitations, source, and license terms.
Intended result
Normalize exported Kindle, Kobo, and CSV highlights with provenance, filters, dedup review, and portable reports.
This summary is reconciled from the current catalog and repository documentation.
Features in v1.1.0
Capabilities below come from the current project README and release documentation.
[](https://github.com/loganpendragonmultiverse/book-annotation-extractor/actions/workflows/ci.yml)
Normalize exported highlights and notes from multiple ebook platforms into portable records. The command uses explicit UTF-8 JSON input and produces reviewable JSON or Markdown output.
Verified examples
Screenshots are shown only when the current README references a local source image. Otherwise, repository example files are linked directly.
Platforms and implementation
The public release claims only the cataloged platforms and technologies.
Supported platforms
- Windows
- macOS
- Linux
Built with
- Python
Quick start
The shortest documented path into the current release.
python -m pip install .
annotation-extract examples/sample.json
annotation-extract examples/sample.json --format json --output report.json
annotation-extract kobo.csv --adapter kobo-csv --format csv --output annotations.csvThe example documents the v1 input shape. Existing report files are never overwritten. Source inputs are read-only except where the documented purpose explicitly creates a new output artifact.
Version 1.1 reads the original JSON format plus exported Kindle text/HTML, Kobo CSV, and generic CSV files. Every normalized item includes source provenance. Repeated --work and --platform filters, date bounds, and --kind highlight|note narrow a report. Markdown groups annotations by work, CSV provides a portable table, and duplicate candidates remain visible for review.
Current limitations
These boundaries are part of the product and prevent the page from implying unverified capability.
The tool only normalizes files the user has already exported. It does not bypass DRM, log into platforms, or recover unavailable annotations.
Privacy and licensing
Review the actual data boundary before using a tool with sensitive inputs.
Privacy and safety
The tool runs locally and does not upload input or include telemetry. Python 3.10 or newer is supported on Windows, macOS, and Linux.
License and release
Book Annotation Extractor is published under MIT. The current cataloged release is v1.1.0, published 2026-07-27.
Related projects
Related projects are selected deterministically from shared catalog tags, category, and implementation technologies—not popularity or paid placement.
Book Annotation Extractor v1.1.0
Use the tagged release for downloads and release notes. Use the repository for source, issues, contribution guidance, security reporting, and complete documentation.
Page source: current Forge catalog plus README and CHANGELOG from the canonical local repository. Fingerprint: e328e9d2bfe1a186.