CHANGELOG

Automate figure extraction from PDFs with high accuracy and customizable output.

Changelog

All notable changes to this project will be documented in this file.

[0.1.0] - 2026-03-11

Added

  • Initial release
  • Automatic figure detection from PDF captions
  • Sub-figure splitting with OCR label recognition
  • High-quality rendering at configurable DPI (default 600)
  • CLI interface for batch extraction
  • Comprehensive error handling and logging
  • Support for custom output formats (PNG, JPG)
  • Robust fallback when OCR fails

License

  • Licensed under AGPL-3.0-or-later due to PyMuPDF dependency