CHANGELOG
Automate figure extraction from PDFs with high accuracy and customizable output.
Changelog
All notable changes to this project will be documented in this file.
[0.1.0] - 2026-03-11
Added
- Initial release
- Automatic figure detection from PDF captions
- Sub-figure splitting with OCR label recognition
- High-quality rendering at configurable DPI (default 600)
- CLI interface for batch extraction
- Comprehensive error handling and logging
- Support for custom output formats (PNG, JPG)
- Robust fallback when OCR fails
License
- Licensed under AGPL-3.0-or-later due to PyMuPDF dependency