mirror of
https://github.com/snapotter-hq/SnapOtter.git
synced 2026-08-03 07:46:42 +02:00
- Pin PaddlePaddle to 3.0.0 on ARM64 to fix segfault in PIR inference engine (3.1+ crashes on aarch64 Debian Bookworm) - Fix text extraction for PaddleOCR 3.4.x result format (rec_texts) - Add Node.js-level fallback chain (best -> balanced -> fast) when Python subprocess crashes - Add multi-image OCR: processes all uploaded files sequentially with per-file progress and filename headers in combined output - Convert input images to PNG via Sharp before OCR so HEIC, AVIF, WebP, TIFF all work transparently - Implement real auto-detect language using Tesseract multi-lang script detection (analyzes Unicode ranges for Hangul, CJK, Kana, Latin) - Default enhance to off (hurts clean digital images)
10 lines
202 B
Plaintext
10 lines
202 B
Plaintext
rembg[cpu]==2.0.62
|
|
realesrgan==0.3.0
|
|
paddleocr[doc-parser]>=3.4.0,<3.5.0
|
|
paddlepaddle>=3.0.0,<3.1.0
|
|
mediapipe==0.10.21
|
|
onnxruntime==1.20.1
|
|
numpy==1.26.4
|
|
Pillow==11.1.0
|
|
opencv-python-headless==4.10.0.84
|