Training unicharset_extractor broken in 5.5.3
Current Behavior
The Training Windows program unicharset_extractor of Tesseract 5.5.3 release has failed silently, not producing unicharset output file as expected. I installed from the link below:
Expected Behavior
Should produce unicharset output file
Suggested Fix
No response
tesseract -v
tesseract v5.5.3.20260724 leptonica-1.87.0 libgif 5.2.2 : libjpeg 8d (libjpeg-turbo 3.1.3) : libpng 1.6.58 : libtiff 4.7.2 : zlib 1.3.1 : libwebp 1.6.0 : libopenjp2 2.5.4 Found AVX2 Found AVX Found FMA Found SSE4.1 Found libarchive 3.8.8 zlib/1.3.2 liblzma/5.8.3 bz2lib/1.0.8 liblz4/1.10.0 libzstd/1.5.7 expat/expat_2.8.2 cng/1.0 libb2/system Found libcurl/8.21.0 Schannel zlib/1.3.2 brotli/1.2.0 zstd/1.5.7 libidn2/2.3.8 libpsl/0.21.5 libssh2/1.11.1 WinLDAP
Operating System
Windows 11
Other Operating System
No response
uname -a
No response
Compiler
No response
CPU
Intel i9 14900HX
Virtualization / Containers
No response
Other Information
No response
Source: tesseract-ocr/tesseract