hn • r/hackernews
Comment on: Zpdf: PDF text extraction in Zig
I've done lots of pdf indexing apps, and mupdf is by far the slowest and least able to open valid pdfs when it came to text extraction.
End users report that existing PDF indexing tools are sluggish and fail to extract text from many valid PDFs, hampering search and analysis. ZigExtract delivers a lightning‑fast, reliable PDF text parsing engine built with Zig, giving developers instant, accurate indexing for enterprise search and analytics.