跳到主要内容
A lot of PDFs already have a text layer — running OCR again is just wasted work.
Firecrawl’s open-source pdf-inspector first checks the PDF type locally (in tens of milliseconds), then decides which files can be extracted directly and which pages actually need OCR.
Readable