将扫描文档转换为可搜索 PDF。使用关键词、数字、姓名等搜索扫描件、文档和 PDF 的内容。
或拖放、粘贴,或从云端选择
What the app does with your file, before you upload anything.
将文档数字化并实现无纸化办公,不仅有利于环保,还能节省你在一层层文件夹中查找所花费的大量时间。然而,许多扫描件只是文档的图像,这类文件除了文件名外,其内容都无法搜索。
通过让扫描件可搜索,会在扫描图像上添加一个额外的文本层,从而创建所谓的混合 PDF。这个文本层允许你在 PDF 中按关键词、金额和其他数字、姓名或主题进行搜索。
A searchable PDF therefore looks exactly like the scan it came from. The pages, the stamps and the handwriting are all still in place. The difference sits underneath, where a layer of real characters waits for your reader, your computer and your document manager to read it.
File extension: .pdf, MIME type: application/pdf
PDF on WikipediaThis app takes a PDF and gives a PDF back. Nothing about the look of the document changes. The pages keep their size, the scans keep their marks, and a hidden layer of recognized text is written in behind them.
That is what separates it from the apps that pull text out of a file. Those hand you the words in a plain text file and leave the layout behind. This one leaves the layout alone, so you keep a document you can pass on, print or sign, and it searches like a text file as well.
The upload has to be a PDF. If you are holding a photo or an image file instead, the JPG to PDF and PNG to PDF apps build the PDF for you and run the same text recognition while they do it.
Text recognition reads a page against a model of the language it expects, so naming that language is the single most useful thing you can do for the result. The setting sits beside the uploader and it is the only choice the app asks you to make.
It takes more than one language at a time. A contract with an English cover page and a German annex reads best with both selected, and so does a scan that mixes a Latin script with Arabic, Cyrillic or Chinese.
The list holds 124 source languages, among them Arabic, Hindi, Thai, Vietnamese, Chinese in simplified and traditional form, and the Cyrillic forms of Azerbaijani, Serbian, and Uzbek. English is selected for you if you change nothing.
Text recognition does not read every page equally well. Here is what to expect from the file you are holding.
| The PDF you have | Can it be made searchable | What to do |
|---|---|---|
| A flat scan of printed pages | Yes | Upload it and pick the document language. This is the file the app is built for. |
| A page photographed with a phone | Usually | Shadows, angles and a curved spine all cost accuracy. Shoot flat and in even light where you can. |
| A faint copy, a fax or handwriting | Partly | Print and handwriting are not the same problem. Expect gaps and read the result before you rely on it. |
| A PDF whose text you can already select | Nothing to add | It is searchable already. Use PDF to Text if what you want is the words in a separate file. |
| A password-protected or damaged PDF | No | The app cannot open it. Remove the protection or repair the file first, then upload it again. |
| A JPG, PNG or other image file | Not here | JPG to PDF and PNG to PDF build the PDF and run the same recognition. |
Recognition quality follows scan quality. A crooked page, a low resolution or a shadow across the paper all cost accuracy, so it is worth reading the result before you file the document away.
A scan and a searchable PDF look the same on screen. These are the three things only one of them can do.
The find command in any PDF reader jumps straight to the word you typed. On a 90-page lease that is the difference between a search and an afternoon.
Desktop search and most cloud drives read the text layer inside a PDF. A searchable PDF turns up in results by what is written in it, not only by its file name.
Dragging your cursor across a scan does nothing, because there is nothing there to select. A searchable PDF lets you pick up a paragraph and paste it into a mail or a note.
The papers you scan and the documents you convert can be personal. Here is exactly how yours are handled.
Every upload and download runs over an encrypted connection, so your files cannot be read in transit.
Text recognition is fully automated. Nobody reads your scans or documents, and nothing is shared with third parties.
Only what a conversion needs is processed. Your files are never used to train AI models.
Upload the PDF you want to make searchable. Scanned pages and photographed pages both work.
从列表中选择文档语言(可选)。
点击 "Start"。稍等片刻,你的文档就会准备好。
A searchable PDF is a scan with a layer of real text behind it. The pages still look exactly like the paper they came from, but the words underneath can be searched, selected and copied. It is also called a hybrid PDF, because it holds the picture and the text in one file.
Upload the PDF above, pick the language of the document and press Start. Text recognition reads the pages, writes what it finds into a hidden layer and hands the PDF back. Several PDFs can go in at once, and each one comes back on its own.
Yes. Making a PDF searchable is free, needs no sign-up and costs no credits. A plan only matters for very large documents and long batches.
No. If you can drag your cursor across a line and the words highlight, the PDF already carries a text layer and is already searchable. Running it through again adds nothing. If what you want is the words in a separate file, use the PDF to Text app instead.
It reads 124 source languages, among them Arabic, Hindi, Thai, Vietnamese, Chinese in simplified and traditional form, and the Cyrillic forms of Azerbaijani, Serbian, and Uzbek. Select every language that appears in the document, because a mixed page reads best when all of them are listed.
Recognition quality follows scan quality. A crooked page, a low resolution, a shadow across the paper or a photo taken by hand all cost accuracy, and handwriting is harder again than print. Scan flat at 300 dpi where you can, and set the document language before you start.
Not with this app. It adds a text layer rather than removing one. Making a PDF non-searchable means flattening every page back into an image, which is the opposite job, and nothing here does that.
Uploads and downloads run over an encrypted connection, processing happens in certified data centres inside the European Union, and the whole pipeline is automated, so nobody reads your documents. Your files are never used to train AI models and you keep the rights to them.
Every one of these runs in the browser. No install, no account, and the same text recognition behind it.