CarePDF Logo
CarePDF Logo

OCR PDF

Extract the text layer from a PDF so you can search, copy, or reuse the content.

Select PDF files

or drag and drop files here

Got a password-protected PDF? Unlock it here first

How to OCR PDF – テキストの抽出

1

Upload your scanned document

Select a PDF that consists of scanned pages or images of text.

2

Select output format

Choose whether you want a fully searchable PDF or a plain, editable text file.

3

Extract text

The Optical Character Recognition (OCR) engine reads the image and converts it back into machine-readable text.

Frequently Asked Questions

Does this work on scanned documents that are just photos of text?
This tool reads text layers that are already embedded in a PDF using the pdf.js engine, so it works best on PDFs that already contain real text. A pure image-only scan with no text layer needs full optical character recognition, which requires server-side processing (like Tesseract OCR) rather than the in-browser extraction this tool performs.
How do I know if my PDF already has a text layer?
通常の PDF ビューアでページ上のテキストを選択または強調表示してみてください。個々の単語を選択できる場合は、テキスト レイヤーが存在し、このツールはそれをきれいに抽出します。
What output formats can I choose?
元の外観を維持しながらコンテンツを選択および検索可能にしたい場合は、検索可能な PDF を選択します。また、生の単語を他の場所に貼り付ける必要があるだけの場合は、プレーン テキスト ファイルを選択します。
Does this handle handwritten notes?
いいえ、このツールは画像ベースの文字認識を実行するのではなく、既存のデジタル テキスト レイヤーを抽出するため、手書き文字は読み取ることができません。
Is my document processed on your servers?
いいえ、テキスト抽出は pdf.js を使用してブラウザ内で完全に行われるため、ドキュメントのコンテンツがどこにもアップロードされることはありません。

Explore More PDF Tools