스캔한 PDF는 글자를 선택·복사할 수 없습니다
복합기로 스캔한 문서에서 글자를 선택하려 해도 아무 일도 일어나지 않습니다. 검색해도 걸리지 않습니다. 끌면 단어가 아니라 사각형이 선택됩니다.
그 파일이 종이의 이미지일 뿐이기 때문입니다. 스캐너는 페이지가 어떻게 보이는지를 저장합니다. 사람 눈에는 글자지만 파일 안에서는 사진과 다를 바 없습니다. 문자 데이터가 아예 없으므로 선택할 것도 복사할 것도 검색할 것도 없습니다.
꺼내려면 이미지에서 글자를 읽어 내야 합니다. 흐린 인쇄, 기울어짐, 낮은 해상도는 읽을 수 없는 부분을 만듭니다. PDFIntact는 그런 곳을 비워 두지 않고 읽기 불가로 표시합니다. 조용히 채워 넣는 편이야말로 잘못을 눈치챌 수 없게 만들기 때문입니다.
지금 시도
가지고 계신 PDF로 확인해 보세요
진단은 브라우저 안에서만 이루어집니다. 파일은 어디로도 전송되지 않으며 가입도 필요 없습니다. 몇 페이지, 표와 그림이 몇 개나 나오는지 결제 전에 확인할 수 있습니다.
실례
まず見分けてください
同じ「PDF」でも、原因によって対処が変わります。次のどれに当てはまるかで判断できます。
- 文字を選択できない・検索しても出てこない
- スキャンPDF(画像)です。このページの対象です。
- 選択もコピーもできるが、貼り付けると読めない文字になる
- 画像ではなく、フォントの対応表が欠けている状態です。PDFをコピーすると文字化けする をご覧ください。
- 文字は取り出せるが、表の行と列がずれる
- PDFが表の構造を持っていないためです。PDFの表をExcelに貼り付けるとずれる をご覧ください。
자주 묻는 질문
- Why can I not copy text from a scanned PDF?
- Because the PDF contains a photograph of the page rather than character data. It looks like text to you, but to the file it is a picture, so there is nothing to select or search.
- How do I make it selectable?
- The characters have to be recognised from the image. PDFIntact analyses each page and exports the text, tables and charts to Excel, Word, Markdown or JSON. You can check in your browser, before paying, whether your file can be read.
- Can you read handwriting?
- Documents made up entirely of handwriting are out of scope. Use it for documents that are mainly printed text. The check tells you before purchase what cannot be converted.
- Can tables in scanned documents be extracted?
- Yes, with rules and merged cells reproduced. Faint, skewed or low-resolution printing will leave passages unread; those are reported as unreadable rather than guessed.
- What happens to the file I send?
- The check runs entirely in your browser and the original is never uploaded. A PDF sent for conversion is deleted within 24 hours, and by design nothing reaches our servers until payment has completed.