FatafatPDF mascot FatafatPDF
No sign-up No file limits No ads
Runs entirely in your browser — your files never leave your device

將 PDF 轉換為文字

將每一個字擷取到一份純文字 .txt 檔案中。

運作方式

上傳您的 PDF 並下載結果——每一頁的文字都會依閱讀順序擷取,輸出成一份純文字 .txt 檔案。

這會保留格式、表格或版面嗎?

不會——它會依閱讀順序擷取原始文字,輸出成單一 .txt 檔案。版面、字型與格式不會被保留。

這適用於掃描的 PDF 嗎?

只有在 PDF 本身已有可選取文字時才可以。單純是圖片的掃描頁(未經 OCR)轉換出的文字會是空白或幾乎空白的內容。

Can I get one .txt file per page instead of one combined file?

Currently it's one combined text file for the whole document, in page order.

Does this work on PDFs in languages other than English?

Yes — it extracts whatever selectable text the PDF contains, in any language, as long as the text is real (not a scanned image without OCR).

Why is some of my extracted text jumbled or out of order?

PDFs don't always store text in strict visual reading order internally, especially multi-column layouts — extraction follows the file's internal text order, which can occasionally differ from the visual layout.

How can I tell where one page ends and the next begins in the output?

Each page is separated by a line reading "--- page break ---" in the downloaded .txt file, so you can always tell where one page's text stops and the next one starts.

Why doesn't the output keep line breaks or paragraphs?

page.pdfToText.faq7Text

What's the difference between this and PDF to Word?

page.pdfToText.faq8Text

My PDF is scanned — how do I still get text out of it?

page.pdfToText.faq9Text