Extract text from a scanned PDF
Import a PDF without a text layer, run OCR, then retrieve the recognized text.
What does this template do?
Import a PDF without a text layer, run OCR, then retrieve the recognized text.
Pipeline steps
This template automatically chains 3 steps in the Workspace. You can review every setting before running it.
- 1Import the file
- 2Recognize text (OCR)
- 3Process
When should you use this template?
- Prepare or transform a PDF through several steps.
- Avoid downloading and re-importing the file between operations.
- Reuse the same processing flow with other documents.
Local processing in your browser
When the selected modules support local processing, your files stay on your device while the Pipeline runs.