Are you creating a PDF or extracting from one?
Choose the document and image converters when the final deliverable is a PDF; choose text, OCR, table, image, raster, or JSON tools when the PDF is the input.
Elysia Tools
Navigation
Workflow Playbook
Turn office files, wiki pages, comments, and images into PDFs, then extract text, images, tables, structure, or OCR layers from existing PDFs.
Hubs
PDF conversion work goes faster when the first decision is whether PDF is the output or the input. If source material lives in wiki pages, comments, office documents, spreadsheets, presentations, or image files, start with the matching document-to-PDF or image-to-PDF converter and verify page order, orientation, and visual fidelity before sending the file onward.
An existing PDF may contain text, scanned images, embedded tables, encrypted pages, or visual layouts that need to be reused. Authorized protected files should be unlocked first with encrypted-pdf-converter. Then choose pdf-to-image or pdf-rasterize-pages when the next tool needs page images, and choose text or structure extraction when the next step needs content rather than pixels.
Use pdf-text-extractor for selectable text, pdf-ocr-text-layer for scanned pages, pdf-to-clean-text-for-llm for retrieval or summarization pipelines, pdf-table-extractor-to-csv-json for analysis, and pdf-to-json-structure-explorer when layout or object structure must be inspected. A reliable finish includes checking representative pages against the source PDF so missing columns, OCR errors, or page-range mistakes are caught before delivery.
Workflow playbook
Start with authored material such as doc comments, wiki pages, presentations, spreadsheets, and text documents so reviewable or shareable PDF packages are created from the original source.
Convert image collections into PDFs when receipts, scans, screenshots, product photos, or archival images need a portable page-based format.
Unlock authorized encrypted files, export pages as images, or rasterize pages when downstream tools need visual page evidence instead of editable document objects.
Choose direct text extraction, OCR layering, clean text for LLMs, table extraction, or JSON structure inspection based on whether the downstream task needs readable text, tabular data, or layout-aware metadata.
Choose the document and image converters when the final deliverable is a PDF; choose text, OCR, table, image, raster, or JSON tools when the PDF is the input.
Use direct text extraction for digital PDFs, OCR text layering for scanned pages, and clean-text or structured extraction when the result must feed search, RAG, analysis, or automation.