pdf-ocr · zhaoheng588-tech/ai-skills-collection

Extract text from PDFs, scanned pages, and images locally

Extracts text from PDFs (including scanned pages via OCR), pulls text from images, and structures tables into CSV, all processed locally without paid APIs; useful for batch-converting documents into plain text or Markdown.

Good for

  • Extract text from a scanned PDF via OCR
  • Pull text out of an image file
  • Convert a table image into CSV
Category
Office

Open-source skills are maintained by their authors and listed as published, with attribution. Results depend on how well the skill fits your task and material.

A good place to start

Extract the text from this scanned PDF, it has no text layer, and save it as Markdown.

Make your next great thing.

Bring a question, a file, or an idea that’s not quite there yet.