extract-slide-text
Extract text from each page of a PDF into a markdown file using pdftotext. Produces slide_ascii.md with a heading, image reference, and extracted text per slide. USE FOR: extract text from PDF slides, get slide text content, PDF to markdown text, slide_ascii.md.
How do I install this agent skill?
npx skills add https://github.com/pamelafox/presentation-skills --skill extract-slide-textIs this agent skill safe to install?
- Gen Agent Trust Hubpass
The skill extracts text from PDFs using a system utility. It is safe for its intended use, though it handles untrusted PDF content without sanitization, posing a minor indirect prompt injection risk.
- Socketpass
No alerts
- Snykpass
Risk: LOW · No issues
What does this agent skill do?
Extract slide text from PDF
Run the extract_slide_text.py script to extract the text content of each PDF page into a structured markdown file:
uv run .agents/skills/extract-slide-text/extract_slide_text.py <pdf_path> <output_path> [images_dir]
Arguments
pdf_path(required): Path to the PDF file.output_path(required): Path to write the output markdown fileimages_dir(optional): Path to the slide images directory. Used to generate correct relative image references. Defaults toslide_images/.
Output format
A markdown file with one section per slide:
## Slide 1

\```
Extracted text content from slide 1
\```
## Slide 2

\```
Extracted text content from slide 2
\```
Pages with no extractable text (e.g., full-bleed images) show (no extractable text).
Why this matters
PDF text extraction is deterministic — it produces ground-truth slide content without relying on vision models. This prevents misidentification of embedded screenshots or demo captures as actual slide content, a common failure mode when using only image-based slide analysis.
Prerequisites
Poppler utilities must be installed (provides the pdftotext command):
- macOS:
brew install poppler - Ubuntu:
apt-get install poppler-utils
How can the creator link this skill?
Add the canonical catalog link to the repository README so users can inspect current installs and available audits. The publishing guide covers the complete discovery path.
<a href="https://skillzs.dev/skills/pamelafox/presentation-skills/extract-slide-text">View extract-slide-text on skillZs</a>