Results for “extraction”
40 skillsIterationlayer
Integrate Iteration Layer APIs for document, website, image, and sheet processing. Composable APIs — Document Extraction, Website Extraction, Document to Markdown, Image Transformation, Image Generation, Document Generation, and Sheet Generation — share one credit pool and chain together into workflows.
2
Minimax PDF
**Trigger**: Use when creating, editing, or formatting PDF documents — generation, template application, content extraction, and validation.
1 · bundle
Book To Skill
Converts technical books and documents (PDF, EPUB, DOCX, HTML, Markdown, RTF, MOBI) into structured agent skills with frameworks, mental models, chapter references, and decision rules. Includes a full extraction pipeline for turning owned documents into reusable skills.
10 · bundle
Ocr
Processes documents through case.dev OCR for text and table extraction. Supports PDF and image files up to 500MB with page-level and word-level output. Use when the user mentions "OCR", "text extraction", "scan document", "digitize", "extract text from PDF", or needs word-level positional data from documents.
34
Defuddle
Extract clean markdown content from web pages by removing clutter, navigation, and ads to save tokens.
39.8k
Defuddle
Extract clean markdown content from web pages using Defuddle CLI, removing clutter and navigation to save tokens.
42.4k
More results
Defuddle
Extracts clean markdown content from web pages using the Defuddle CLI, removing navigation and clutter to reduce token usage.
20
PDF
Read, create, and review PDF files with visual rendering checks using Poppler, reportlab, pdfplumber, and pypdf.
23.3k · bundle
PPTX
Create, edit, and analyze PowerPoint presentations (.pptx files) with support for text extraction, raw XML access, and design customization.
66.9k · bundle
PPTX Generator
Generate, edit, and read PowerPoint presentations using PptxGenJS for creation, XML workflows for editing, and markitdown for text extraction.
12.9k · bundle
DOCX
Create, edit, and analyze .docx files with support for tracked changes, comments, formatting preservation, and text extraction.
66.9k · bundle
PPTX
Creates, reads, and edits PowerPoint (.pptx) files, including text extraction, visual inspection, and design guidance for polished slide decks.
61
Pdf2tex
Reconstructs editable LaTeX source from compiled PDFs by extracting text, math, tables, figures, and structure with pymupdf and AI.
17
PPTX Official
Creates, edits, and analyzes PowerPoint .pptx files, including text extraction, raw XML access, and design guidance for building presentations from scratch.
2 · bundle
PDF
Read, create, inspect, render, and verify PDF files where visual layout matters. Use Poppler rendering plus Python tools such as reportlab, pdfplumber, and pypdf for generation and extraction.
0 · bundle
PDF
Process PDFs with Python and command-line tools: extract text and tables, merge, split, rotate, create, watermark, OCR, and handle passwords.
54 · bundle
Defuddle
Extracts clean markdown from web pages via the Defuddle CLI, removing navigation and clutter to reduce token usage for reading or analyzing URLs.
3
PDF
Process PDFs with Python libraries and command-line tools: extract text and tables, create, merge, split, rotate, watermark, encrypt, and OCR documents.
0 · bundle
Ocr And Documents
Extracts text from PDFs and scanned documents, choosing the cheapest method that works, from direct file reads to full OCR.
2 · bundle
Extract Article Text
Extract clean article content — title, author, date, and body text — from PDFs, Word docs, and web pages.
2
Extract Aging Report Data
Extract structured fields from aging report documents.
2
Citation Management
Search Google Scholar and PubMed for papers, extract accurate metadata, validate citations, and generate properly formatted BibTeX entries.
30.2k · bundle
Baoyu Danger X To Markdown
Converts X (Twitter) tweets, threads, and articles to markdown with YAML front matter using a reverse-engineered API.
23.1k · bundle
PDF
Read, extract, merge, split, rotate, encrypt, and create PDF files. Supports markdown-to-PDF conversion, table extraction, form filling, and OCR for scanned documents.
1.5k · bundle
Defuddle
Extracts clean markdown content from web pages using the Defuddle CLI, removing clutter and navigation to save tokens. Prefer over WebFetch for reading or analyzing standard web pages.
2
PDF
Use when tasks involve reading, creating, or reviewing PDF files where rendering and layout matter; prefer visual checks by rendering pages (Poppler) and use Python tools such as `reportlab`, `pdfplumber`, and `pypdf` for generation and extraction.
0 · bundle
Obsidian Claw
Transforms chaotic notes, voice dumps, and meeting transcripts into a structured, interlinked knowledge base inside Obsidian by extracting entities, generating wikilinks, and surfacing forgotten context.
17 · bundle
PDF
Use when tasks involve reading, creating, or reviewing PDF files where rendering and layout matter; prefer visual checks by rendering pages (Poppler) and use Python tools such as `reportlab`, `pdfplumber`, and `pypdf` for generation and extraction.
0
Analyzing Malicious PDF With Peepdf
Perform static analysis of malicious PDF documents using peepdf, pdfid, and pdf-parser to extract embedded JavaScript, shellcode, and suspicious objects.
24.6k · bundle
Defuddle
Extracts clean markdown content from web pages using the Defuddle CLI, removing navigation and clutter to reduce token usage. Prefer over WebFetch for reading or analyzing standard web pages.
5
Tw Ghost
Extracts a language-agnostic ghost package (spec, tests, install and verify docs) from an existing repository, preserving behavior via tests.yaml and evidence bundles.
7 · bundle
Building Phishing Reporting Button Workflow
Deploy a phishing report button in email clients and build an automated triage workflow that analyzes user-reported suspicious emails, extracts IOCs, and provides feedback to reporters.
24.6k · bundle
DOCX
Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with ...
6 · bundle
DOCX
Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with tracked changes, (4) Adding comments, or any other document tasks
1 · bundle
DOCX
Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with tracked changes, (4) Adding comments, or any other document tasks
11 · bundle
DOCX
Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with tracked changes, (4) Adding comments, or any other document tasks
3 · bundle