Results for “extraction”

40 skills
More results
galyarderlabs
Defuddle
Extracts clean markdown content from web pages using the Defuddle CLI, removing navigation and clutter to reduce token usage.
20
openai
PDF
Read, create, and review PDF files with visual rendering checks using Poppler, reportlab, pdfplumber, and pypdf.
23.3k · bundle
composiohq
PPTX
Create, edit, and analyze PowerPoint presentations (.pptx files) with support for text extraction, raw XML access, and design customization.
66.9k · bundle
minimax-ai
PPTX Generator
Generate, edit, and read PowerPoint presentations using PptxGenJS for creation, XML workflows for editing, and markitdown for text extraction.
12.9k · bundle
composiohq
DOCX
Create, edit, and analyze .docx files with support for tracked changes, comments, formatting preservation, and text extraction.
66.9k · bundle
comeonoliver
PPTX
Creates, reads, and edits PowerPoint (.pptx) files, including text extraction, visual inspection, and design guidance for polished slide decks.
61
gabrielmoreira
Pdf2tex
Reconstructs editable LaTeX source from compiled PDFs by extracting text, math, tables, figures, and structure with pymupdf and AI.
17
nimoqup046-collab
PPTX Official
Creates, edits, and analyzes PowerPoint .pptx files, including text extraction, raw XML access, and design guidance for building presentations from scratch.
2 · bundle
concertonotes
PDF
Read, create, inspect, render, and verify PDF files where visual layout matters. Use Poppler rendering plus Python tools such as reportlab, pdfplumber, and pypdf for generation and extraction.
0 · bundle
diegosouzapw
PDF
Process PDFs with Python and command-line tools: extract text and tables, merge, split, rotate, create, watermark, OCR, and handle passwords.
54 · bundle
phoroth
Defuddle
Extracts clean markdown from web pages via the Defuddle CLI, removing navigation and clutter to reduce token usage for reading or analyzing URLs.
3
jorcan
PDF
Process PDFs with Python libraries and command-line tools: extract text and tables, create, merge, split, rotate, watermark, encrypt, and OCR documents.
0 · bundle
infinition
Ocr And Documents
Extracts text from PDFs and scanned documents, choosing the cheapest method that works, from direct file reads to full OCR.
2 · bundle
iterationlayer
Extract Article Text
Extract clean article content — title, author, date, and body text — from PDFs, Word docs, and web pages.
2
iterationlayer
Extract Aging Report Data
Extract structured fields from aging report documents.
2
k-dense-ai
Citation Management
Search Google Scholar and PubMed for papers, extract accurate metadata, validate citations, and generate properly formatted BibTeX entries.
30.2k · bundle
jimliu
Baoyu Danger X To Markdown
Converts X (Twitter) tweets, threads, and articles to markdown with YAML front matter using a reverse-engineered API.
23.1k · bundle
memento-teams
PDF
Read, extract, merge, split, rotate, encrypt, and create PDF files. Supports markdown-to-PDF conversion, table extraction, form filling, and OCR for scanned documents.
1.5k · bundle
nimoqup046-collab
Defuddle
Extracts clean markdown content from web pages using the Defuddle CLI, removing clutter and navigation to save tokens. Prefer over WebFetch for reading or analyzing standard web pages.
2
jarbitechture
PDF
Use when tasks involve reading, creating, or reviewing PDF files where rendering and layout matter; prefer visual checks by rendering pages (Poppler) and use Python tools such as `reportlab`, `pdfplumber`, and `pypdf` for generation and extraction.
0 · bundle
gabrielmoreira
Obsidian Claw
Transforms chaotic notes, voice dumps, and meeting transcripts into a structured, interlinked knowledge base inside Obsidian by extracting entities, generating wikilinks, and surfacing forgotten context.
17 · bundle
eliferjunior
PDF
Use when tasks involve reading, creating, or reviewing PDF files where rendering and layout matter; prefer visual checks by rendering pages (Poppler) and use Python tools such as `reportlab`, `pdfplumber`, and `pypdf` for generation and extraction.
0
mukul975
Analyzing Malicious PDF With Peepdf
Perform static analysis of malicious PDF documents using peepdf, pdfid, and pdf-parser to extract embedded JavaScript, shellcode, and suspicious objects.
24.6k · bundle
lucaspmarie-a11y
Defuddle
Extracts clean markdown content from web pages using the Defuddle CLI, removing navigation and clutter to reduce token usage. Prefer over WebFetch for reading or analyzing standard web pages.
5
sdiamante13
Tw Ghost
Extracts a language-agnostic ghost package (spec, tests, install and verify docs) from an existing repository, preserving behavior via tests.yaml and evidence bundles.
7 · bundle
mukul975
Building Phishing Reporting Button Workflow
Deploy a phishing report button in email clients and build an automated triage workflow that analyzes user-reported suspicious emails, extracts IOCs, and provides feedback to reporters.
24.6k · bundle
jantoniofc
DOCX
Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with ...
6 · bundle
timlai666
DOCX
Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with tracked changes, (4) Adding comments, or any other document tasks
1 · bundle
sinhoneyy
DOCX
Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with tracked changes, (4) Adding comments, or any other document tasks
11 · bundle
levalencia
DOCX
Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with tracked changes, (4) Adding comments, or any other document tasks
3 · bundle