Results for “extraction”
293 skillsDocument AI
Comprehensive patterns for AI-powered document understanding including PDF parsing, OCR, invoice/receipt extraction, table extraction, multimodal RAG with vision models, and structured data output. Use when "document parsing, PDF extraction, OCR, invoice processing, receipt extraction, document understanding, LlamaParse, Unstructured, vision document, table extraction, structured output from PDF, " mentioned.
128 · bundle
Document Extraction API
Extract structured data from documents using AI-powered field extraction.
2
Sec Edgar Pipeline
SEC EDGAR extraction pipeline: setup, filing discovery by CIK, recipe-driven extraction, and report generation.
71 · bundle
Agent Web Scraper
Web Scraper IA — Expert en extraction web (Scrapy, BeautifulSoup, Playwright, anti-bot, proxy rotation, data extraction)
6
Agent Browser
Automates browser interactions for web testing, form filling, screenshots, and data extraction.
7 · bundle
Web Scraper
Web scraping and content comprehension agent — multi-strategy extraction with cascade fallback, news detection, boilerplate removal, structured metadata, and LLM entity extraction
228 · bundle
More results
Taggun Automation
Automate Taggun document data extraction operations through Composio's Taggun toolkit via Rube MCP.
66.9k
Diffbot Automation
Automate Diffbot data extraction and analysis through Composio's Diffbot toolkit via Rube MCP.
66.9k
Nlp Advanced
Use when extracting structured information from text - named entity recognition, relation extraction, coreference resolution, knowledge graph construction, and information extraction pipelinesUse when ", " mentioned.
128 · bundle
Dynamic Workflow Mode
Design task-local harnesses, eval gates, and reusable skill extraction for adaptive agent workflows.
226k
Crawl4ai MCP Server
Self-hosted web crawling and content extraction exposed as MCP tools, with depth control and clean markdown output.
28
Tavily Best Practices
Reference for building Tavily-powered search, extraction, crawling, and research into agentic workflows and RAG systems.
2 · bundle
Fabric Extract V2
Parallel Fabric Workspace Metadata Extraction
0
Scrape Do Automation
Automate web scraping and data extraction tasks using the Scrape Do toolkit via Rube MCP and Composio.
66.9k
Mediabunny
Extract audio and video metadata such as duration and dimensions using the Mediabunny library in the browser.
3.9k · bundle
Hwp
Use kordoc for agent-native HWP/HWPX document parsing, JSON extraction, diffing, form-field extraction, and Markdown→HWPX reverse conversion (read/convert only — for binary editing use rhwp-edit).
3 · bundle
Parsehub Automation
Automates Parsehub data extraction tasks through Composio's Parsehub toolkit via Rube MCP, with tool discovery and connection management.
66.9k
Extract Receipt Data
Extract merchant, date, line items, tax, and total from receipts.
2
Reduce
Extract structured knowledge from source material. Comprehensive extraction is the default — every insight that serves the domain gets extracted. For domain-relevant sources, skip rate must be below 10%. Zero extraction from a domain-relevant source is a BUG. Triggers on "/reduce", "/reduce [file]", "extract insights", "mine this", "process this".
3 · bundle
Extract Nomina Data
Extract employee, employer, earnings, deductions, tax, and net pay from Spanish or Latin American payroll slips.
2
Extract Packing List Data
Extract shipper, consignee, order references, carton lines, weights, and package totals from packing lists.
2
Stitch Code To Design
Converts existing frontend code into a Stitch Design by chaining static HTML extraction, design system extraction, and file upload.
0
Hwp
Use kordoc for agent-native HWP/HWPX document parsing, JSON extraction, diffing, form-field extraction, and Markdown→HWPX reverse conversion (read/convert only — for binary editing use rhwp-edit).
0 · bundle
Extract Content
Fetch a published article from a URL and extract its title, headings, body content, and metadata into a markdown file the update pipeline can audit.
0
Extract
Crawl a live website, extract its design system, brand surface, and page inventory, and save the snapshot under stardust/current/ for downstream redesign or migration.
142 · bundle
Extract Modello F24 Data
Extract structured fields from modello f24 documents.
2
Extract Article Text
Extract clean article content — title, author, date, and body text — from PDFs, Word docs, and web pages.
2
Extract Rechnung Data
Extract structured fields from rechnung documents.
2
Website Extraction API
Extract typed JSON from public website pages using a schema.
2
Extract
Extract and consolidate reusable components, design tokens, and patterns into your design system. Identifies opportunities for systematic reuse and enriches your component library.
1 · bundle
Extract
Extract and consolidate reusable components, design tokens, and patterns into your design system. Identifies opportunities for systematic reuse and enriches your component library.
55 · bundle
Extract P60 Data
Extract structured fields from p60 documents.
2
Extract 1099 Nec Data
Extract structured fields from 1099-nec documents.
2
Extract Traffic Fine Data
Extract violation details, fine amounts, vehicle information, and payment deadlines from traffic fine notices.
2
Extract I 9 Data
Extract structured fields from i-9 documents.
2
Extract Recu Data
Extract receipt number, issuer, payer, date, VAT, and total from French receipts.
2