Results for “full-text-extraction”

51 skills
More results
iterationlayer
Extract Article Text
Extract clean article content — title, author, date, and body text — from PDFs, Word docs, and web pages.
2
iterationlayer
Extract Modello F24 Data
Extract structured fields from modello f24 documents.
2
iterationlayer
Extract Receipt Data
Extract merchant, date, line items, tax, and total from receipts.
2
iterationlayer
Extract Modello 730 Data
Extract structured fields from modello 730 documents.
2
flyfiref
Ocr And Documents
Extract text from PDFs/scans (pymupdf, marker-pdf).
0 · bundle
iterationlayer
Extract Nda Terms
Extract parties, obligations, restrictions, permitted disclosures, and expiry dates from non-disclosure agreements.
2
micsapp
Reduce
Extract structured knowledge from source material. Comprehensive extraction is the default — every insight that serves the domain gets extracted. For domain-relevant sources, skip rate must be below 10%. Zero extraction from a domain-relevant source is a BUG. Triggers on "/reduce", "/reduce [file]", "extract insights", "mine this", "process this".
3 · bundle
upayanghosh
Synapse Web Scrape
Fetches and extracts readable text from a URL, then summarizes or presents key information with source citation.
14 · bundle
browser-act
Wechat Article Search API Skill
Extract full article contents from WeChat using the BrowserAct API, with keyword search and date filtering.
3.7k · bundle
browser-act
Tiktok Video Detail
Extracts complete metadata from a TikTok video page by reading SSR-embedded data, including author profile, engagement stats, music info, hashtags, and slideshow images.
3.7k · bundle
iterationlayer
Extract Nomina Data
Extract employee, employer, earnings, deductions, tax, and net pay from Spanish or Latin American payroll slips.
2
scoheart
Tavily Extract
Extracts clean markdown or text from one or more URLs using the Tavily CLI, with support for JavaScript-rendered pages and query-focused chunking.
2
scoheart
Firecrawl Crawl
Bulk extract content from an entire website or site section by crawling pages that follow links, with configurable depth, path filters, and concurrency.
2
browser-act
Zhihu Search API Skill
Extracts structured article details and full content from Zhihu search results via the BrowserAct API, with keyword and date filtering.
3.7k · bundle
browser-act
Ecommerce Product Detail
Extract complete product information from any e-commerce product page using a universal multi-layer extraction strategy.
3.7k · bundle
gonglingrui
Text Truncator
智能截断文本,保持内容的完整性和语义连贯性。适用于长文本预处理、确保文本不超过指定长度限制
349 · bundle
leandrobenjaminl
Youtube Transcript
Extrae transcripciones de videos de YouTube y genera resúmenes estructurados con timestamps y análisis temático.
0
bitwikiorg
Inventory
Imported skill inventory from anthropic
3
baofeng-tech
Tavily Extract
Extract clean, readable content from one or more URLs using Tavily Extract via AIsa API. Useful for reading full articles without visiting the page. Use when: the user needs web search, research, source discovery, or content extraction.
1 · bundle
ichichuang
Ocr And Documents
Extract text from PDFs and scanned documents. Use web_extract for remote URLs, pymupdf for local text-based PDFs, marker-pdf for OCR/scanned docs. For DOCX use python-docx, for PPTX see the powerpoint skill.
0 · bundle
lionelndong
Extract Content
Fetch a published article from a URL and extract its title, headings, body content, and metadata into a markdown file the update pipeline can audit.
0
drnabeelkhan
Quote Engine 5 Ready To Post Quotes From Any Long Text
Extracts five strong, ready-to-post quotes from any long text, interview, or transcript, scoring candidates for standalone impact and tagging each with the best platform and psychological lever.
2
iterationlayer
Extract Invoice Data
Extract vendor name, line items, totals, and dates from invoice documents.
2
iterationlayer
Extract Resume Data
Extract candidate name, contact details, work history, and skills from resumes.
2
iterationlayer
Extract Balance Sheet Data
Extract structured fields from balance sheet documents.
2
iterationlayer
Extract Busta Paga Data
Extract employee, employer, gross pay, INPS/IRPEF deductions, and net pay from Italian payslips.
2
google-labs-code
Stitch Extract Static HTML
Extract self-contained static HTML from a built web application or React components by inlining CSS and images.
6.4k · bundle
iterationlayer
Extract Ct600 Data
Extract structured fields from ct600 documents.
2
mesteriis
Design System Extractor
Extracts tokens, typography, spacing, component states, assets, layout, responsive rules, and interactions from a UI reference.
0 · bundle
browser-act
Business Contact Social Links Skill
Extract official websites and social media profiles (LinkedIn, Facebook, X/Twitter, Instagram, YouTube, TikTok) from a company name or website URL using automated browser scripts.
3.7k · bundle
iterationlayer
Extract Traffic Fine Data
Extract violation details, fine amounts, vehicle information, and payment deadlines from traffic fine notices.
2
qhjqhj00
Adhx
Fetches any X/Twitter post as structured JSON via the ADHX API, including full article content, author info, and engagement metrics, without scraping or a browser.
3 · bundle
mmehdi0606
Adhx
Fetch any X/Twitter post as clean LLM-friendly JSON. Converts x.com, twitter.com, or adhx.com links into structured data with full article content, author info, and engagement metrics. No scraping or browser required.
2
welitonevoc
Adhx
Fetch any X/Twitter post as clean LLM-friendly JSON. Converts x.com, twitter.com, or adhx.com links into structured data with full article content, author info, and engagement metrics. No scraping or browser required.
1
ranbot-ai
Adhx
Fetch any X/Twitter post as clean LLM-friendly JSON. Converts x.com, twitter.com, or adhx.com links into structured data with full article content, author info, and engagement metrics. No scraping or
6