Results for “extraction”

60 skills
More results
tools-only
200 Aeon E7807df1
Guides feature extraction and preprocessing for time series data using aeon transformers, covering collection and series transformers with code examples.
7 · bundle
github
Mini Context Graph
Build a persistent, compounding knowledge base that combines a wiki, knowledge graph, and raw source storage for structured retrieval with provenance.
36.2k · bundle
composiohq
Firecrawl Automation
Automate web crawling and data extraction with Firecrawl: scrape pages, crawl sites, extract structured data, batch scrape URLs, and map website structures.
66.9k
scoheart
Firecrawl Scrape
Extracts clean, LLM-optimized markdown from any URL, including JavaScript-rendered SPAs, with support for concurrent scraping of multiple URLs and options like main-content-only extraction and custom output formats.
2
browser-act
Browser Act Skill Forge
Turns any website's data extraction or operation needs into reusable Agent-callable Skill packages by exploring API endpoints or DOM methods, then generating SKILL.md and Python scripts.
3.7k · bundle
mukul975
Performing Firmware Extraction With Binwalk
Extracts and analyzes firmware images using binwalk to identify embedded filesystems, compressed archives, bootloaders, kernel images, and cryptographic material. Covers entropy analysis, recursive extraction, filesystem mounting, and string analysis for credential and configuration discovery.
24.6k · bundle
mukul975
Analyzing Malware Behavior With Cuckoo Sandbox
Executes malware samples in Cuckoo Sandbox to observe runtime behavior including process creation, file system modifications, registry changes, network communications, and API calls. Generates comprehensive behavioral reports for malware classification and IOC extraction.
24.6k · bundle
majiayu000
Jq
Process JSON data from files or standard input using jq filters for extraction, filtering, and transformation.
567 · bundle
gabrielmoreira
Pdf2tex
Reconstructs editable LaTeX source from compiled PDFs by extracting text, math, tables, figures, and structure with pymupdf and AI.
17
composiohq
Apify Automation
Run Apify web scraping Actors, manage datasets, create reusable tasks, and retrieve crawl results directly from the terminal.
66.9k
lingxling
Pydicom
Read, write, and modify DICOM medical imaging files, including pixel data extraction, anonymization, format conversion, and compression handling.
253 · bundle
nvidia
Dicom Metadata Extract
Extracts selected metadata from a DICOM file and flags standard-tag PHI presence. Not for anonymization or clinical use.
2.2k · bundle
diegosouzapw
PDF
Process PDFs with Python and command-line tools: extract text and tables, merge, split, rotate, create, watermark, OCR, and handle passwords.
54 · bundle
qhjqhj00
Pydicom
Read, write, and manipulate DICOM medical imaging files, including pixel data extraction, metadata editing, anonymization, format conversion, and compression handling.
3 · bundle
auto-skiller
Data Scraping
Builds a configurable scraping agent that collects data from APIs, HTML, or RSS, enriches it with Gemini AI scoring, and stores results in Notion, Google Sheets, Supabase, or local files.
1 · bundle
browser-act
Producthunt Launches
Extract structured product launch data from Product Hunt leaderboards, enriched with maker profiles and website contact information.
3.7k · bundle
k-dense-ai
Pydicom
Read, write, and modify DICOM medical imaging files, including pixel data extraction, metadata manipulation, anonymization, and format conversion.
30.2k · bundle
k-dense-ai
Exa Search
Search the web and extract content from URLs using Exa, with support for academic and scientific sources.
30.2k · bundle
scoheart
Firecrawl Crawl
Bulk extract content from an entire website or site section by crawling pages that follow links, with configurable depth, path filters, and concurrency.
2
jorcan
PDF
Process PDFs with Python libraries and command-line tools: extract text and tables, create, merge, split, rotate, watermark, encrypt, and OCR documents.
0 · bundle
infinition
Ocr And Documents
Extracts text from PDFs and scanned documents, choosing the cheapest method that works, from direct file reads to full OCR.
2 · bundle
alirezarezvani
Browser Automation
Automate browser tasks, scrape websites, fill forms, capture screenshots, and extract structured data from web pages using Playwright.
20.4k · bundle
browser-act
Youtube Search API Skill
Extracts structured data from YouTube search results, including videos, shorts, channels, and playlists, using the BrowserAct API.
3.7k · bundle
k-dense-ai
Histolab
Process whole slide images for digital pathology: detect tissue, extract tiles, and prepare datasets for deep learning pipelines.
30.2k · bundle
muratcankoylan
Book Sft Pipeline
Convert books into supervised fine-tuning datasets and train style-transfer models that replicate an author's voice.
16.9k · bundle
scoheart
Firecrawl Agent
Extracts structured JSON data from complex multi-page websites using an AI agent that navigates pages and returns results matching a schema.
2
mariadb-corporation
Mariadb String Functions
Reference for MariaDB string functions: extraction, length, search, concatenation, modification, case, PCRE2 regex, formatting, and encoding, with common pitfalls and version notes.
0
tools-only
209 SQL 23f1987a
Provides SQL window function examples for ranking, aggregation, lag/lead, value extraction, frame specifications, and advanced analytics.
7 · bundle
browser-act
Youtube Batch Transcript Extractor API Skill
Extracts YouTube video transcripts and metadata in batch via the BrowserAct API, using search keywords and date filters.
3.7k · bundle
browser-act
Web Search Scraper API Skill
Extracts clean Markdown content from any website URL using the BrowserAct Web Search Scraper API, with automatic retry and error handling.
3.7k · bundle