Results for “web-scraping”
81 skillsscrapingbee-automation
Automate web scraping tasks using Scrapingbee through Composio's Rube MCP toolkit, with tool discovery and connection management.
66.9k
scrapingant-automation
Automate web scraping tasks using Scrapingant via Composio's Rube MCP toolkit, with dynamic tool discovery and connection management.
66.9k
scrapfly-automation
Automate Scrapfly web scraping operations through Composio's Scrapfly toolkit via Rube MCP.
66.9k
webscraping-ai-automation
Automates web scraping AI operations through Composio's Webscraping AI toolkit via Rube MCP, with tool discovery and connection management.
66.9k
hasdata
Extract public web data, search engine results, and structured data from platforms like Google, Amazon, and Zillow using HasData APIs.
42.4k · bundle
apify-actor-runner
Runs Apify cloud actors for structured web scraping and exports datasets to S3, with input schema validation and webhook notifications.
28
More results
firecrawl-automation
Automate web crawling and data extraction with Firecrawl: scrape pages, crawl sites, extract structured data, batch scrape URLs, and map website structures.
66.9k
brightdata-automation
Automate Brightdata web scraping and data collection operations through Composio's Brightdata toolkit via Rube MCP.
66.9k
apify-automation
Run Apify web scraping Actors, manage datasets, create reusable tasks, and retrieve crawl results directly from the terminal.
66.9k
scrape-do-automation
Automate web scraping and data extraction tasks using the Scrape Do toolkit via Rube MCP and Composio.
66.9k
hasdata-cli
Provides command-line access to search, scraping, and structured web data from over 40 APIs including Google, Amazon, Yelp, and Zillow.
42.4k · bundle
hasdata
Extract public web data via HasData APIs, including search engine results, structured data from ecommerce, travel, jobs, and local business platforms, with support for web scraping, pre-parsed APIs, and async jobs.
3 · bundle
web-scraper
Web scraping inteligente multi-estrategia. Extrai dados estruturados de paginas web (tabelas, listas, precos). Paginacao, monitoramento e export CSV/JSON.
1
scrapers
Monitors websites, RSS feeds, and e-commerce pages for content changes and price drops, with social listening and custom scraping for competitive intelligence and market research.
10
data-scraping
Builds a configurable scraping agent that collects data from APIs, HTML, or RSS, enriches it with Gemini AI scoring, and stores results in Notion, Google Sheets, Supabase, or local files.
1 · bundle
scrape-webpage
Extract content, metadata, and images from a webpage for import or migration to AEM Edge Delivery Services.
142 · bundle
firecrawl-scrape
Extracts clean, LLM-optimized markdown from any URL, including JavaScript-rendered SPAs, with support for concurrent scraping of multiple URLs and options like main-content-only extraction and custom output formats.
2
web-search-scraper-api-skill
Extracts clean Markdown content from any website URL using the BrowserAct Web Search Scraper API, with automatic retry and error handling.
3.7k · bundle
firecrawl-crawl
Bulk extract content from an entire website or site section by crawling pages that follow links, with configurable depth, path filters, and concurrency.
2
zhihu-fetcher
抓取知乎收藏夹与文章正文为 Markdown,支持图片本地化、断点续传及写入 Obsidian 知识库。
28 · bundle
webcrawler-deep-crawl
Deep-crawl any website from start URLs, returning per-page LLM-ready text, markdown, or HTML with metadata and in-scope outbound links.
3.7k · bundle
tiktok-hashtag-videos
Extract paginated video lists for a TikTok hashtag, including author profiles, engagement stats, music, and video metadata.
3.7k · bundle
browser-automation
Automate browser tasks, scrape websites, fill forms, capture screenshots, and extract structured data from web pages using Playwright.
20.4k · bundle
firecrawl-map
Discovers and lists all URLs on a website, with optional search filtering to find specific pages within large sites.
2
python-executor
Execute Python code in a safe sandboxed environment with 100+ pre-installed libraries for data processing, web scraping, image manipulation, video creation, 3D model processing, PDF generation, API calls, and automation.
584
google-news-api-skill
Extracts structured news data from Google News via the BrowserAct API, including headlines, sources, publication times, and article links.
3.7k · bundle
firecrawl
Search the web, scrape pages, crawl sites, and interact with dynamic content via the Firecrawl CLI, returning clean markdown for LLM contexts.
2 · bundle
facebook-page-profile-posts
Extracts posts from public Facebook pages or profiles, returning structured data with text, engagement metrics, reaction breakdowns, media thumbnails, and pagination support.
3.7k · bundle
producthunt-launches
Extract structured product launch data from Product Hunt leaderboards, enriched with maker profiles and website contact information.
3.7k · bundle
firecrawl-agent
Extracts structured JSON data from complex multi-page websites using an AI agent that navigates pages and returns results matching a schema.
2
page-reduce
Reduces a webpage to a structural skeleton by tokenizing content in the browser and applying LLM reasoning to collapse repeated patterns.
142 · bundle
diffbot-automation
Automate Diffbot data extraction and analysis through Composio's Diffbot toolkit via Rube MCP.
66.9k
business-contact-social-links-skill
Extract official websites and social media profiles (LinkedIn, Facebook, X/Twitter, Instagram, YouTube, TikTok) from a company name or website URL using automated browser scripts.
3.7k · bundle
xhs-explore
Search, browse, and analyze content on Xiaohongshu (Little Red Book) including feeds, note details, comments, and user profiles.
1.8k
parsehub-automation
Automates Parsehub data extraction tasks through Composio's Parsehub toolkit via Rube MCP, with tool discovery and connection management.
66.9k
zyte-api-automation
Automate Zyte API operations through Composio's Rube MCP integration, including tool discovery, connection management, and execution.
66.9k