Results for “web-scraping”
40 skillsScrapingbee Automation
Automate web scraping tasks using Scrapingbee through Composio's Rube MCP toolkit, with tool discovery and connection management.
66.9k
Scrapingant Automation
Automate web scraping tasks using Scrapingant via Composio's Rube MCP toolkit, with dynamic tool discovery and connection management.
66.9k
Scrapfly Automation
Automate Scrapfly web scraping operations through Composio's Scrapfly toolkit via Rube MCP.
66.9k
Webscraping AI Automation
Automates web scraping AI operations through Composio's Webscraping AI toolkit via Rube MCP, with tool discovery and connection management.
66.9k
Brightdata Automation
Automate Brightdata web scraping and data collection operations through Composio's Brightdata toolkit via Rube MCP.
66.9k
Zenrows Automation
Automates Zenrows web scraping operations through Composio's Zenrows toolkit via Rube MCP, with tool discovery and connection management.
66.9k
More results
Scrape Do Automation
Automate web scraping and data extraction tasks using the Scrape Do toolkit via Rube MCP and Composio.
66.9k
Parsera Automation
Automate Parsera web scraping tasks through Composio's Parsera toolkit via Rube MCP, with dynamic tool discovery and connection management.
66.9k
Data Scraping
Builds a configurable scraping agent that collects data from APIs, HTML, or RSS, enriches it with Gemini AI scoring, and stores results in Notion, Google Sheets, Supabase, or local files.
1 · bundle
Web Scraper
Web scraping and content comprehension agent — multi-strategy extraction with cascade fallback, news detection, boilerplate removal, structured metadata, and LLM entity extraction
228 · bundle
Google News API Skill
Extracts structured news data from Google News via the BrowserAct API, including headlines, sources, publication times, and article links.
3.7k · bundle
Crawl4ai MCP Server
Self-hosted web crawling and content extraction exposed as MCP tools, with depth control and clean markdown output.
28
Firecrawl Agent
Extracts structured JSON data from complex multi-page websites using an AI agent that navigates pages and returns results matching a schema.
2
Diffbot Automation
Automate Diffbot data extraction and analysis through Composio's Diffbot toolkit via Rube MCP.
66.9k
Business Contact Social Links Skill
Extract official websites and social media profiles (LinkedIn, Facebook, X/Twitter, Instagram, YouTube, TikTok) from a company name or website URL using automated browser scripts.
3.7k · bundle
007 File 265f6921
Converts documentation websites, GitHub repositories, and PDF files into AI-ready skills for multiple LLM platforms.
7 · bundle
Browserbase Tool Automation
Automates Browserbase Tool operations through Composio's Rube MCP integration, including tool discovery, connection management, and execution.
66.9k
Agent Web Scraper
Web Scraper IA — Expert en extraction web (Scrapy, BeautifulSoup, Playwright, anti-bot, proxy rotation, data extraction)
6
Parsehub Automation
Automates Parsehub data extraction tasks through Composio's Parsehub toolkit via Rube MCP, with tool discovery and connection management.
66.9k
Zyte API Automation
Automate Zyte API operations through Composio's Rube MCP integration, including tool discovery, connection management, and execution.
66.9k
Browser Use
Use when an AI agent needs to control a browser, automate web tasks, scrape pages, fill forms, or click buttons autonomously. Triggers on: 'browser automation', 'web agent', 'browser-use', 'AI browse', 'tự động duyệt web', 'điều khiển trình duyệt', 'scrape with AI', 'click button automatically', 'fill form automatically', 'web task automation'.
2
Scrapegraph AI Automation
Automates Scrapegraph AI operations through Composio's toolkit via Rube MCP, with tool discovery and connection management.
66.9k
Youtube Search API Skill
Extracts structured data from YouTube search results, including videos, shorts, channels, and playlists, using the BrowserAct API.
3.7k · bundle
Tavily Extract
Extracts clean markdown or text from one or more URLs using the Tavily CLI, with support for JavaScript-rendered pages and query-focused chunking.
2
Youtube Transcript Extractor API Skill
Extracts YouTube video transcripts and metadata (title, publisher, likes) via the BrowserAct API without CAPTCHA or IP restrictions.
3.7k · bundle
Autobrowse
Builds reliable browser automation skills through iterative experimentation, running an inner agent to browse sites and improving navigation instructions until tasks pass consistently.
3.6k · bundle
Google Maps API Skill
Extracts structured business data from Google Maps, including names, categories, contact info, ratings, and addresses, using the BrowserAct API.
3.7k · bundle
Zhihu Search API Skill
Extracts structured article details and full content from Zhihu search results via the BrowserAct API, with keyword and date filtering.
3.7k · bundle
Aipex Browser
Controls a Chrome browser through the AIPex extension and MCP bridge, enabling navigation, clicking, form filling, screenshots, tab management, and content downloads.
17 · bundle
Google Image API Skill
Extracts structured image metadata from Google Images via the BrowserAct API, including thumbnails, titles, source logos, and click-through URLs.
3.7k · bundle
Amazon Product Search API Skill
Extracts structured product data from Amazon search results, including prices, ratings, sales estimates, and shipping info, using the BrowserAct API.
3.7k · bundle
Amazon Reviews API Skill
Extract Amazon product reviews by ASIN using the BrowserAct API, returning structured data including ratings, text, reviewer info, and verified purchase status.
3.7k · bundle
Amazon Best Selling Products Finder API Skill
Extract structured best-selling product data from Amazon, including titles, prices, ratings, reviews, sales volume, and promotions, using the BrowserAct API.
3.7k · bundle
Agent Browser
Control a headless browser to navigate pages, click elements, fill forms, take screenshots, record video, and execute JavaScript using Playwright and inference.sh.
584 · bundle
Data Scraper Agent
Builds a scheduled, AI-powered data collection agent that scrapes public sources, enriches results with Gemini Flash, and stores them in Notion, Sheets, or Supabase.
1 · bundle
Adhx
Fetch any X/Twitter post as clean LLM-friendly JSON, converting x.com, twitter.com, or adhx.com links into structured data with full article content, author info, and engagement metrics.
42.4k