ETL & Pipelines
-
bankrbot Bundle Token Scam AnalysisPerform forensic on-chain analysis of EVM tokens to detect scams, rug pulls, and soft rugs by cross-referencing on-chain state against team narratives.
1.2k -
composiohq Skill Datagma AutomationAutomate Datagma operations through Composio's Datagma toolkit via Rube MCP, with tool discovery and connection management.
66.9k -
composiohq Skill Parseur AutomationAutomate Parseur document parsing operations through Composio's Parseur toolkit via Rube MCP.
66.9k -
composiohq Skill Listclean AutomationAutomates Listclean data-cleaning operations through Composio's toolkit via Rube MCP, with tool discovery and connection management.
66.9k -
composiohq Skill Scrape Do AutomationAutomate web scraping and data extraction tasks using the Scrape Do toolkit via Rube MCP and Composio.
66.9k -
composiohq Skill Brightdata AutomationAutomate Brightdata web scraping and data collection operations through Composio's Brightdata toolkit via Rube MCP.
66.9k -
composiohq Skill Stormglass Io AutomationAutomate Stormglass IO operations through Composio's Stormglass IO toolkit via Rube MCP, including tool discovery, connection management, and execution.
66.9k -
browser-act Bundle Browser Act Skill ForgeTurns any website's data extraction or operation needs into reusable Agent-callable Skill packages by exploring API endpoints or DOM methods, then generating SKILL.md and Python scripts.
Audited 3.7k -
browser-act Bundle Webcrawler Deep CrawlDeep-crawl any website from start URLs, returning per-page LLM-ready text, markdown, or HTML with metadata and in-scope outbound links.
3.7k -
browser-act Bundle Amazon Asin Lookup API SkillExtract structured product details from Amazon using an ASIN, including title, price, ratings, brand, and description via the BrowserAct API.
3.7k -
k-dense-ai Bundle DaskScale pandas and NumPy workflows to larger-than-memory datasets using parallel and distributed computing.
Audited 30.2k -
k-dense-ai Bundle VaexProcess and analyze large tabular datasets (billions of rows) that exceed available RAM using lazy, out-of-core DataFrames with fast aggregations, visualization, and machine learning integration.
Audited 30.2k -
k-dense-ai Bundle PolarsProcess data with high-performance DataFrames using Polars' expression-based API, lazy evaluation, and parallel execution for ETL, analytics, and pandas migration.
Audited 30.2k -
k-dense-ai Bundle NextflowBuild, run, and debug Nextflow data pipelines and nf-core workflows end to end, covering processes, channels, operators, configuration, testing, and deployment to HPC or cloud.
30.2k -
k-dense-ai Bundle PacsomaticValidates inputs, generates samplesheets and launch scripts, and optionally executes nf-core/pacsomatic matched tumor-normal workflows from BAM files, supporting local runs and scheduler submission (LSF/Slurm/PBS/SGE).
30.2k -
k-dense-ai Bundle PylabrobotControl liquid handling robots, plate readers, pumps, and other lab equipment through a unified Python interface across platforms.
Audited 30.2k -
k-dense-ai Bundle BioservicesQuery 40+ bioinformatics services (UniProt, KEGG, ChEMBL, Reactome) with a unified Python interface for cross-database analysis, identifier mapping, and sequence analysis.
30.2k -
k-dense-ai Bundle Bulk RnaseqOrchestrates a complete bulk RNA-seq differential-expression study from raw FASTQ reads through QC, alignment, quantification, differential expression, pathway enrichment, and publication figures.
30.2k -
k-dense-ai Bundle Zarr PythonStore and process large N-dimensional arrays with chunking, compression, and parallel I/O, integrating with NumPy, Dask, and Xarray for cloud-native scientific computing.
Audited 30.2k -
k-dense-ai Bundle Dnanexus IntegrationBuild and deploy apps/applets on the DNAnexus cloud genomics platform, manage data objects, run workflows, and use the dxpy Python SDK for genomics pipeline development and execution.
30.2k -
k-dense-ai Bundle Latchbio IntegrationBuild and deploy bioinformatics workflows as serverless pipelines on the Latch platform using Python decorators, cloud data management, and GPU support.
Audited 30.2k -
k-dense-ai Bundle Benchling IntegrationIntegrate with Benchling's Python SDK and REST API to manage registry entities, inventory, ELN entries, workflows, and Data Warehouse queries for life sciences R&D automation.
30.2k -
k-dense-ai Bundle Opentrons IntegrationWrite Opentrons Protocol API v2 protocols for Flex and OT-2 robots to automate liquid handling, control hardware modules, and manage labware configurations.
Audited 30.2k -
k-dense-ai Bundle Labarchive IntegrationAccess and manage LabArchives electronic lab notebooks programmatically via REST API. Create entries, upload attachments, backup notebooks, generate reports, and integrate with Protocols.io, Jupyter, REDCap, and other scientific tools.
30.2k -
mukul975 Bundle Processing Stix Taxii FeedsProcesses STIX 2.1 threat intelligence bundles from TAXII 2.1 servers, normalizing objects into platform-native schemas and routing them to consuming systems.
24.6k -
mukul975 Bundle Analyzing Threat Intelligence FeedsIngests, normalizes, and enriches structured and unstructured threat intelligence feeds into STIX 2.1 format, evaluating feed quality and deduplicating indicators for distribution to SIEM, firewall, and EDR platforms.
24.6k -
mukul975 Bundle Performing Ioc Enrichment AutomationAutomates multi-source enrichment of IPs, domains, URLs, and file hashes using VirusTotal, AbuseIPDB, Shodan, GreyNoise, URLScan.io, and MISP to provide contextual risk scoring and disposition recommendations for SOC analysts.
24.6k -
mukul975 Bundle Implementing Stix Taxii Feed IntegrationConsume and produce STIX/TAXII 2.1 cyber threat intelligence feeds using Python, including server discovery, collection polling, object parsing, and SIEM/TIP integration.
Audited 24.6k -
mukul975 Bundle Implementing Siem Correlation Rules For AptDetect APT lateral movement by chaining Windows authentication events, process execution telemetry, and network connection logs across hosts using Splunk SPL and Sigma rule format.
24.6k -
jeffallan Bundle Ml PipelineDesigns and implements production-grade ML pipeline infrastructure: configures experiment tracking, creates orchestration DAGs, builds feature store schemas, deploys model registries, and automates retraining and validation workflows.
Audited 10.4k -
muratcankoylan Bundle Book Sft PipelineConvert books into supervised fine-tuning datasets and train style-transfer models that replicate an author's voice.
16.9k -
orchestra-research Bundle Ray DataProcess large-scale ML datasets with distributed streaming execution across CPU/GPU, supporting Parquet, CSV, JSON, images, and integration with PyTorch, TensorFlow, and Ray Train.
Audited 10.4k -
orchestra-research Bundle Nemo CuratorGPU-accelerated data curation for LLM training, supporting text, image, video, and audio with fuzzy deduplication, quality filtering, semantic deduplication, PII redaction, and NSFW detection.
Audited 10.4k -
czlonkowski Bundle N8n Binary And DataHandle files and binary data in n8n workflows correctly, covering the $binary vs $json split, reading/writing binary, preserving binary across transforms, and the agent-tool binary boundary.
Audited 5.7k -
tradermonty Bundle Edge Pipeline OrchestratorCoordinate multi-stage edge research pipelines from candidate detection through strategy design, review, revision, and export.
2.3k -
albedo-tabai Bundle Lets Go RssAggregate RSS feeds from YouTube, Vimeo, Behance, Twitter/X, Bilibili, Weibo, Douyin, Xiaohongshu, and Zhihu with incremental updates, deduplication, and AI classification.
99