Packs
3 packscurated
Social Media Scraping
Extract structured data from social media platforms via browser automation.
12 skills · pack
curated
SEO Audit and Fix
Audit a website for SEO issues, fix metadata and structured data, and verify improvements.
10 skills · pack
@adobe
Edge Delivery Services Content Ops
Content operations skills for AEM Edge Delivery Services: page auditing, SEO optimization, AI search (GEO), WCAG accessibility, bulk metadata, structured data, sitemap validation, and content diffing
12 skills · pack
Results for “data-structure”
42 skillshasdata
Extract public web data, search engine results, and structured data from platforms like Google, Amazon, and Zillow using HasData APIs.
42.4k · bundle
google-news-api-skill
Extracts structured news data from Google News via the BrowserAct API, including headlines, sources, publication times, and article links.
3.7k · bundle
dbt
Provides dbt patterns for data transformation and analytics engineering, including model structures, incremental models, and testing.
54 · bundle
pymatgen
Analyze and manipulate crystal structures, compute phase diagrams, and access the Materials Project database using the pymatgen library.
30.2k · bundle
tao-validate-dataset-format
Validates NVIDIA TAO DAFT datasets for structure, schema, and cross-reference errors using the `tao-daft validate` CLI tool.
2.2k · bundle
data-profiling
Profiles datasets automatically to assess data quality, structure, and completeness, generating reports with ydata-profiling, pandera, or manual pandas methods.
0 · bundle
More results
spark-engineer
Write, optimize, and debug Apache Spark jobs for high-performance distributed data processing, ETL pipelines, and big data workloads.
10.4k · bundle
data-analysis
Guide through a structured data analysis workflow: define the question, validate data quality, select the appropriate analytical method, and produce decision-ready findings with caveats.
42 · bundle
bids
Organize, query, validate, and convert neuroscience and biomedical data using the Brain Imaging Data Structure (BIDS) standard.
30.2k · bundle
bids
Organize, query, validate, and convert neuroscience datasets following the Brain Imaging Data Structure (BIDS) standard, including metadata sidecars and derivatives.
253 · bundle
google-maps-search-api-skill
Extracts structured business data from Google Maps search results using the BrowserAct API. Provide search keywords, language, and country filters to get clean, usable business data.
3.7k · bundle
datadog-logs
Query and filter Datadog logs from the shell using the Composio CLI, enabling scoped log searches, pivoting across services and environments, and exporting structured JSON for downstream analysis.
66.9k
gget
Queries 20+ bioinformatics databases from the command line or Python for gene info, sequences, BLAST/BLAT, protein structures, viral data, and expression metrics.
253 · bundle
youtube-search-api-skill
Extracts structured data from YouTube search results, including videos, shorts, channels, and playlists, using the BrowserAct API.
3.7k · bundle
youtube-video-api-skill
Extracts structured channel-level and video detail data from a YouTube channel via the BrowserAct API, including metrics like views, likes, comments, and subscriber count.
3.7k · bundle
google-maps-api-skill
Extracts structured business data from Google Maps, including names, categories, contact info, ratings, and addresses, using the BrowserAct API.
3.7k · bundle
indeed-job-search
Extract structured job listing data from Indeed search results, including job titles, companies, salaries, descriptions, benefits, and apply links.
3.7k · bundle
amazon-product-search-api-skill
Extracts structured product data from Amazon search results, including prices, ratings, sales estimates, and shipping info, using the BrowserAct API.
3.7k · bundle
youtube-channel-api-skill
Extracts structured channel data from YouTube search results via the BrowserAct API, including subscriber counts, verification status, and channel descriptions.
3.7k · bundle
162-use-e70e84f3
Provides guidance on using Apache Spark RDDs, including creation, transformations, actions, and performance considerations.
7 · bundle
big-data
Designs and implements big data architectures, processes large-scale datasets with distributed systems, and optimizes data pipelines for throughput using Hadoop, Spark, and cloud platforms.
1
mathguard
Guides AI agents to apply advanced mathematical and probabilistic techniques (Bloom filters, HyperLogLog, FFT, etc.) for large-scale data problems where classical algorithms are optimal but math offers better asymptotic bounds.
42.4k
ddia-systems
Design reliable, scalable, and maintainable data systems by applying principles from storage engines, replication, partitioning, transactions, and consistency models.
1.6k · bundle
data-validation
Define and enforce data schemas and quality checks using pandera, Great Expectations, or manual assertions, with clear error handling and documentation.
0 · bundle
amazon-best-selling-products-finder-api-skill
Extract structured best-selling product data from Amazon, including titles, prices, ratings, reviews, sales volume, and promotions, using the BrowserAct API.
3.7k · bundle
hasdata
Extract public web data via HasData APIs, including search engine results, structured data from ecommerce, travel, jobs, and local business platforms, with support for web scraping, pre-parsed APIs, and async jobs.
3 · bundle
hla-typing
Performs HLA allele genotyping from WGS/WES VCF data, producing a structured markdown report and machine-readable JSON results.
17 · bundle
biopython
Provides reference documentation and code patterns for using Biopython to handle biological sequences, file formats, database access, alignments, structures, and phylogenetics.
2
amazon-product-api-skill
Extracts structured product listings from Amazon, including titles, ASINs, prices, ratings, and specifications, using the BrowserAct API.
3.7k · bundle
producthunt-launches
Extract structured product launch data from Product Hunt leaderboards, enriched with maker profiles and website contact information.
3.7k · bundle
gget
Query 20+ bioinformatics databases from the command line or Python for gene information, sequences, protein structures, enrichment analysis, and more.
30.2k · bundle
hasdata-cli
Provides command-line access to search, scraping, and structured web data from over 40 APIs including Google, Amazon, Yelp, and Zillow.
42.4k · bundle
power-bi-performance-troubleshooting
Systematically diagnose and resolve performance issues in Power BI models, reports, and queries using a structured troubleshooting methodology.
36.2k
azure-ai-document-intelligence-ts
Extract text, tables, and structured data from documents using Azure Document Intelligence. Process invoices, receipts, IDs, forms, or build custom document models.
2.7k
biopython
Provides reference documentation and code patterns for Biopython, covering sequence handling, alignments, NCBI database access, BLAST, protein structures, phylogenetics, and other bioinformatics tasks.
5
amazon-reviews-api-skill
Extract Amazon product reviews by ASIN using the BrowserAct API, returning structured data including ratings, text, reviewer info, and verified purchase status.
3.7k · bundle