Results for “data-intensive”

49 skills
More results
projectious-work
data-science
Data analysis workflow from import through modeling and communication. Use when analyzing a dataset, exploring data, building a statistical model, selecting features, or communicating findings to stakeholders.
0 · bundle
intense-visions
db-time-series
Time-Series Data
18 · bundle
bouclem
data-scientist
Expert data scientist for advanced analytics, machine learning, and statistical modeling. Handles complex data analysis, predictive modeling, and business intelligence.
7
bouclem
big-data
Apache Spark, Hadoop, distributed computing, and large-scale data processing for petabyte-scale workloads
7 · bundle
bdm-15
data-analyzer
Advanced data analysis, pattern detection, and insight generation from structured and unstructured datasets. Use when the user wants to analyze data, perform statistical analysis, find insights, detect patterns, identify anomalies, compare segments, test hypotheses, or generate data-driven recommendations. Triggers on phrases like 'analyze data', 'data analysis', 'find insights', 'analyze dataset', 'statistical analysis', 'find patterns', 'compare groups', 'test hypothesis', 'correlation analysis', or 'trend analysis'.
0 · bundle
jeffallan
pandas-pro
Perform efficient pandas DataFrame operations for data analysis, manipulation, and transformation with production-grade patterns.
10.4k · bundle
dokhacgiakhoa
data-engineer
Build scalable data pipelines, modern data warehouses, and real-time streaming architectures. Implements Apache Spark, dbt, Airflow, and cloud-native data platforms. Use PROACTIVELY for data pipeline design, analytics infrastructure, or modern data stack implementation.
505 · bundle
thanakijwanavit
data-workflow
Use this skill for any data or analytics task — querying databases, analyzing metrics, exploring data warehouses, processing datasets, or creating visualizations.
0
neuralblitz
big-data
Designs and implements big data architectures, processes large-scale datasets with distributed systems, and optimizes data pipelines for throughput using Hadoop, Spark, and cloud platforms.
1
antood69
data-workflow
Use this skill for any data or analytics task — querying databases, analyzing metrics, exploring data warehouses, processing datasets, or creating visualizations.
0
danstrem2
nosql-expert
Expert guidance for distributed NoSQL databases (Cassandra, DynamoDB). Focuses on mental models, query-first modeling, single-table design, and avoiding hot partitions in high-scale systems.
2
sirnosh
bmad-ml-gekko
Data pipeline specialist for ML experiments. Use when the user asks to talk to Gekko, requests the data engineer, or needs DataLoader optimization.
0 · bundle
bouclem
data-engineer
Build scalable data pipelines, modern data warehouses, and real-time streaming architectures. Implements Apache Spark, dbt, Airflow, and cloud-native data platforms.
7
galyarderlabs
data-room
Prepares, audits, and organizes fundraising materials for investor due diligence, generating a readiness checklist and folder structure.
20
rootcastleco
nosql-expert
Expert guidance for distributed NoSQL databases (Cassandra, DynamoDB). Focuses on mental models, query-first modeling, single-table design, and avoiding hot partitions in high-scale systems.
6
deanpeters
company-intel
Research companies, industries, or competitor sets using web search and seven analytical lenses to produce structured intelligence for downstream product management tasks.
5.6k
jiachen-t-wang
idefics2-an-8b-parameters-multimodal-model-arxiv-2405-02246v
Idefics2: An 8B Parameters Multimodal Model
6
ahang1598
doubao-data-analysis
结构化业务数据分析:附件读取与口径核验、定向筛选、规则/阈值判定、指标异动归因、漏斗/留存/实验分析、经营复盘及可审计报告。当用户提供 Excel、CSV、PDF、图片或多份业务材料,要求查数、判异常、解释变化、比较方案或形成行动建议时使用。Use for evidence-grounded analysis of structured business data, including filtering, rule checks, reconciliation, diagnostics, experiments, and decision reports.
9 · bundle
modbender
cdo-chief-data-officer
Drive data strategy with governance frameworks, analytics platforms, AI/ML initiatives, and privacy compliance.
12 · bundle
heath-gtm
heath-no-fluff
heath-no-fluff
0
dylanckawalec
postgresql-expert
Expert-level PostgreSQL database administration, advanced queries, performance tuning, and production operations
3
bobmatnyc
json-data-handling
Working effectively with JSON data structures.
71 · bundle
neuralblitz
big-data-based-analysis
Big Data Based Analysis Skill
1 · bundle
inskillflow
sql-pro
Master modern SQL with cloud-native databases, OLTP/OLAP optimization, and advanced query techniques. Expert in performance tuning, data modeling, and hybrid analytical systems.
1
26bb
inngest
Inngest expert for serverless-first background jobs, event-driven workflows, and durable execution without managing queues or workers.
0
jeffallan
sql-pro
Optimizes SQL queries, designs database schemas, and troubleshoots performance issues using execution plan analysis, indexing strategies, and set-based operations.
10.4k · bundle
nous-hermeshub
inngest
Inngest expert for serverless-first background jobs, event-driven
1
nexu-io
social-media-matrix
Build a cinematic, data-dense multi-platform social media dashboard with KPI matrix, interactive charts, insights drawer, and dark/light theme toggle.
· bundle
nexu-io
data-report
Converts CSV, Excel, or JSON data into a polished, interactive visual report page with KPI cards, charts, data tables, and insights.
· bundle
seb1n
data-labeling
Set up and manage data labeling workflows using manual annotation tools, semi-automated pipelines, active learning, and programmatic weak supervision. Use when the user requests data labeling or provides relevant inputs for this workflow.
159
bouclem
data-analyst
Data analysis best practices with pandas, numpy, matplotlib, seaborn, and Jupyter notebooks.
7
herdiansah
database-architect
Expert database architect specializing in data layer design from scratch, technology selection, schema modeling, and scalable database architectures. Masters SQL/NoSQL/TimeSeries database selection, normalization strategies, migration planning, and performance-first design. Handles both greenfield architectures and re-architecture of existing systems. Use PROACTIVELY for database architecture, technology selection, or data modeling decisions.
23
kensaurus
data-pipeline
Wire ETL, ingestion, cron, edge-function, and queue jobs correctly. Use for "build a pipeline", "sync X into Y", "nightly aggregation", "cron double-counts", "dedupe", "backfill", "the numbers are wrong after a retry". Bakes in idempotency, atomic writes, data contracts, dead-letter, and observability.
8
ranbot-ai
sql-pro
Master modern SQL with cloud-native databases, OLTP/OLAP optimization, and advanced query techniques. Expert in performance tuning, data modeling, and hybrid analytical systems.
6
jiachen-t-wang
autoaugment-learning-augmentation-strategies-from-data-arxiv
AutoAugment: Learning Augmentation Strategies from Data
6