Ray Data

Scalable data processing for ML workloads. Streaming execution across CPU/GPU, supports Parquet/CSV/JSON/images. Integrates with Ray Train, PyTorch, TensorFlow. Scales from single machine to 100s of nodes. Use for batch inference, data preprocessing, multi-modal data loading, or distributed ETL pipelines. Use when this capability is needed.

tomevault-io 2ab740e 2 files · 7.8 KB Updated

File contents

tomevault-io/skills-registry/tree/main/orchestra-research--ai-research-skills--ray-data commit 2ab740e9a9

Frequently asked questions

npx skillmds@latest add tomevault-io/ray-data