Polars

Fast in-memory DataFrame library for datasets that fit in RAM. Use when pandas is too slow but data still fits in memory. Lazy evaluation, parallel execution, Apache Arrow backend. Best for 1-100GB datasets, ETL pipelines, faster pandas replacement. For larger-than-RAM data use dask or vaex.

synthetic-sciences 9a95c00 7 files · 76.9 KB Updated

File contents

synthetic-sciences/openscience/tree/main/backend/cli/skills/data-engineering/polars commit 9a95c00a4e

Frequently asked questions

npx skillmds@latest add synthetic-sciences/polars