pdfplumber Python PDF Text and Table Extraction Library
pdfplumber is a Python library for extracting detailed information from PDFs — text, tables, lines, rectangles, and curves — with visual debugging support. Built on pdfminer.six, it excels at structured table extraction from machine-generated PDFs and includes both a Python API and CLI.
Installation
Use the upstream install or setup path that matches your environment:
- pip install pdfplumber
Requirements and caveats from upstream:
- [![Code coverage](https://codecov.io/gh/jsvi...
- Currently tested on Python 3.10, 3.11, 3.12, 3.13, 3.14.
- Python library
Basic usage or getting-started notes:
sh
Command line interface
Basic example
Extracted from upstream docs: https://raw.githubusercontent.com/jsvine/pdfplumber/HEAD/README.md