PPTX creation, editing, and analysis
Overview
A .pptx file is a ZIP archive containing XML files and resources. Create, edit, or analyze PowerPoint presentations using text extraction, raw XML access, or html2pptx workflows. Apply this skill for programmatic presentation creation and modification.
Visual Enhancement with Scientific Schematics
When creating documents with this skill, always consider adding scientific diagrams and schematics to enhance visual communication.
If your document does not already contain schematics or diagrams:
- Use the scientific-schematics skill to generate AI-powered publication-quality diagrams
- Simply describe your desired diagram in natural language
- Nano Banana Pro will automatically generate, review, and refine the schematic
For new documents: Scientific schematics should be generated by default to visually represent key concepts, workflows, architectures, or relationships described in the text.
How to generate schematics:
python scripts/generate_schematic.py "your diagram description" -o figures/output.png
The AI will automatically:
- Create publication-quality images with proper formatting
- Review and refine through multiple iterations
- Ensure accessibility (colorblind-friendly, high contrast)
- Save outputs in the figures/ directory
When to add schematics:
- Presentation workflow diagrams for slides
- Slide design process flowcharts
- Content organization diagrams
- System architecture illustrations
- Process flow visualizations
- Any complex concept that benefits from visualization
For detailed guidance on creating schematics, refer to the scientific-schematics skill documentation.
Reading and analyzing content
Text extraction
To read the text contents of a presentation, convert the document to markdown:
# Convert document to markdown
python -m markitdown path-to-file.pptx
Raw XML access
Raw XML access is required for: comments, speaker notes, slide layouts, animations, design elements, and complex formatting. For any of these features, unpack a presentation and read its raw XML contents.
Unpacking a file
python ooxml/scripts/unpack.py <office_file> <output_dir>
Note: The unpack.py script is located at skills/pptx/ooxml/scripts/unpack.py relative to the project root. If the script doesn't exist at this path, use find . -name "unpack.py" to locate it.
Key file structures
ppt/presentation.xml - Main presentation metadata and slide references
ppt/slides/slide{N}.xml - Individual slide contents (slide1.xml, slide2.xml, etc.)
ppt/notesSlides/notesSlide{N}.xml - Speaker notes for each slide
ppt/comments/modernComment_*.xml - Comments for specific slides
ppt/slideLayouts/ - Layout templates for slides
ppt/slideMasters/ - Master slide templates
ppt/theme/ - Theme and styling information
ppt/media/ - Images and other media files
Typography and color extraction
When given an example design to emulate: Always analyze the presentation's typography and colors first using the methods below:
- Read theme file: Check
ppt/theme/theme1.xml for colors (<a:clrScheme>) and fonts (<a:fontScheme>)
- Sample slide content: Examine
ppt/slides/slide1.xml for actual font usage (<a:rPr>) and colors
- Search for patterns: Use grep to find color (
<a:solidFill>, <a:srgbClr>) and font references across all XML files
Creating a new PowerPoint presentation without a template
When creating a new PowerPoint presentation from scratch, use the html2pptx workflow to convert HTML slides to PowerPoint with accurate positioning.
Design Principles
CRITICAL: Before creating any presentation, analyze the content and choose appropriate design elements:
- Consider the subject matter: What is this presentation about? What tone, industry, or mood does it suggest?
- Check for branding: If the user mentions a company/organization, consider their brand colors and identity
- Match palette to content: Select colors that reflect the subject
- State your approach: Explain your design choices before writing code
Requirements:
- ✅ State your content-informed design approach BEFORE writing code
- ✅ Use web-safe fonts only: Arial, Helvetica, Times New Roman, Georgia, Courier New, Verdana, Tahoma, Trebuchet MS, Impact
- ✅ Create clear visual hierarchy through size, weight, and color
- ✅ Ensure readability: strong contrast, appropriately sized text, clean alignment
- ✅ Be consistent: repeat patterns, spacing, and visual language across slides
Color Palette Selection
Choosing colors creatively:
- Think beyond defaults: What colors genuinely match this specific topic? Avoid autopilot choices.
- Consider multiple angles: Topic, industry, mood, energy level, target audience, brand identity (if mentioned)
- Be adventurous: Try unexpected combinations - a healthcare presentation doesn't have to be green, finance doesn't have to be navy
- Build your palette: Pick 3-5 colors that work together (dominant colors + supporting tones + accent)
- Ensure contrast: Text must be clearly readable on backgrounds
Example color palettes (use these to spark creativity - choose one, adapt it, or create your own):
- Classic Blue: Deep navy (#1C2833), slate gray (#2E4053), silver (#AAB7B8), off-white (#F4F6F6)
- Teal & Coral: Teal (#5EA8A7), deep teal (#277884), coral (#FE4447), white (#FFFFFF)
- Bold Red: Red (#C0392B), bright red (#E74C3C), orange (#F39C12), yellow (#F1C40F), green (#2ECC71)
- Warm Blush: Mauve (#A49393), blush (#EED6D3), rose (#E8B4B8), cream (#FAF7F2)
- Burgundy Luxury: Burgundy (#5D1D2E), crimson (#951233), rust (#C15937), gold (#997929)
- Deep Purple & Emerald: Purple (#B165FB), dark blue (#181B24), emerald (#40695B), white (#FFFFFF)
- Cream & Forest Green: Cream (#FFE1C7), forest green (#40695B), white (#FCFCFC)
- Pink & Purple: Pink (#F8275B), coral (#FF574A), rose (#FF737D), purple (#3D2F68)
- Lime & Plum: Lime (#C5DE82), plum (#7C3A5F), coral (#FD8C6E), blue-gray (#98ACB5)
- Black & Gold: Gold (#BF9A4A), black (#000000), cream (#F4F6F6)
- Sage & Terracotta: Sage (#87A96B), terracotta (#E07A5F), cream (#F4F1DE), charcoal (#2C2C2C)
- Charcoal & Red: Charcoal (#292929), red (#E33737), light gray (#CCCBCB)
- Vibrant Orange: Orange (#F96D00), light gray (#F2F2F2), charcoal (#222831)
- Forest Green: Black (#191A19), green (#4E9F3D), dark green (#1E5128), white (#FFFFFF)
- Retro Rainbow: Purple (#722880), pink (#D72D51), orange (#EB5C18), amber (#F08800), gold (#DEB600)
- Vintage Earthy: Mustard (#E3B448), sage (#CBD18F), forest green (#3A6B35), cream (#F4F1DE)
- Coastal Rose: Old rose (#AD7670), beaver (#B49886), eggshell (#F3ECDC), ash gray (#BFD5BE)
- Orange & Turquoise: Light orange (#FC993E), grayish turquoise (#667C6F), white (#FCFCFC)
Visual Details Options
Geometric Patterns:
- Diagonal section dividers instead of horizontal
- Asymmetric column widths (30/70, 40/60, 25/75)
- Rotated text headers at 90° or 270°
- Circular/hexagonal frames for images
- Triangular accent shapes in corners
- Overlapping shapes for depth
Border & Frame Treatments:
- Thick single-color borders (10-20pt) on one side only
- Double-line borders with contrasting colors
- Corner brackets instead of full frames
- L-shaped borders (top+left or bottom+right)
- Underline accents beneath headers (3-5pt thick)
Typography Treatments:
- Extreme size contrast (72pt headlines vs 11pt body)
- All-caps headers with wide letter spacing
- Numbered sections in oversized display type
- Monospace (Courier New) for data/stats/technical content
- Condensed fonts (Arial Narrow) for dense information
- Outlined text for emphasis
Chart & Data Styling:
- Monochrome charts with single accent color for key data
- Horizontal bar charts instead of vertical
- Dot plots instead of bar charts
- Minimal gridlines or none at all
- Data labels directly on elements (no legends)
- Oversized numbers for key metrics
Layout Innovations:
- Full-bleed images with text overlays
- Sidebar column (20-30% width) for navigation/context
- Modular grid systems (3×3, 4×4 blocks)
- Z-pattern or F-pattern content flow
- Floating text boxes over colored shapes
- Magazine-style multi-column layouts
Background Treatments:
- Solid color blocks occupying 40-60% of slide
- Gradient fills (vertical or diagonal only)
- Split backgrounds (two colors, diagonal or vertical)
- Edge-to-edge color bands
- Negative space as a design element
Layout Tips
For slides with charts or tables:
- Two-column layout (PREFERRED): Use a header spanning the full width, then two columns below - text/bullets in one column and the featured content in the other. This provides better balance and makes charts/tables more readable. Use flexbox with unequal column widths (e.g., 40%/60% split) to optimize space for each content type.
- Full-slide layout: Let the featured content (chart/table) take up the entire slide for maximum impact and readability
- NEVER vertically stack: Do not place charts/tables below text in a single column - this causes poor readability and layout issues
Workflow
- MANDATORY - READ ENTIRE FILE: Read
html2pptx.md completely from start to finish. NEVER set any range limits when reading this file. Read the full file content for detailed syntax, critical formatting rules, and best practices before proceeding with presentation creation.
- Create an HTML file for each slide with proper dimensions (e.g., 720pt × 405pt for 16:9)
- Use
<p>, <h1>-<h6>, <ul>, <ol> for all text content
- Use
class="placeholder" for areas where charts/tables will be added (render with gray background for visibility)
- CRITICAL: Rasterize gradients and icons as PNG images FIRST using Sharp, then reference in HTML
- LAYOUT: For slides with charts/tables/images, use either full-slide layout or two-column layout for better readability
- Create and run a JavaScript file using the
html2pptx.js library to convert HTML slides to PowerPoint and save the presentation
- Use the
html2pptx() function to process each HTML file
- Add charts and tables to placeholder areas using PptxGenJS API
- Save the presentation using
pptx.writeFile()
- Visual validation: Generate thumbnails and inspect for layout issues
- Create thumbnail grid:
python scripts/thumbnail.py output.pptx workspace/thumbnails --cols 4
- Read and carefully examine the thumbnail image for:
- Text cutoff: Text being cut off by header bars, shapes, or slide edges
- Text overlap: Text overlapping with other text or shapes
- Positioning issues: Content too close to slide boundaries or other elements
- Contrast issues: Insufficient contrast between text and backgrounds
- If issues found, adjust HTML margins/spacing/colors and regenerate the presentation
- Repeat until all slides are visually correct
Editing an existing PowerPoint presentation
To edit slides in an existing PowerPoint presentation, work with the raw Office Open XML (OOXML) format. This involves unpacking the .pptx file, editing the XML content, and repacking it.
Workflow
- MANDATORY - READ ENTIRE FILE: Read
ooxml.md (~500 lines) completely from start to finish. NEVER set any range limits when reading this file. Read the full file content for detailed guidance on OOXML structure and editing workflows before any presentation editing.
- Unpack the presentation:
python ooxml/scripts/unpack.py <office_file> <output_dir>
- Edit the XML files (primarily
ppt/slides/slide{N}.xml and related files)
- CRITICAL: Validate immediately after each edit and fix any validation errors before proceeding:
python ooxml/scripts/validate.py <dir> --original <file>
- Pack the final presentation:
python ooxml/scripts/pack.py <input_directory> <office_file>
Creating a new PowerPoint presentation using a template
To create a presentation that follows an existing template's design, duplicate and re-arrange template slides before replacing placeholder context.
Workflow
Extract template text AND create visual thumbnail grid:
- Extract text:
python -m markitdown template.pptx > template-content.md
- Read
template-content.md: Read the entire file to understand the contents of the template presentation. NEVER set any range limits when reading this file.
- Create thumbnail grids:
python scripts/thumbnail.py template.pptx
- See Creating Thumbnail Grids section for more details
Analyze template and save inventory to a file:
- Visual Analysis: Review thumbnail grid(s) to understand slide layouts, design patterns, and visual structure
- Create and save a template inventory file at
template-inventory.md containing:# Template Inventory Analysis
**Total Slides: [count]**
**IMPORTANT: Slides are 0-indexed (first slide = 0, last slide = count-1)**
## [Category Name]
- Slide 0: [Layout code if available] - Description/purpose
- Slide 1: [Layout code] - Description/purpose
- Slide 2: [Layout code] - Description/purpose
[... EVERY slide must be listed individually with its index ...]
- Using the thumbnail grid: Reference the visual thumbnails to identify:
- Layout patterns (title slides, content layouts, section dividers)
- Image placeholder locations and counts
- Design consistency across slide groups
- Visual hierarchy and structure
- This inventory file is REQUIRED for selecting appropriate templates in the next step
Create presentation outline based on template inventory:
- Review available templates from step 2.
- Choose an intro or title template for the first slide. This should be one of the first templates.
- Choose safe, text-based layouts for the other slides.
- CRITICAL: Match layout structure to actual content:
- Single-column layouts: Use for unified narrative or single topic
- Two-column layouts: Use ONLY when there are exactly 2 distinct items/concepts
- Three-column layouts: Use ONLY when there are exactly 3 distinct items/concepts
- Image + text layouts: Use ONLY when actual images are available to insert
- Quote layouts: Use ONLY for actual quotes from people (with attribution), never for emphasis
- Never use layouts with more placeholders than available content
- If there are 2 items, don't force them into a 3-column layout
- If there are 4+ items, consider breaking into multiple slides or using a list format
- Count actual content pieces BEFORE selecting the layout
- Verify each placeholder in the chosen layout will be filled with meaningful content
- Select one option representing the best layout for each content section.
- Save
outline.md with content AND template mapping that leverages available designs
- Example template mapping:
# Template slides to use (0-based indexing)
# WARNING: Verify indices are within range! Template with 73 slides has indices 0-72
# Mapping: slide numbers from outline -> template slide indices
template_mapping = [
0, # Use slide 0 (Title/Cover)
34, # Use slide 34 (B1: Title and body)
34, # Use slide 34 again (duplicate for second B1)
50, # Use slide 50 (E1: Quote)
54, # Use slide 54 (F2: Closing + Text)
]
Duplicate, reorder, and delete slides using rearrange.py:
- Use the
scripts/rearrange.py script to create a new presentation with slides in the desired order:python scripts/rearrange.py template.pptx working.pptx 0,34,34,50,52
- The script handles duplicating repeated slides, deleting unused slides, and reordering automatically
- Slide indices are 0-based (first slide is 0, second is 1, etc.)
- The same slide index can appear multiple times to duplicate that slide
Extract ALL text using the inventory.py script:
Run inventory extraction:
python scripts/inventory.py working.pptx text-inventory.json
Read text-inventory.json: Read the entire text-inventory.json file to understand all shapes and their properties. NEVER set any range limits when reading this file.
The inventory JSON structure:
{
"slide-0": {
"shape-0": {
"placeholder_type": "TITLE", // or null for non-placeholders
"left": 1.5, // position in inches
"top": 2.0,
"width": 7.5,
"height": 1.2,
"paragraphs": [
{
"text": "Paragraph text",
// Optional properties (only included when non-default):
"bullet": true, // explicit bullet detected
"level": 0, // only included when bullet is true
"alignment": "CENTER", // CENTER, RIGHT (not LEFT)
"space_before": 10.0, // space before paragraph in points
"space_after": 6.0, // space after paragraph in points
"line_spacing": 22.4, // line spacing in points
"font_name": "Arial", // from first run
"font_size": 14.0, // in points
"bold": true,
"italic": false,
"underline": false,
"color": "FF0000" // RGB color
}
]
}
}
}
Key features:
- Slides: Named as "slide-0", "slide-1", etc.
- Shapes: Ordered by visual position (top-to-bottom, left-to-right) as "shape-0", "shape-1", etc.
- Placeholder types: TITLE, CENTER_TITLE, SUBTITLE, BODY, OBJECT, or null
- Default font size:
default_font_size in points extracted from layout placeholders (when available)
- Slide numbers are filtered: Shapes with SLIDE_NUMBER placeholder type are automatically excluded from inventory
- Bullets: When
bullet: true, level is always included (even if 0)
- Spacing:
space_before, space_after, and line_spacing in points (only included when set)
- Colors:
color for RGB (e.g., "FF0000"), theme_color for theme colors (e.g., "DARK_1")
- Properties: Only non-default values are included in the output
Generate replacement text and save the data to a JSON file
Based on the text inventory from the previous step:
- CRITICAL: First verify which shapes exist in the inventory - only reference shapes that are actually present
- VALIDATION: The replace.py script will validate that all shapes in the replacement JSON exist in the inventory
- If a non-existent shape is referenced, an error will show available shapes
- If a non-existent slide is referenced, an error will indicate the slide doesn't exist
- All validation errors are shown at once before the script exits
- IMPORTANT: The replace.py script uses inventory.py internally to identify ALL text shapes
- AUTOMATIC CLEARING: ALL text shapes from the inventory will be cleared unless you provide "paragraphs" for them
- Add a "paragraphs" field to shapes that need content (not "replacement_paragraphs")
- Shapes without "paragraphs" in the replacement JSON will have their text cleared automatically
- Paragraphs with bullets will be automatically left aligned. Don't set the
alignment property on when "bullet": true
- Generate appropriate replacement content for placeholder text
- Use shape size to determine appropriate content length
- CRITICAL: Include paragraph properties from the original inventory - don't just provide text
- IMPORTANT: When bullet: true, do NOT include bullet symbols (•, -, *) in text - they are added automatically
- ESSENTIAL FORMATTING RULES:
- Headers/titles should typically have
"bold": true
- List items should have
"bullet": true, "level": 0 (level is required when bullet is true)
- Preserve any alignment properties (e.g.,
"alignment": "CENTER" for centered text)
- Include font properties when different from default (e.g.,
"font_size": 14.0, "font_name": "Lora")
- Colors: Use
"color": "FF0000" for RGB or "theme_color": "DARK_1" for theme colors
- The replacement script expects properly formatted paragraphs, not just text strings
- Overlapping shapes: Prefer shapes with larger default_font_size or more appropriate placeholder_type
- Save the updated inventory with replacements to
replacement-text.json
- WARNING: Different template layouts have different shape counts - always check the actual inventory before creating replacements
Example paragraphs field showing proper formatting:
"paragraphs": [
{
"text": "New presentation title text",
"alignment": "CENTER",
"bold": true
},
{
"text": "Section Header",
"bold": true
},
{
"text": "First bullet point without bullet symbol",
"bullet": true,
"level": 0
},
{
"text": "Red colored text",
"color": "FF0000"
},
{
"text": "Theme colored text",
"theme_color": "DARK_1"
},
{
"text": "Regular paragraph text without special formatting"
}
]
Shapes not listed in the replacement JSON are automatically cleared:
{
"slide-0": {
"shape-0": {
"paragraphs": [...] // This shape gets new text
}
// shape-1 and shape-2 from inventory will be cleared automatically
}
}
Common formatting patterns for presentations:
- Title slides: Bold text, sometimes centered
- Section headers within slides: Bold text
- Bullet lists: Each item needs
"bullet": true, "level": 0
- Body text: Usually no special properties needed
- Quotes: May have special alignment or font properties
Apply replacements using the replace.py script
python scripts/replace.py working.pptx replacement-text.json output.pptx
The script will:
- First extract the inventory of ALL text shapes using functions from inventory.py
- Validate that all shapes in the replacement JSON exist in the inventory
- Clear text from ALL shapes identified in the inventory
- Apply new text only to shapes with "paragraphs" defined in the replacement JSON
- Preserve formatting by applying paragraph properties from the JSON
- Handle bullets, alignment, font properties, and colors automatically
- Save the updated presentation
Example validation errors:
ERROR: Invalid shapes in replacement JSON:
- Shape 'shape-99' not found on 'slide-0'. Available shapes: shape-0, shape-1, shape-4
- Slide 'slide-999' not found in inventory
ERROR: Replacement text made overflow worse in these shapes:
- slide-0/shape-2: overflow worsened by 1.25" (was 0.00", now 1.25")
Creating Thumbnail Grids
To create visual thumbnail grids of PowerPoint slides for quick analysis and reference:
python scripts/thumbnail.py template.pptx [output_prefix]
Features:
- Creates:
thumbnails.jpg (or thumbnails-1.jpg, thumbnails-2.jpg, etc. for large decks)
- Default: 5 columns, max 30 slides per grid (5×6)
- Custom prefix:
python scripts/thumbnail.py template.pptx my-grid
- Note: The output prefix should include the path if you want output in a specific directory (e.g.,
workspace/my-grid)
- Adjust columns:
--cols 4 (range: 3-6, affects slides per grid)
- Grid limits: 3 cols = 12 slides/grid, 4 cols = 20, 5 cols = 30, 6 cols = 42
- Slides are zero-indexed (Slide 0, Slide 1, etc.)
Use cases:
- Template analysis: Quickly understand slide layouts and design patterns
- Content review: Visual overview of entire presentation
- Navigation reference: Find specific slides by their visual appearance
- Quality check: Verify all slides are properly formatted
Examples:
# Basic usage
python scripts/thumbnail.py presentation.pptx
# Combine options: custom name, columns
python scripts/thumbnail.py template.pptx analysis --cols 4
Converting Slides to Images
To visually analyze PowerPoint slides, convert them to images using a two-step process:
Convert PPTX to PDF:
soffice --headless --convert-to pdf template.pptx
Convert PDF pages to JPEG images:
pdftoppm -jpeg -r 150 template.pdf slide
This creates files like slide-1.jpg, slide-2.jpg, etc.
Options:
-r 150: Sets resolution to 150 DPI (adjust for quality/size balance)
-jpeg: Output JPEG format (use -png for PNG if preferred)
-f N: First page to convert (e.g., -f 2 starts from page 2)
-l N: Last page to convert (e.g., -l 5 stops at page 5)
slide: Prefix for output files
Example for specific range:
pdftoppm -jpeg -r 150 -f 2 -l 5 template.pdf slide # Converts only pages 2-5
Code Style Guidelines
IMPORTANT: When generating code for PPTX operations:
- Write concise code
- Avoid verbose variable names and redundant operations
- Avoid unnecessary print statements
Dependencies
Required dependencies (should already be installed):
- markitdown:
pip install "markitdown[pptx]" (for text extraction from presentations)
- pptxgenjs:
npm install -g pptxgenjs (for creating presentations via html2pptx)
- playwright:
npm install -g playwright (for HTML rendering in html2pptx)
- react-icons:
npm install -g react-icons react react-dom (for icons)
- sharp:
npm install -g sharp (for SVG rasterization and image processing)
- LibreOffice:
sudo apt-get install libreoffice (for PDF conversion)
- Poppler:
sudo apt-get install poppler-utils (for pdftoppm to convert PDF to images)
- defusedxml:
pip install defusedxml (for secure XML parsing)
1---2name: pptx3description: Presentation toolkit (.pptx). Create/edit slides, layouts, content, speaker notes, comments, for programmatic presentation creation and modification.4license: Proprietary. LICENSE.txt has complete terms5---6
7# PPTX creation, editing, and analysis
8
9## Overview
10
11A .pptx file is a ZIP archive containing XML files and resources. Create, edit, or analyze PowerPoint presentations using text extraction, raw XML access, or html2pptx workflows. Apply this skill for programmatic presentation creation and modification.
12
13## Visual Enhancement with Scientific Schematics
14
15**When creating documents with this skill, always consider adding scientific diagrams and schematics to enhance visual communication.**
16
17If your document does not already contain schematics or diagrams:
18- Use the **scientific-schematics** skill to generate AI-powered publication-quality diagrams
19- Simply describe your desired diagram in natural language
20- Nano Banana Pro will automatically generate, review, and refine the schematic
21
22**For new documents:** Scientific schematics should be generated by default to visually represent key concepts, workflows, architectures, or relationships described in the text.
23
24**How to generate schematics:**
25```bash
26python scripts/generate_schematic.py "your diagram description" -o figures/output.png
27```
28
29The AI will automatically:
30- Create publication-quality images with proper formatting
31- Review and refine through multiple iterations
32- Ensure accessibility (colorblind-friendly, high contrast)
33- Save outputs in the figures/ directory
34
35**When to add schematics:**
36- Presentation workflow diagrams for slides
37- Slide design process flowcharts
38- Content organization diagrams
39- System architecture illustrations
40- Process flow visualizations
41- Any complex concept that benefits from visualization
42
43For detailed guidance on creating schematics, refer to the scientific-schematics skill documentation.
44
45---
46
47## Reading and analyzing content
48
49### Text extraction
50To read the text contents of a presentation, convert the document to markdown:
51
52```bash
53# Convert document to markdown
54python -m markitdown path-to-file.pptx
55```
56
57### Raw XML access
58Raw XML access is required for: comments, speaker notes, slide layouts, animations, design elements, and complex formatting. For any of these features, unpack a presentation and read its raw XML contents.
59
60#### Unpacking a file
61`python ooxml/scripts/unpack.py <office_file> <output_dir>`
62
63**Note**: The unpack.py script is located at `skills/pptx/ooxml/scripts/unpack.py` relative to the project root. If the script doesn't exist at this path, use `find . -name "unpack.py"` to locate it.
64
65#### Key file structures
66* `ppt/presentation.xml` - Main presentation metadata and slide references
67* `ppt/slides/slide{N}.xml` - Individual slide contents (slide1.xml, slide2.xml, etc.)
68* `ppt/notesSlides/notesSlide{N}.xml` - Speaker notes for each slide
69* `ppt/comments/modernComment_*.xml` - Comments for specific slides
70* `ppt/slideLayouts/` - Layout templates for slides
71* `ppt/slideMasters/` - Master slide templates
72* `ppt/theme/` - Theme and styling information
73* `ppt/media/` - Images and other media files
74
75#### Typography and color extraction
76**When given an example design to emulate**: Always analyze the presentation's typography and colors first using the methods below:
771. **Read theme file**: Check `ppt/theme/theme1.xml` for colors (`<a:clrScheme>`) and fonts (`<a:fontScheme>`)
782. **Sample slide content**: Examine `ppt/slides/slide1.xml` for actual font usage (`<a:rPr>`) and colors
793. **Search for patterns**: Use grep to find color (`<a:solidFill>`, `<a:srgbClr>`) and font references across all XML files
80
81## Creating a new PowerPoint presentation **without a template**
82
83When creating a new PowerPoint presentation from scratch, use the **html2pptx** workflow to convert HTML slides to PowerPoint with accurate positioning.
84
85### Design Principles
86
87**CRITICAL**: Before creating any presentation, analyze the content and choose appropriate design elements:
881. **Consider the subject matter**: What is this presentation about? What tone, industry, or mood does it suggest?
892. **Check for branding**: If the user mentions a company/organization, consider their brand colors and identity
903. **Match palette to content**: Select colors that reflect the subject
914. **State your approach**: Explain your design choices before writing code
92
93**Requirements**:
94- ✅ State your content-informed design approach BEFORE writing code
95- ✅ Use web-safe fonts only: Arial, Helvetica, Times New Roman, Georgia, Courier New, Verdana, Tahoma, Trebuchet MS, Impact
96- ✅ Create clear visual hierarchy through size, weight, and color
97- ✅ Ensure readability: strong contrast, appropriately sized text, clean alignment
98- ✅ Be consistent: repeat patterns, spacing, and visual language across slides
99
100#### Color Palette Selection
101
102**Choosing colors creatively**:
103- **Think beyond defaults**: What colors genuinely match this specific topic? Avoid autopilot choices.
104- **Consider multiple angles**: Topic, industry, mood, energy level, target audience, brand identity (if mentioned)
105- **Be adventurous**: Try unexpected combinations - a healthcare presentation doesn't have to be green, finance doesn't have to be navy
106- **Build your palette**: Pick 3-5 colors that work together (dominant colors + supporting tones + accent)
107- **Ensure contrast**: Text must be clearly readable on backgrounds
108
109**Example color palettes** (use these to spark creativity - choose one, adapt it, or create your own):
110
1111. **Classic Blue**: Deep navy (#1C2833), slate gray (#2E4053), silver (#AAB7B8), off-white (#F4F6F6)
1122. **Teal & Coral**: Teal (#5EA8A7), deep teal (#277884), coral (#FE4447), white (#FFFFFF)
1133. **Bold Red**: Red (#C0392B), bright red (#E74C3C), orange (#F39C12), yellow (#F1C40F), green (#2ECC71)
1144. **Warm Blush**: Mauve (#A49393), blush (#EED6D3), rose (#E8B4B8), cream (#FAF7F2)
1155. **Burgundy Luxury**: Burgundy (#5D1D2E), crimson (#951233), rust (#C15937), gold (#997929)
1166. **Deep Purple & Emerald**: Purple (#B165FB), dark blue (#181B24), emerald (#40695B), white (#FFFFFF)
1177. **Cream & Forest Green**: Cream (#FFE1C7), forest green (#40695B), white (#FCFCFC)
1188. **Pink & Purple**: Pink (#F8275B), coral (#FF574A), rose (#FF737D), purple (#3D2F68)
1199. **Lime & Plum**: Lime (#C5DE82), plum (#7C3A5F), coral (#FD8C6E), blue-gray (#98ACB5)
12010. **Black & Gold**: Gold (#BF9A4A), black (#000000), cream (#F4F6F6)
12111. **Sage & Terracotta**: Sage (#87A96B), terracotta (#E07A5F), cream (#F4F1DE), charcoal (#2C2C2C)
12212. **Charcoal & Red**: Charcoal (#292929), red (#E33737), light gray (#CCCBCB)
12313. **Vibrant Orange**: Orange (#F96D00), light gray (#F2F2F2), charcoal (#222831)
12414. **Forest Green**: Black (#191A19), green (#4E9F3D), dark green (#1E5128), white (#FFFFFF)
12515. **Retro Rainbow**: Purple (#722880), pink (#D72D51), orange (#EB5C18), amber (#F08800), gold (#DEB600)
12616. **Vintage Earthy**: Mustard (#E3B448), sage (#CBD18F), forest green (#3A6B35), cream (#F4F1DE)
12717. **Coastal Rose**: Old rose (#AD7670), beaver (#B49886), eggshell (#F3ECDC), ash gray (#BFD5BE)
12818. **Orange & Turquoise**: Light orange (#FC993E), grayish turquoise (#667C6F), white (#FCFCFC)
129
130#### Visual Details Options
131
132**Geometric Patterns**:
133- Diagonal section dividers instead of horizontal
134- Asymmetric column widths (30/70, 40/60, 25/75)
135- Rotated text headers at 90° or 270°
136- Circular/hexagonal frames for images
137- Triangular accent shapes in corners
138- Overlapping shapes for depth
139
140**Border & Frame Treatments**:
141- Thick single-color borders (10-20pt) on one side only
142- Double-line borders with contrasting colors
143- Corner brackets instead of full frames
144- L-shaped borders (top+left or bottom+right)
145- Underline accents beneath headers (3-5pt thick)
146
147**Typography Treatments**:
148- Extreme size contrast (72pt headlines vs 11pt body)
149- All-caps headers with wide letter spacing
150- Numbered sections in oversized display type
151- Monospace (Courier New) for data/stats/technical content
152- Condensed fonts (Arial Narrow) for dense information
153- Outlined text for emphasis
154
155**Chart & Data Styling**:
156- Monochrome charts with single accent color for key data
157- Horizontal bar charts instead of vertical
158- Dot plots instead of bar charts
159- Minimal gridlines or none at all
160- Data labels directly on elements (no legends)
161- Oversized numbers for key metrics
162
163**Layout Innovations**:
164- Full-bleed images with text overlays
165- Sidebar column (20-30% width) for navigation/context
166- Modular grid systems (3×3, 4×4 blocks)
167- Z-pattern or F-pattern content flow
168- Floating text boxes over colored shapes
169- Magazine-style multi-column layouts
170
171**Background Treatments**:
172- Solid color blocks occupying 40-60% of slide
173- Gradient fills (vertical or diagonal only)
174- Split backgrounds (two colors, diagonal or vertical)
175- Edge-to-edge color bands
176- Negative space as a design element
177
178### Layout Tips
179**For slides with charts or tables:**
180- **Two-column layout (PREFERRED)**: Use a header spanning the full width, then two columns below - text/bullets in one column and the featured content in the other. This provides better balance and makes charts/tables more readable. Use flexbox with unequal column widths (e.g., 40%/60% split) to optimize space for each content type.
181- **Full-slide layout**: Let the featured content (chart/table) take up the entire slide for maximum impact and readability
182- **NEVER vertically stack**: Do not place charts/tables below text in a single column - this causes poor readability and layout issues
183
184### Workflow
1851. **MANDATORY - READ ENTIRE FILE**: Read [`html2pptx.md`](html2pptx.md) completely from start to finish. **NEVER set any range limits when reading this file.** Read the full file content for detailed syntax, critical formatting rules, and best practices before proceeding with presentation creation.
1862. Create an HTML file for each slide with proper dimensions (e.g., 720pt × 405pt for 16:9)
187 - Use `<p>`, `<h1>`-`<h6>`, `<ul>`, `<ol>` for all text content
188 - Use `class="placeholder"` for areas where charts/tables will be added (render with gray background for visibility)
189 - **CRITICAL**: Rasterize gradients and icons as PNG images FIRST using Sharp, then reference in HTML
190 - **LAYOUT**: For slides with charts/tables/images, use either full-slide layout or two-column layout for better readability
1913. Create and run a JavaScript file using the [`html2pptx.js`](scripts/html2pptx.js) library to convert HTML slides to PowerPoint and save the presentation
192 - Use the `html2pptx()` function to process each HTML file
193 - Add charts and tables to placeholder areas using PptxGenJS API
194 - Save the presentation using `pptx.writeFile()`
1954. **Visual validation**: Generate thumbnails and inspect for layout issues
196 - Create thumbnail grid: `python scripts/thumbnail.py output.pptx workspace/thumbnails --cols 4`
197 - Read and carefully examine the thumbnail image for:
198 - **Text cutoff**: Text being cut off by header bars, shapes, or slide edges
199 - **Text overlap**: Text overlapping with other text or shapes
200 - **Positioning issues**: Content too close to slide boundaries or other elements
201 - **Contrast issues**: Insufficient contrast between text and backgrounds
202 - If issues found, adjust HTML margins/spacing/colors and regenerate the presentation
203 - Repeat until all slides are visually correct
204
205## Editing an existing PowerPoint presentation
206
207To edit slides in an existing PowerPoint presentation, work with the raw Office Open XML (OOXML) format. This involves unpacking the .pptx file, editing the XML content, and repacking it.
208
209### Workflow
2101. **MANDATORY - READ ENTIRE FILE**: Read [`ooxml.md`](ooxml.md) (~500 lines) completely from start to finish. **NEVER set any range limits when reading this file.** Read the full file content for detailed guidance on OOXML structure and editing workflows before any presentation editing.
2112. Unpack the presentation: `python ooxml/scripts/unpack.py <office_file> <output_dir>`
2123. Edit the XML files (primarily `ppt/slides/slide{N}.xml` and related files)
2134. **CRITICAL**: Validate immediately after each edit and fix any validation errors before proceeding: `python ooxml/scripts/validate.py <dir> --original <file>`
2145. Pack the final presentation: `python ooxml/scripts/pack.py <input_directory> <office_file>`
215
216## Creating a new PowerPoint presentation **using a template**
217
218To create a presentation that follows an existing template's design, duplicate and re-arrange template slides before replacing placeholder context.
219
220### Workflow
2211. **Extract template text AND create visual thumbnail grid**:
222 * Extract text: `python -m markitdown template.pptx > template-content.md`
223 * Read `template-content.md`: Read the entire file to understand the contents of the template presentation. **NEVER set any range limits when reading this file.**
224 * Create thumbnail grids: `python scripts/thumbnail.py template.pptx`
225 * See [Creating Thumbnail Grids](#creating-thumbnail-grids) section for more details
226
2272. **Analyze template and save inventory to a file**:
228 * **Visual Analysis**: Review thumbnail grid(s) to understand slide layouts, design patterns, and visual structure
229 * Create and save a template inventory file at `template-inventory.md` containing:
230 ```markdown
231 # Template Inventory Analysis
232 **Total Slides: [count]**
233 **IMPORTANT: Slides are 0-indexed (first slide = 0, last slide = count-1)**
234
235 ## [Category Name]
236 - Slide 0: [Layout code if available] - Description/purpose
237 - Slide 1: [Layout code] - Description/purpose
238 - Slide 2: [Layout code] - Description/purpose
239 [... EVERY slide must be listed individually with its index ...]
240 ```
241 * **Using the thumbnail grid**: Reference the visual thumbnails to identify:
242 - Layout patterns (title slides, content layouts, section dividers)
243 - Image placeholder locations and counts
244 - Design consistency across slide groups
245 - Visual hierarchy and structure
246 * This inventory file is REQUIRED for selecting appropriate templates in the next step
247
2483. **Create presentation outline based on template inventory**:
249 * Review available templates from step 2.
250 * Choose an intro or title template for the first slide. This should be one of the first templates.
251 * Choose safe, text-based layouts for the other slides.
252 * **CRITICAL: Match layout structure to actual content**:
253 - Single-column layouts: Use for unified narrative or single topic
254 - Two-column layouts: Use ONLY when there are exactly 2 distinct items/concepts
255 - Three-column layouts: Use ONLY when there are exactly 3 distinct items/concepts
256 - Image + text layouts: Use ONLY when actual images are available to insert
257 - Quote layouts: Use ONLY for actual quotes from people (with attribution), never for emphasis
258 - Never use layouts with more placeholders than available content
259 - If there are 2 items, don't force them into a 3-column layout
260 - If there are 4+ items, consider breaking into multiple slides or using a list format
261 * Count actual content pieces BEFORE selecting the layout
262 * Verify each placeholder in the chosen layout will be filled with meaningful content
263 * Select one option representing the **best** layout for each content section.
264 * Save `outline.md` with content AND template mapping that leverages available designs
265 * Example template mapping:
266 ```
267 # Template slides to use (0-based indexing)
268 # WARNING: Verify indices are within range! Template with 73 slides has indices 0-72
269 # Mapping: slide numbers from outline -> template slide indices
270 template_mapping = [
271 0, # Use slide 0 (Title/Cover)
272 34, # Use slide 34 (B1: Title and body)
273 34, # Use slide 34 again (duplicate for second B1)
274 50, # Use slide 50 (E1: Quote)
275 54, # Use slide 54 (F2: Closing + Text)
276 ]
277 ```
278
2794. **Duplicate, reorder, and delete slides using `rearrange.py`**:
280 * Use the `scripts/rearrange.py` script to create a new presentation with slides in the desired order:
281 ```bash
282 python scripts/rearrange.py template.pptx working.pptx 0,34,34,50,52
283 ```
284 * The script handles duplicating repeated slides, deleting unused slides, and reordering automatically
285 * Slide indices are 0-based (first slide is 0, second is 1, etc.)
286 * The same slide index can appear multiple times to duplicate that slide
287
2885. **Extract ALL text using the `inventory.py` script**:
289 * **Run inventory extraction**:
290 ```bash
291 python scripts/inventory.py working.pptx text-inventory.json
292 ```
293 * **Read text-inventory.json**: Read the entire text-inventory.json file to understand all shapes and their properties. **NEVER set any range limits when reading this file.**
294
295 * The inventory JSON structure:
296 ```json
297 {
298 "slide-0": {
299 "shape-0": {
300 "placeholder_type": "TITLE", // or null for non-placeholders
301 "left": 1.5, // position in inches
302 "top": 2.0,
303 "width": 7.5,
304 "height": 1.2,
305 "paragraphs": [
306 {
307 "text": "Paragraph text",
308 // Optional properties (only included when non-default):
309 "bullet": true, // explicit bullet detected
310 "level": 0, // only included when bullet is true
311 "alignment": "CENTER", // CENTER, RIGHT (not LEFT)
312 "space_before": 10.0, // space before paragraph in points
313 "space_after": 6.0, // space after paragraph in points
314 "line_spacing": 22.4, // line spacing in points
315 "font_name": "Arial", // from first run
316 "font_size": 14.0, // in points
317 "bold": true,
318 "italic": false,
319 "underline": false,
320 "color": "FF0000" // RGB color
321 }
322 ]
323 }
324 }
325 }
326 ```
327
328 * Key features:
329 - **Slides**: Named as "slide-0", "slide-1", etc.
330 - **Shapes**: Ordered by visual position (top-to-bottom, left-to-right) as "shape-0", "shape-1", etc.
331 - **Placeholder types**: TITLE, CENTER_TITLE, SUBTITLE, BODY, OBJECT, or null
332 - **Default font size**: `default_font_size` in points extracted from layout placeholders (when available)
333 - **Slide numbers are filtered**: Shapes with SLIDE_NUMBER placeholder type are automatically excluded from inventory
334 - **Bullets**: When `bullet: true`, `level` is always included (even if 0)
335 - **Spacing**: `space_before`, `space_after`, and `line_spacing` in points (only included when set)
336 - **Colors**: `color` for RGB (e.g., "FF0000"), `theme_color` for theme colors (e.g., "DARK_1")
337 - **Properties**: Only non-default values are included in the output
338
3396. **Generate replacement text and save the data to a JSON file**
340 Based on the text inventory from the previous step:
341 - **CRITICAL**: First verify which shapes exist in the inventory - only reference shapes that are actually present
342 - **VALIDATION**: The replace.py script will validate that all shapes in the replacement JSON exist in the inventory
343 - If a non-existent shape is referenced, an error will show available shapes
344 - If a non-existent slide is referenced, an error will indicate the slide doesn't exist
345 - All validation errors are shown at once before the script exits
346 - **IMPORTANT**: The replace.py script uses inventory.py internally to identify ALL text shapes
347 - **AUTOMATIC CLEARING**: ALL text shapes from the inventory will be cleared unless you provide "paragraphs" for them
348 - Add a "paragraphs" field to shapes that need content (not "replacement_paragraphs")
349 - Shapes without "paragraphs" in the replacement JSON will have their text cleared automatically
350 - Paragraphs with bullets will be automatically left aligned. Don't set the `alignment` property on when `"bullet": true`
351 - Generate appropriate replacement content for placeholder text
352 - Use shape size to determine appropriate content length
353 - **CRITICAL**: Include paragraph properties from the original inventory - don't just provide text
354 - **IMPORTANT**: When bullet: true, do NOT include bullet symbols (•, -, *) in text - they are added automatically
355 - **ESSENTIAL FORMATTING RULES**:
356 - Headers/titles should typically have `"bold": true`
357 - List items should have `"bullet": true, "level": 0` (level is required when bullet is true)
358 - Preserve any alignment properties (e.g., `"alignment": "CENTER"` for centered text)
359 - Include font properties when different from default (e.g., `"font_size": 14.0`, `"font_name": "Lora"`)
360 - Colors: Use `"color": "FF0000"` for RGB or `"theme_color": "DARK_1"` for theme colors
361 - The replacement script expects **properly formatted paragraphs**, not just text strings
362 - **Overlapping shapes**: Prefer shapes with larger default_font_size or more appropriate placeholder_type
363 - Save the updated inventory with replacements to `replacement-text.json`
364 - **WARNING**: Different template layouts have different shape counts - always check the actual inventory before creating replacements
365
366 Example paragraphs field showing proper formatting:
367 ```json
368 "paragraphs": [
369 {
370 "text": "New presentation title text",
371 "alignment": "CENTER",
372 "bold": true
373 },
374 {
375 "text": "Section Header",
376 "bold": true
377 },
378 {
379 "text": "First bullet point without bullet symbol",
380 "bullet": true,
381 "level": 0
382 },
383 {
384 "text": "Red colored text",
385 "color": "FF0000"
386 },
387 {
388 "text": "Theme colored text",
389 "theme_color": "DARK_1"
390 },
391 {
392 "text": "Regular paragraph text without special formatting"
393 }
394 ]
395 ```
396
397 **Shapes not listed in the replacement JSON are automatically cleared**:
398 ```json
399 {
400 "slide-0": {
401 "shape-0": {
402 "paragraphs": [...] // This shape gets new text
403 }
404 // shape-1 and shape-2 from inventory will be cleared automatically
405 }
406 }
407 ```
408
409 **Common formatting patterns for presentations**:
410 - Title slides: Bold text, sometimes centered
411 - Section headers within slides: Bold text
412 - Bullet lists: Each item needs `"bullet": true, "level": 0`
413 - Body text: Usually no special properties needed
414 - Quotes: May have special alignment or font properties
415
4167. **Apply replacements using the `replace.py` script**
417 ```bash
418 python scripts/replace.py working.pptx replacement-text.json output.pptx
419 ```
420
421 The script will:
422 - First extract the inventory of ALL text shapes using functions from inventory.py
423 - Validate that all shapes in the replacement JSON exist in the inventory
424 - Clear text from ALL shapes identified in the inventory
425 - Apply new text only to shapes with "paragraphs" defined in the replacement JSON
426 - Preserve formatting by applying paragraph properties from the JSON
427 - Handle bullets, alignment, font properties, and colors automatically
428 - Save the updated presentation
429
430 Example validation errors:
431 ```
432 ERROR: Invalid shapes in replacement JSON:
433 - Shape 'shape-99' not found on 'slide-0'. Available shapes: shape-0, shape-1, shape-4
434 - Slide 'slide-999' not found in inventory
435 ```
436
437 ```
438 ERROR: Replacement text made overflow worse in these shapes:
439 - slide-0/shape-2: overflow worsened by 1.25" (was 0.00", now 1.25")
440 ```
441
442## Creating Thumbnail Grids
443
444To create visual thumbnail grids of PowerPoint slides for quick analysis and reference:
445
446```bash
447python scripts/thumbnail.py template.pptx [output_prefix]
448```
449
450**Features**:
451- Creates: `thumbnails.jpg` (or `thumbnails-1.jpg`, `thumbnails-2.jpg`, etc. for large decks)
452- Default: 5 columns, max 30 slides per grid (5×6)
453- Custom prefix: `python scripts/thumbnail.py template.pptx my-grid`
454 - Note: The output prefix should include the path if you want output in a specific directory (e.g., `workspace/my-grid`)
455- Adjust columns: `--cols 4` (range: 3-6, affects slides per grid)
456- Grid limits: 3 cols = 12 slides/grid, 4 cols = 20, 5 cols = 30, 6 cols = 42
457- Slides are zero-indexed (Slide 0, Slide 1, etc.)
458
459**Use cases**:
460- Template analysis: Quickly understand slide layouts and design patterns
461- Content review: Visual overview of entire presentation
462- Navigation reference: Find specific slides by their visual appearance
463- Quality check: Verify all slides are properly formatted
464
465**Examples**:
466```bash
467# Basic usage
468python scripts/thumbnail.py presentation.pptx
469
470# Combine options: custom name, columns
471python scripts/thumbnail.py template.pptx analysis --cols 4
472```
473
474## Converting Slides to Images
475
476To visually analyze PowerPoint slides, convert them to images using a two-step process:
477
4781. **Convert PPTX to PDF**:
479 ```bash
480 soffice --headless --convert-to pdf template.pptx
481 ```
482
4832. **Convert PDF pages to JPEG images**:
484 ```bash
485 pdftoppm -jpeg -r 150 template.pdf slide
486 ```
487 This creates files like `slide-1.jpg`, `slide-2.jpg`, etc.
488
489Options:
490- `-r 150`: Sets resolution to 150 DPI (adjust for quality/size balance)
491- `-jpeg`: Output JPEG format (use `-png` for PNG if preferred)
492- `-f N`: First page to convert (e.g., `-f 2` starts from page 2)
493- `-l N`: Last page to convert (e.g., `-l 5` stops at page 5)
494- `slide`: Prefix for output files
495
496Example for specific range:
497```bash
498pdftoppm -jpeg -r 150 -f 2 -l 5 template.pdf slide # Converts only pages 2-5
499```
500
501## Code Style Guidelines
502**IMPORTANT**: When generating code for PPTX operations:
503- Write concise code
504- Avoid verbose variable names and redundant operations
505- Avoid unnecessary print statements
506
507## Dependencies
508
509Required dependencies (should already be installed):
510
511- **markitdown**: `pip install "markitdown[pptx]"` (for text extraction from presentations)
512- **pptxgenjs**: `npm install -g pptxgenjs` (for creating presentations via html2pptx)
513- **playwright**: `npm install -g playwright` (for HTML rendering in html2pptx)
514- **react-icons**: `npm install -g react-icons react react-dom` (for icons)
515- **sharp**: `npm install -g sharp` (for SVG rasterization and image processing)
516- **LibreOffice**: `sudo apt-get install libreoffice` (for PDF conversion)
517- **Poppler**: `sudo apt-get install poppler-utils` (for pdftoppm to convert PDF to images)
518- **defusedxml**: `pip install defusedxml` (for secure XML parsing)