PPTX creation, editing, and analysis
Overview
A user may ask you to create, edit, or analyze the contents of a .pptx file. A .pptx file is essentially a ZIP archive containing XML files and other resources that you can read or edit. You have different tools and workflows available for different tasks.
Reading and analyzing content
Text extraction
If you just need to read the text contents of a presentation, you should convert the document to markdown:
# Convert document to markdown
python -m markitdown path-to-file.pptx
Raw XML access
You need raw XML access for: comments, speaker notes, slide layouts, animations, design elements, and complex formatting. For any of these features, you'll need to unpack a presentation and read its raw XML contents.
Unpacking a file
python ooxml/scripts/unpack.py <office_file> <output_dir>
Note: The unpack.py script is located at skills/pptx/ooxml/scripts/unpack.py relative to the project root. If the script doesn't exist at this path, use find . -name "unpack.py" to locate it.
Key file structures
ppt/presentation.xml - Main presentation metadata and slide references
ppt/slides/slide{N}.xml - Individual slide contents (slide1.xml, slide2.xml, etc.)
ppt/notesSlides/notesSlide{N}.xml - Speaker notes for each slide
ppt/comments/modernComment_*.xml - Comments for specific slides
ppt/slideLayouts/ - Layout templates for slides
ppt/slideMasters/ - Master slide templates
ppt/theme/ - Theme and styling information
ppt/media/ - Images and other media files
Typography and color extraction
When given an example design to emulate: Always analyze the presentation's typography and colors first using the methods below:
- Read theme file: Check
ppt/theme/theme1.xml for colors (<a:clrScheme>) and fonts (<a:fontScheme>)
- Sample slide content: Examine
ppt/slides/slide1.xml for actual font usage (<a:rPr>) and colors
- Search for patterns: Use grep to find color (
<a:solidFill>, <a:srgbClr>) and font references across all XML files
Creating a new PowerPoint presentation without a template
When creating a new PowerPoint presentation from scratch, use the html2pptx workflow to convert HTML slides to PowerPoint with accurate positioning.
Design Principles
CRITICAL: Before creating any presentation, analyze the content and choose appropriate design elements:
- Consider the subject matter: What is this presentation about? What tone, industry, or mood does it suggest?
- Check for branding: If the user mentions a company/organization, consider their brand colors and identity
- Match palette to content: Select colors that reflect the subject
- State your approach: Explain your design choices before writing code
Requirements:
- ✅ State your content-informed design approach BEFORE writing code
- ✅ Use web-safe fonts only: Arial, Helvetica, Times New Roman, Georgia, Courier New, Verdana, Tahoma, Trebuchet MS, Impact
- ✅ Create clear visual hierarchy through size, weight, and color
- ✅ Ensure readability: strong contrast, appropriately sized text, clean alignment
- ✅ Be consistent: repeat patterns, spacing, and visual language across slides
Color Palette Selection
Choosing colors creatively:
- Think beyond defaults: What colors genuinely match this specific topic? Avoid autopilot choices.
- Consider multiple angles: Topic, industry, mood, energy level, target audience, brand identity (if mentioned)
- Be adventurous: Try unexpected combinations - a healthcare presentation doesn't have to be green, finance doesn't have to be navy
- Build your palette: Pick 3-5 colors that work together (dominant colors + supporting tones + accent)
- Ensure contrast: Text must be clearly readable on backgrounds
Example color palettes (use these to spark creativity - choose one, adapt it, or create your own):
- Classic Blue: Deep navy (#1C2833), slate gray (#2E4053), silver (#AAB7B8), off-white (#F4F6F6)
- Teal & Coral: Teal (#5EA8A7), deep teal (#277884), coral (#FE4447), white (#FFFFFF)
- Bold Red: Red (#C0392B), bright red (#E74C3C), orange (#F39C12), yellow (#F1C40F), green (#2ECC71)
- Warm Blush: Mauve (#A49393), blush (#EED6D3), rose (#E8B4B8), cream (#FAF7F2)
- Burgundy Luxury: Burgundy (#5D1D2E), crimson (#951233), rust (#C15937), gold (#997929)
- Deep Purple & Emerald: Purple (#B165FB), dark blue (#181B24), emerald (#40695B), white (#FFFFFF)
- Cream & Forest Green: Cream (#FFE1C7), forest green (#40695B), white (#FCFCFC)
- Pink & Purple: Pink (#F8275B), coral (#FF574A), rose (#FF737D), purple (#3D2F68)
- Lime & Plum: Lime (#C5DE82), plum (#7C3A5F), coral (#FD8C6E), blue-gray (#98ACB5)
- Black & Gold: Gold (#BF9A4A), black (#000000), cream (#F4F6F6)
- Sage & Terracotta: Sage (#87A96B), terracotta (#E07A5F), cream (#F4F1DE), charcoal (#2C2C2C)
- Charcoal & Red: Charcoal (#292929), red (#E33737), light gray (#CCCBCB)
- Vibrant Orange: Orange (#F96D00), light gray (#F2F2F2), charcoal (#222831)
- Forest Green: Black (#191A19), green (#4E9F3D), dark green (#1E5128), white (#FFFFFF)
- Retro Rainbow: Purple (#722880), pink (#D72D51), orange (#EB5C18), amber (#F08800), gold (#DEB600)
- Vintage Earthy: Mustard (#E3B448), sage (#CBD18F), forest green (#3A6B35), cream (#F4F1DE)
- Coastal Rose: Old rose (#AD7670), beaver (#B49886), eggshell (#F3ECDC), ash gray (#BFD5BE)
- Orange & Turquoise: Light orange (#FC993E), grayish turquoise (#667C6F), white (#FCFCFC)
Visual Details Options
Geometric Patterns:
- Diagonal section dividers instead of horizontal
- Asymmetric column widths (30/70, 40/60, 25/75)
- Rotated text headers at 90° or 270°
- Circular/hexagonal frames for images
- Triangular accent shapes in corners
- Overlapping shapes for depth
Border & Frame Treatments:
- Thick single-color borders (10-20pt) on one side only
- Double-line borders with contrasting colors
- Corner brackets instead of full frames
- L-shaped borders (top+left or bottom+right)
- Underline accents beneath headers (3-5pt thick)
Typography Treatments:
- Extreme size contrast (72pt headlines vs 11pt body)
- All-caps headers with wide letter spacing
- Numbered sections in oversized display type
- Monospace (Courier New) for data/stats/technical content
- Condensed fonts (Arial Narrow) for dense information
- Outlined text for emphasis
Chart & Data Styling:
- Monochrome charts with single accent color for key data
- Horizontal bar charts instead of vertical
- Dot plots instead of bar charts
- Minimal gridlines or none at all
- Data labels directly on elements (no legends)
- Oversized numbers for key metrics
Layout Innovations:
- Full-bleed images with text overlays
- Sidebar column (20-30% width) for navigation/context
- Modular grid systems (3×3, 4×4 blocks)
- Z-pattern or F-pattern content flow
- Floating text boxes over colored shapes
- Magazine-style multi-column layouts
Background Treatments:
- Solid color blocks occupying 40-60% of slide
- Gradient fills (vertical or diagonal only)
- Split backgrounds (two colors, diagonal or vertical)
- Edge-to-edge color bands
- Negative space as a design element
Layout Tips
When creating slides with charts or tables:
- Two-column layout (PREFERRED): Use a header spanning the full width, then two columns below - text/bullets in one column and the featured content in the other. This provides better balance and makes charts/tables more readable. Use flexbox with unequal column widths (e.g., 40%/60% split) to optimize space for each content type.
- Full-slide layout: Let the featured content (chart/table) take up the entire slide for maximum impact and readability
- NEVER vertically stack: Do not place charts/tables below text in a single column - this causes poor readability and layout issues
Workflow: Step-by-Step Deterministic Process
⚠️ CRITICAL: Follow these steps IN ORDER. Each step creates ONE file before moving to the next.
Step 1: Read Documentation (MANDATORY)
Read html2pptx.md completely from start to finish. NEVER set any range limits. This file contains:
- HTML syntax rules and supported elements
- Dimension requirements for different layouts
- Placeholder usage for charts/images
- PptxGenJS chart/table examples
Step 2: Create HTML Files (One HTML file per slide)
IMPORTANT: Create separate HTML files - one file for each slide.
Example for a 3-slide presentation:
<!-- slide1.html -->
<!DOCTYPE html>
<html>
<head>
<style>
body {
width: 720pt;
height: 405pt; /* 16:9 layout */
margin: 0;
padding: 48pt;
box-sizing: border-box;
font-family: Arial, sans-serif;
background: #1C2833;
}
h1 { font-size: 36pt; color: #F4F6F6; margin: 0 0 24pt 0; }
p { font-size: 18pt; color: #F4F6F6; }
</style>
</head>
<body>
<h1>Slide 1 Title</h1>
<p>Content for first slide</p>
</body>
</html>
<!-- slide2.html -->
<!DOCTYPE html>
<html>
<head>
<style>
body {
width: 720pt;
height: 405pt;
margin: 0;
padding: 48pt;
box-sizing: border-box;
font-family: Arial, sans-serif;
background: #2E4053;
}
h1 { font-size: 36pt; color: #F4F6F6; margin: 0 0 24pt 0; }
p { font-size: 18pt; color: #F4F6F6; }
</style>
</head>
<body>
<h1>Slide 2 Title</h1>
<p>Content for second slide</p>
<!-- Placeholder for chart -->
<div id="chart1" class="placeholder" style="width: 600pt; height: 260pt;"></div>
</body>
</html>
<!-- slide3.html - similar structure -->
Critical HTML Rules:
- ✅ ONE HTML file = ONE slide
- ✅ ALL text MUST be in
<p>, <h1>-<h6>, <ul>, <ol> tags
- ✅ Use
class="placeholder" with unique id for chart/image areas
- ✅ Use web-safe fonts only (Arial, Helvetica, Times New Roman, etc.)
- ✅ Rasterize gradients/icons as PNG FIRST using Sharp, then reference
- ✅ Two-column or full-slide layout for charts/tables (NEVER vertical stack)
Step 3: Create JavaScript Presentation Script (ONE file)
CRITICAL - Script Path Resolution
When generating a script that uses html2pptx.js, you MUST use the correct path pattern. The skills connector sets the SKILL_DIR environment variable pointing to this skill's root directory.
Required pattern for path resolution:
const path = require('path');
// CRITICAL: Use SKILL_DIR environment variable to locate html2pptx
// SKILL_DIR points to the skill's root (e.g., D:\ops-data\agent-skills\pptx)
// This works regardless of where the script is run from
const skillDir = process.env.SKILL_DIR;
if (!skillDir) {
console.error('ERROR: SKILL_DIR environment variable not set');
console.error('This script must be run via the skills connector');
process.exit(1);
}
const html2pptx = require(path.join(skillDir, 'scripts', 'html2pptx'));
Why this pattern is required:
- ❌
require('./scripts/html2pptx') - Fails if script runs from a different directory
- ❌
require('../agent-skills/pptx/scripts/html2pptx') - Breaks if connector mount path changes
- ✅
require(path.join(process.env.SKILL_DIR, 'scripts', 'html2pptx')) - Always works correctly
Full example script:
// create-presentation.js
const pptxgen = require('pptxgenjs');
const path = require('path');
// CRITICAL: Use SKILL_DIR environment variable to locate html2pptx
const skillDir = process.env.SKILL_DIR;
if (!skillDir) {
console.error('ERROR: SKILL_DIR environment variable not set');
console.error('This script must be run via the skills connector');
process.exit(1);
}
const html2pptx = require(path.join(skillDir, 'scripts', 'html2pptx'));
async function main() {
try {
console.log('Creating presentation...');
console.log('SKILL_DIR:', skillDir);
const pptx = new pptxgen();
pptx.layout = 'LAYOUT_16x9';
pptx.title = 'My Presentation';
// Convert each HTML file to a slide
// ONE html2pptx() call per HTML file
await html2pptx(path.join(__dirname, 'slide1.html'), pptx);
await html2pptx(path.join(__dirname, 'slide2.html'), pptx);
await html2pptx(path.join(__dirname, 'slide3.html'), pptx);
// Add charts to placeholders if needed
// (See complete example in html2pptx.md)
await pptx.writeFile({ fileName: 'output.pptx' });
console.log('✓ Presentation created: output.pptx');
process.exit(0);
} catch (error) {
console.error('✗ Error:', error.message);
console.error('Stack trace:', error.stack);
process.exit(1);
}
}
main();
Step 4: Run the JavaScript File
Use run_script with absolute path from write response
After writing the script in Step 3, the write response includes the full absolute path. Use this exact path with run_script and add the skillId parameter.
✅ Correct approach:
// Step 3 write response: { "path": "D:\\ops-data\\create-presentation.js" }
// Step 4: Run using the FULL PATH from response
{
"action": "remote_tool_call",
"target": "filesystem",
"params": {
"agent_id": "local-dev",
"connector_id": "local-files", // Same connector you used for write
"command": "run_script",
"path": "D:\\ops-data\\create-presentation.js", // FULL path from write response
"skillId": "pptx", // Sets SKILL_DIR environment variable
"type": "node"
}
}
How skillId works:
- Sets environment variable:
SKILL_DIR=/path/to/agent-skills/pptx
- Your script uses
process.env.SKILL_DIR to locate html2pptx.js
- Works regardless of which connector the script is stored in
Check the result:
- ✅
exitCode: 0 → Success! PPTX created
- ❌
exitCode: 1 → Read stdout/stderr error message, fix issues, retry
- ❌ "Script not found" → Check connector_id matches where you wrote the file
Step 5: Visual Validation
Generate thumbnails and inspect:
python scripts/thumbnail.py output.pptx workspace/thumbnails --cols 4
Read thumbnail image and check for:
- Text cutoff or overlap
- Positioning issues
- Contrast problems
If issues found, adjust HTML and regenerate (return to Step 2).
📋 Summary - Files to Create:
- HTML files:
slide1.html, slide2.html, slide3.html, ... (ONE file per slide)
- JavaScript file:
create-presentation.js (ONE script that calls html2pptx() for each HTML file)
- Run:
node create-presentation.js → output.pptx created
- Validate: Generate thumbnails to check layout
Troubleshooting Script Errors
If your script fails with exit code 1:
The error output from run_script should contain diagnostic information. Look for:
Module not found errors:
Error: Cannot find module './scripts/html2pptx'
Cause: Not using process.env.SKILL_DIR for path resolution
Fix: Update require statement to use path.join(process.env.SKILL_DIR, 'scripts', 'html2pptx')
SKILL_DIR not set:
ERROR: SKILL_DIR environment variable not set
Cause: Missing skillId parameter in run_script call
Fix: Add "skillId": "pptx" to the params when calling run_script. This sets the SKILL_DIR environment variable that your script needs to locate html2pptx.
HTML conversion errors:
Error: HTML file must have exactly one body element
Cause: Invalid HTML structure
Fix: Verify HTML files have proper structure with single <html>, <head>, and <body> tags
Dimension errors:
Error: Body dimensions must be 720pt × 405pt for 16:9
Cause: Incorrect HTML body dimensions
Fix: Ensure body { width: 720pt; height: 405pt; } in CSS
Debug template to add to your script:
// Add this at the start of your script for debugging
console.log('=== Debug Info ===');
console.log('Current directory:', __dirname);
console.log('SKILL_DIR:', process.env.SKILL_DIR);
console.log('html2pptx path:', path.join(process.env.SKILL_DIR, 'scripts', 'html2pptx'));
console.log('==================');
Common mistakes to avoid:
- ❌ Using relative paths:
require('./scripts/html2pptx')
- ❌ Using hard-coded paths:
require('D:\\ops-data\\agent-skills\\pptx\\scripts\\html2pptx')
- ❌ Missing error handler: Not catching errors means no diagnostic output
- ❌ Not logging SKILL_DIR: Makes debugging path issues harder
- ✅ Always use:
require(path.join(process.env.SKILL_DIR, 'scripts', 'html2pptx'))
- ✅ Always log: Console output helps diagnose failures
- ✅ Always catch: Wrap main logic in try-catch with detailed error logging
Reference example:
See example-create-presentation.js for a complete working example with:
- Correct SKILL_DIR usage
- Comprehensive error handling
- Detailed diagnostic output
- Proper exit codes
Editing an existing PowerPoint presentation
When edit slides in an existing PowerPoint presentation, you need to work with the raw Office Open XML (OOXML) format. This involves unpacking the .pptx file, editing the XML content, and repacking it.
Workflow
- MANDATORY - READ ENTIRE FILE: Read
ooxml.md (~500 lines) completely from start to finish. NEVER set any range limits when reading this file. Read the full file content for detailed guidance on OOXML structure and editing workflows before any presentation editing.
- Unpack the presentation:
python ooxml/scripts/unpack.py <office_file> <output_dir>
- Edit the XML files (primarily
ppt/slides/slide{N}.xml and related files)
- CRITICAL: Validate immediately after each edit and fix any validation errors before proceeding:
python ooxml/scripts/validate.py <dir> --original <file>
- Pack the final presentation:
python ooxml/scripts/pack.py <input_directory> <office_file>
Creating a new PowerPoint presentation using a template
When you need to create a presentation that follows an existing template's design, you'll need to duplicate and re-arrange template slides before then replacing placeholder context.
Workflow
Extract template text AND create visual thumbnail grid:
- Extract text:
python -m markitdown template.pptx > template-content.md
- Read
template-content.md: Read the entire file to understand the contents of the template presentation. NEVER set any range limits when reading this file.
- Create thumbnail grids:
python scripts/thumbnail.py template.pptx
- See Creating Thumbnail Grids section for more details
Analyze template and save inventory to a file:
- Visual Analysis: Review thumbnail grid(s) to understand slide layouts, design patterns, and visual structure
- Create and save a template inventory file at
template-inventory.md containing:# Template Inventory Analysis
**Total Slides: [count]**
**IMPORTANT: Slides are 0-indexed (first slide = 0, last slide = count-1)**
## [Category Name]
- Slide 0: [Layout code if available] - Description/purpose
- Slide 1: [Layout code] - Description/purpose
- Slide 2: [Layout code] - Description/purpose
[... EVERY slide must be listed individually with its index ...]
- Using the thumbnail grid: Reference the visual thumbnails to identify:
- Layout patterns (title slides, content layouts, section dividers)
- Image placeholder locations and counts
- Design consistency across slide groups
- Visual hierarchy and structure
- This inventory file is REQUIRED for selecting appropriate templates in the next step
Create presentation outline based on template inventory:
- Review available templates from step 2.
- Choose an intro or title template for the first slide. This should be one of the first templates.
- Choose safe, text-based layouts for the other slides.
- CRITICAL: Match layout structure to actual content:
- Single-column layouts: Use for unified narrative or single topic
- Two-column layouts: Use ONLY when you have exactly 2 distinct items/concepts
- Three-column layouts: Use ONLY when you have exactly 3 distinct items/concepts
- Image + text layouts: Use ONLY when you have actual images to insert
- Quote layouts: Use ONLY for actual quotes from people (with attribution), never for emphasis
- Never use layouts with more placeholders than you have content
- If you have 2 items, don't force them into a 3-column layout
- If you have 4+ items, consider breaking into multiple slides or using a list format
- Count your actual content pieces BEFORE selecting the layout
- Verify each placeholder in the chosen layout will be filled with meaningful content
- Select one option representing the best layout for each content section.
- Save
outline.md with content AND template mapping that leverages available designs
- Example template mapping:
# Template slides to use (0-based indexing)
# WARNING: Verify indices are within range! Template with 73 slides has indices 0-72
# Mapping: slide numbers from outline -> template slide indices
template_mapping = [
0, # Use slide 0 (Title/Cover)
34, # Use slide 34 (B1: Title and body)
34, # Use slide 34 again (duplicate for second B1)
50, # Use slide 50 (E1: Quote)
54, # Use slide 54 (F2: Closing + Text)
]
Duplicate, reorder, and delete slides using rearrange.py:
- Use the
scripts/rearrange.py script to create a new presentation with slides in the desired order:python scripts/rearrange.py template.pptx working.pptx 0,34,34,50,52
- The script handles duplicating repeated slides, deleting unused slides, and reordering automatically
- Slide indices are 0-based (first slide is 0, second is 1, etc.)
- The same slide index can appear multiple times to duplicate that slide
Extract ALL text using the inventory.py script:
Run inventory extraction:
python scripts/inventory.py working.pptx text-inventory.json
Read text-inventory.json: Read the entire text-inventory.json file to understand all shapes and their properties. NEVER set any range limits when reading this file.
The inventory JSON structure:
{
"slide-0": {
"shape-0": {
"placeholder_type": "TITLE", // or null for non-placeholders
"left": 1.5, // position in inches
"top": 2.0,
"width": 7.5,
"height": 1.2,
"paragraphs": [
{
"text": "Paragraph text",
// Optional properties (only included when non-default):
"bullet": true, // explicit bullet detected
"level": 0, // only included when bullet is true
"alignment": "CENTER", // CENTER, RIGHT (not LEFT)
"space_before": 10.0, // space before paragraph in points
"space_after": 6.0, // space after paragraph in points
"line_spacing": 22.4, // line spacing in points
"font_name": "Arial", // from first run
"font_size": 14.0, // in points
"bold": true,
"italic": false,
"underline": false,
"color": "FF0000" // RGB color
}
]
}
}
}
Key features:
- Slides: Named as "slide-0", "slide-1", etc.
- Shapes: Ordered by visual position (top-to-bottom, left-to-right) as "shape-0", "shape-1", etc.
- Placeholder types: TITLE, CENTER_TITLE, SUBTITLE, BODY, OBJECT, or null
- Default font size:
default_font_size in points extracted from layout placeholders (when available)
- Slide numbers are filtered: Shapes with SLIDE_NUMBER placeholder type are automatically excluded from inventory
- Bullets: When
bullet: true, level is always included (even if 0)
- Spacing:
space_before, space_after, and line_spacing in points (only included when set)
- Colors:
color for RGB (e.g., "FF0000"), theme_color for theme colors (e.g., "DARK_1")
- Properties: Only non-default values are included in the output
Generate replacement text and save the data to a JSON file
Based on the text inventory from the previous step:
- CRITICAL: First verify which shapes exist in the inventory - only reference shapes that are actually present
- VALIDATION: The replace.py script will validate that all shapes in your replacement JSON exist in the inventory
- If you reference a non-existent shape, you'll get an error showing available shapes
- If you reference a non-existent slide, you'll get an error indicating the slide doesn't exist
- All validation errors are shown at once before the script exits
- IMPORTANT: The replace.py script uses inventory.py internally to identify ALL text shapes
- AUTOMATIC CLEARING: ALL text shapes from the inventory will be cleared unless you provide "paragraphs" for them
- Add a "paragraphs" field to shapes that need content (not "replacement_paragraphs")
- Shapes without "paragraphs" in the replacement JSON will have their text cleared automatically
- Paragraphs with bullets will be automatically left aligned. Don't set the
alignment property on when "bullet": true
- Generate appropriate replacement content for placeholder text
- Use shape size to determine appropriate content length
- CRITICAL: Include paragraph properties from the original inventory - don't just provide text
- IMPORTANT: When bullet: true, do NOT include bullet symbols (•, -, *) in text - they're added automatically
- ESSENTIAL FORMATTING RULES:
- Headers/titles should typically have
"bold": true
- List items should have
"bullet": true, "level": 0 (level is required when bullet is true)
- Preserve any alignment properties (e.g.,
"alignment": "CENTER" for centered text)
- Include font properties when different from default (e.g.,
"font_size": 14.0, "font_name": "Lora")
- Colors: Use
"color": "FF0000" for RGB or "theme_color": "DARK_1" for theme colors
- The replacement script expects properly formatted paragraphs, not just text strings
- Overlapping shapes: Prefer shapes with larger default_font_size or more appropriate placeholder_type
- Save the updated inventory with replacements to
replacement-text.json
- WARNING: Different template layouts have different shape counts - always check the actual inventory before creating replacements
Example paragraphs field showing proper formatting:
"paragraphs": [
{
"text": "New presentation title text",
"alignment": "CENTER",
"bold": true
},
{
"text": "Section Header",
"bold": true
},
{
"text": "First bullet point without bullet symbol",
"bullet": true,
"level": 0
},
{
"text": "Red colored text",
"color": "FF0000"
},
{
"text": "Theme colored text",
"theme_color": "DARK_1"
},
{
"text": "Regular paragraph text without special formatting"
}
]
Shapes not listed in the replacement JSON are automatically cleared:
{
"slide-0": {
"shape-0": {
"paragraphs": [...] // This shape gets new text
}
// shape-1 and shape-2 from inventory will be cleared automatically
}
}
Common formatting patterns for presentations:
- Title slides: Bold text, sometimes centered
- Section headers within slides: Bold text
- Bullet lists: Each item needs
"bullet": true, "level": 0
- Body text: Usually no special properties needed
- Quotes: May have special alignment or font properties
Apply replacements using the replace.py script
python scripts/replace.py working.pptx replacement-text.json output.pptx
The script will:
- First extract the inventory of ALL text shapes using functions from inventory.py
- Validate that all shapes in the replacement JSON exist in the inventory
- Clear text from ALL shapes identified in the inventory
- Apply new text only to shapes with "paragraphs" defined in the replacement JSON
- Preserve formatting by applying paragraph properties from the JSON
- Handle bullets, alignment, font properties, and colors automatically
- Save the updated presentation
Example validation errors:
ERROR: Invalid shapes in replacement JSON:
- Shape 'shape-99' not found on 'slide-0'. Available shapes: shape-0, shape-1, shape-4
- Slide 'slide-999' not found in inventory
ERROR: Replacement text made overflow worse in these shapes:
- slide-0/shape-2: overflow worsened by 1.25" (was 0.00", now 1.25")
Creating Thumbnail Grids
To create visual thumbnail grids of PowerPoint slides for quick analysis and reference:
python scripts/thumbnail.py template.pptx [output_prefix]
Features:
- Creates:
thumbnails.jpg (or thumbnails-1.jpg, thumbnails-2.jpg, etc. for large decks)
- Default: 5 columns, max 30 slides per grid (5×6)
- Custom prefix:
python scripts/thumbnail.py template.pptx my-grid
- Note: The output prefix should include the path if you want output in a specific directory (e.g.,
workspace/my-grid)
- Adjust columns:
--cols 4 (range: 3-6, affects slides per grid)
- Grid limits: 3 cols = 12 slides/grid, 4 cols = 20, 5 cols = 30, 6 cols = 42
- Slides are zero-indexed (Slide 0, Slide 1, etc.)
Use cases:
- Template analysis: Quickly understand slide layouts and design patterns
- Content review: Visual overview of entire presentation
- Navigation reference: Find specific slides by their visual appearance
- Quality check: Verify all slides are properly formatted
Examples:
# Basic usage
python scripts/thumbnail.py presentation.pptx
# Combine options: custom name, columns
python scripts/thumbnail.py template.pptx analysis --cols 4
Converting Slides to Images
To visually analyze PowerPoint slides, convert them to images using a two-step process:
Convert PPTX to PDF:
soffice --headless --convert-to pdf template.pptx
Convert PDF pages to JPEG images:
pdftoppm -jpeg -r 150 template.pdf slide
This creates files like slide-1.jpg, slide-2.jpg, etc.
Options:
-r 150: Sets resolution to 150 DPI (adjust for quality/size balance)
-jpeg: Output JPEG format (use -png for PNG if preferred)
-f N: First page to convert (e.g., -f 2 starts from page 2)
-l N: Last page to convert (e.g., -l 5 stops at page 5)
slide: Prefix for output files
Example for specific range:
pdftoppm -jpeg -r 150 -f 2 -l 5 template.pdf slide # Converts only pages 2-5
Code Style Guidelines
IMPORTANT: When generating code for PPTX operations:
- Write concise code
- Avoid verbose variable names and redundant operations
- Avoid unnecessary print statements
Dependencies
Required dependencies (should already be installed):
- markitdown:
pip install "markitdown[pptx]" (for text extraction from presentations)
- pptxgenjs:
npm install -g pptxgenjs (for creating presentations via html2pptx)
- playwright:
npm install -g playwright (for HTML rendering in html2pptx)
- react-icons:
npm install -g react-icons react react-dom (for icons)
- sharp:
npm install -g sharp (for SVG rasterization and image processing)
- LibreOffice:
sudo apt-get install libreoffice (for PDF conversion)
- Poppler:
sudo apt-get install poppler-utils (for pdftoppm to convert PDF to images)
- defusedxml:
pip install defusedxml (for secure XML parsing)
1---2name: pptx-43description: Presentation creation, editing, and analysis. When Claude needs to work with presentations (.pptx files) for: (1) Creating new presentations, (2) Modifying or editing content, (3) Working with layouts, (4) Adding comments or speaker notes, or any other presentation tasks4license: Proprietary. LICENSE.txt has complete terms5---6
7# PPTX creation, editing, and analysis
8
9## Overview
10
11A user may ask you to create, edit, or analyze the contents of a .pptx file. A .pptx file is essentially a ZIP archive containing XML files and other resources that you can read or edit. You have different tools and workflows available for different tasks.
12
13## Reading and analyzing content
14
15### Text extraction
16If you just need to read the text contents of a presentation, you should convert the document to markdown:
17
18```bash
19# Convert document to markdown
20python -m markitdown path-to-file.pptx
21```
22
23### Raw XML access
24You need raw XML access for: comments, speaker notes, slide layouts, animations, design elements, and complex formatting. For any of these features, you'll need to unpack a presentation and read its raw XML contents.
25
26#### Unpacking a file
27`python ooxml/scripts/unpack.py <office_file> <output_dir>`
28
29**Note**: The unpack.py script is located at `skills/pptx/ooxml/scripts/unpack.py` relative to the project root. If the script doesn't exist at this path, use `find . -name "unpack.py"` to locate it.
30
31#### Key file structures
32* `ppt/presentation.xml` - Main presentation metadata and slide references
33* `ppt/slides/slide{N}.xml` - Individual slide contents (slide1.xml, slide2.xml, etc.)
34* `ppt/notesSlides/notesSlide{N}.xml` - Speaker notes for each slide
35* `ppt/comments/modernComment_*.xml` - Comments for specific slides
36* `ppt/slideLayouts/` - Layout templates for slides
37* `ppt/slideMasters/` - Master slide templates
38* `ppt/theme/` - Theme and styling information
39* `ppt/media/` - Images and other media files
40
41#### Typography and color extraction
42**When given an example design to emulate**: Always analyze the presentation's typography and colors first using the methods below:
431. **Read theme file**: Check `ppt/theme/theme1.xml` for colors (`<a:clrScheme>`) and fonts (`<a:fontScheme>`)
442. **Sample slide content**: Examine `ppt/slides/slide1.xml` for actual font usage (`<a:rPr>`) and colors
453. **Search for patterns**: Use grep to find color (`<a:solidFill>`, `<a:srgbClr>`) and font references across all XML files
46
47## Creating a new PowerPoint presentation **without a template**
48
49When creating a new PowerPoint presentation from scratch, use the **html2pptx** workflow to convert HTML slides to PowerPoint with accurate positioning.
50
51### Design Principles
52
53**CRITICAL**: Before creating any presentation, analyze the content and choose appropriate design elements:
541. **Consider the subject matter**: What is this presentation about? What tone, industry, or mood does it suggest?
552. **Check for branding**: If the user mentions a company/organization, consider their brand colors and identity
563. **Match palette to content**: Select colors that reflect the subject
574. **State your approach**: Explain your design choices before writing code
58
59**Requirements**:
60- ✅ State your content-informed design approach BEFORE writing code
61- ✅ Use web-safe fonts only: Arial, Helvetica, Times New Roman, Georgia, Courier New, Verdana, Tahoma, Trebuchet MS, Impact
62- ✅ Create clear visual hierarchy through size, weight, and color
63- ✅ Ensure readability: strong contrast, appropriately sized text, clean alignment
64- ✅ Be consistent: repeat patterns, spacing, and visual language across slides
65
66#### Color Palette Selection
67
68**Choosing colors creatively**:
69- **Think beyond defaults**: What colors genuinely match this specific topic? Avoid autopilot choices.
70- **Consider multiple angles**: Topic, industry, mood, energy level, target audience, brand identity (if mentioned)
71- **Be adventurous**: Try unexpected combinations - a healthcare presentation doesn't have to be green, finance doesn't have to be navy
72- **Build your palette**: Pick 3-5 colors that work together (dominant colors + supporting tones + accent)
73- **Ensure contrast**: Text must be clearly readable on backgrounds
74
75**Example color palettes** (use these to spark creativity - choose one, adapt it, or create your own):
76
771. **Classic Blue**: Deep navy (#1C2833), slate gray (#2E4053), silver (#AAB7B8), off-white (#F4F6F6)
782. **Teal & Coral**: Teal (#5EA8A7), deep teal (#277884), coral (#FE4447), white (#FFFFFF)
793. **Bold Red**: Red (#C0392B), bright red (#E74C3C), orange (#F39C12), yellow (#F1C40F), green (#2ECC71)
804. **Warm Blush**: Mauve (#A49393), blush (#EED6D3), rose (#E8B4B8), cream (#FAF7F2)
815. **Burgundy Luxury**: Burgundy (#5D1D2E), crimson (#951233), rust (#C15937), gold (#997929)
826. **Deep Purple & Emerald**: Purple (#B165FB), dark blue (#181B24), emerald (#40695B), white (#FFFFFF)
837. **Cream & Forest Green**: Cream (#FFE1C7), forest green (#40695B), white (#FCFCFC)
848. **Pink & Purple**: Pink (#F8275B), coral (#FF574A), rose (#FF737D), purple (#3D2F68)
859. **Lime & Plum**: Lime (#C5DE82), plum (#7C3A5F), coral (#FD8C6E), blue-gray (#98ACB5)
8610. **Black & Gold**: Gold (#BF9A4A), black (#000000), cream (#F4F6F6)
8711. **Sage & Terracotta**: Sage (#87A96B), terracotta (#E07A5F), cream (#F4F1DE), charcoal (#2C2C2C)
8812. **Charcoal & Red**: Charcoal (#292929), red (#E33737), light gray (#CCCBCB)
8913. **Vibrant Orange**: Orange (#F96D00), light gray (#F2F2F2), charcoal (#222831)
9014. **Forest Green**: Black (#191A19), green (#4E9F3D), dark green (#1E5128), white (#FFFFFF)
9115. **Retro Rainbow**: Purple (#722880), pink (#D72D51), orange (#EB5C18), amber (#F08800), gold (#DEB600)
9216. **Vintage Earthy**: Mustard (#E3B448), sage (#CBD18F), forest green (#3A6B35), cream (#F4F1DE)
9317. **Coastal Rose**: Old rose (#AD7670), beaver (#B49886), eggshell (#F3ECDC), ash gray (#BFD5BE)
9418. **Orange & Turquoise**: Light orange (#FC993E), grayish turquoise (#667C6F), white (#FCFCFC)
95
96#### Visual Details Options
97
98**Geometric Patterns**:
99- Diagonal section dividers instead of horizontal
100- Asymmetric column widths (30/70, 40/60, 25/75)
101- Rotated text headers at 90° or 270°
102- Circular/hexagonal frames for images
103- Triangular accent shapes in corners
104- Overlapping shapes for depth
105
106**Border & Frame Treatments**:
107- Thick single-color borders (10-20pt) on one side only
108- Double-line borders with contrasting colors
109- Corner brackets instead of full frames
110- L-shaped borders (top+left or bottom+right)
111- Underline accents beneath headers (3-5pt thick)
112
113**Typography Treatments**:
114- Extreme size contrast (72pt headlines vs 11pt body)
115- All-caps headers with wide letter spacing
116- Numbered sections in oversized display type
117- Monospace (Courier New) for data/stats/technical content
118- Condensed fonts (Arial Narrow) for dense information
119- Outlined text for emphasis
120
121**Chart & Data Styling**:
122- Monochrome charts with single accent color for key data
123- Horizontal bar charts instead of vertical
124- Dot plots instead of bar charts
125- Minimal gridlines or none at all
126- Data labels directly on elements (no legends)
127- Oversized numbers for key metrics
128
129**Layout Innovations**:
130- Full-bleed images with text overlays
131- Sidebar column (20-30% width) for navigation/context
132- Modular grid systems (3×3, 4×4 blocks)
133- Z-pattern or F-pattern content flow
134- Floating text boxes over colored shapes
135- Magazine-style multi-column layouts
136
137**Background Treatments**:
138- Solid color blocks occupying 40-60% of slide
139- Gradient fills (vertical or diagonal only)
140- Split backgrounds (two colors, diagonal or vertical)
141- Edge-to-edge color bands
142- Negative space as a design element
143
144### Layout Tips
145**When creating slides with charts or tables:**
146- **Two-column layout (PREFERRED)**: Use a header spanning the full width, then two columns below - text/bullets in one column and the featured content in the other. This provides better balance and makes charts/tables more readable. Use flexbox with unequal column widths (e.g., 40%/60% split) to optimize space for each content type.
147- **Full-slide layout**: Let the featured content (chart/table) take up the entire slide for maximum impact and readability
148- **NEVER vertically stack**: Do not place charts/tables below text in a single column - this causes poor readability and layout issues
149
150### Workflow: Step-by-Step Deterministic Process
151
152**⚠️ CRITICAL: Follow these steps IN ORDER. Each step creates ONE file before moving to the next.**
153
154#### Step 1: Read Documentation (MANDATORY)
155Read [`html2pptx.md`](html2pptx.md) completely from start to finish. **NEVER set any range limits.** This file contains:
156- HTML syntax rules and supported elements
157- Dimension requirements for different layouts
158- Placeholder usage for charts/images
159- PptxGenJS chart/table examples
160
161#### Step 2: Create HTML Files (One HTML file per slide)
162
163**IMPORTANT: Create separate HTML files - one file for each slide.**
164
165Example for a 3-slide presentation:
166
167```html
168<!-- slide1.html -->
169<!DOCTYPE html>
170<html>
171<head>
172<style>
173body {
174 width: 720pt;
175 height: 405pt; /* 16:9 layout */
176 margin: 0;
177 padding: 48pt;
178 box-sizing: border-box;
179 font-family: Arial, sans-serif;
180 background: #1C2833;
181}
182h1 { font-size: 36pt; color: #F4F6F6; margin: 0 0 24pt 0; }
183p { font-size: 18pt; color: #F4F6F6; }
184</style>
185</head>
186<body>
187 <h1>Slide 1 Title</h1>
188 <p>Content for first slide</p>
189</body>
190</html>
191```
192
193```html
194<!-- slide2.html -->
195<!DOCTYPE html>
196<html>
197<head>
198<style>
199body {
200 width: 720pt;
201 height: 405pt;
202 margin: 0;
203 padding: 48pt;
204 box-sizing: border-box;
205 font-family: Arial, sans-serif;
206 background: #2E4053;
207}
208h1 { font-size: 36pt; color: #F4F6F6; margin: 0 0 24pt 0; }
209p { font-size: 18pt; color: #F4F6F6; }
210</style>
211</head>
212<body>
213 <h1>Slide 2 Title</h1>
214 <p>Content for second slide</p>
215 <!-- Placeholder for chart -->
216 <div id="chart1" class="placeholder" style="width: 600pt; height: 260pt;"></div>
217</body>
218</html>
219```
220
221```html
222<!-- slide3.html - similar structure -->
223```
224
225**Critical HTML Rules:**
226- ✅ ONE HTML file = ONE slide
227- ✅ ALL text MUST be in `<p>`, `<h1>`-`<h6>`, `<ul>`, `<ol>` tags
228- ✅ Use `class="placeholder"` with unique `id` for chart/image areas
229- ✅ Use web-safe fonts only (Arial, Helvetica, Times New Roman, etc.)
230- ✅ Rasterize gradients/icons as PNG FIRST using Sharp, then reference
231- ✅ Two-column or full-slide layout for charts/tables (NEVER vertical stack)
232
233#### Step 3: Create JavaScript Presentation Script (ONE file)
234
235**CRITICAL - Script Path Resolution**
236
237When generating a script that uses `html2pptx.js`, you MUST use the correct path pattern. The skills connector sets the `SKILL_DIR` environment variable pointing to this skill's root directory.
238
239**Required pattern for path resolution:**
240```javascript
241const path = require('path');
242
243// CRITICAL: Use SKILL_DIR environment variable to locate html2pptx
244// SKILL_DIR points to the skill's root (e.g., D:\ops-data\agent-skills\pptx)
245// This works regardless of where the script is run from
246const skillDir = process.env.SKILL_DIR;
247if (!skillDir) {
248 console.error('ERROR: SKILL_DIR environment variable not set');
249 console.error('This script must be run via the skills connector');
250 process.exit(1);
251}
252
253const html2pptx = require(path.join(skillDir, 'scripts', 'html2pptx'));
254```
255
256**Why this pattern is required:**
257- ❌ `require('./scripts/html2pptx')` - Fails if script runs from a different directory
258- ❌ `require('../agent-skills/pptx/scripts/html2pptx')` - Breaks if connector mount path changes
259- ✅ `require(path.join(process.env.SKILL_DIR, 'scripts', 'html2pptx'))` - Always works correctly
260
261**Full example script:**
262
263```javascript
264// create-presentation.js
265const pptxgen = require('pptxgenjs');
266const path = require('path');
267
268// CRITICAL: Use SKILL_DIR environment variable to locate html2pptx
269const skillDir = process.env.SKILL_DIR;
270if (!skillDir) {
271 console.error('ERROR: SKILL_DIR environment variable not set');
272 console.error('This script must be run via the skills connector');
273 process.exit(1);
274}
275
276const html2pptx = require(path.join(skillDir, 'scripts', 'html2pptx'));
277
278async function main() {
279 try {
280 console.log('Creating presentation...');
281 console.log('SKILL_DIR:', skillDir);
282
283 const pptx = new pptxgen();
284 pptx.layout = 'LAYOUT_16x9';
285 pptx.title = 'My Presentation';
286
287 // Convert each HTML file to a slide
288 // ONE html2pptx() call per HTML file
289 await html2pptx(path.join(__dirname, 'slide1.html'), pptx);
290 await html2pptx(path.join(__dirname, 'slide2.html'), pptx);
291 await html2pptx(path.join(__dirname, 'slide3.html'), pptx);
292
293 // Add charts to placeholders if needed
294 // (See complete example in html2pptx.md)
295
296 await pptx.writeFile({ fileName: 'output.pptx' });
297 console.log('✓ Presentation created: output.pptx');
298 process.exit(0);
299
300 } catch (error) {
301 console.error('✗ Error:', error.message);
302 console.error('Stack trace:', error.stack);
303 process.exit(1);
304 }
305}
306
307main();
308```
309
310#### Step 4: Run the JavaScript File
311
312**Use run_script with absolute path from write response**
313
314After writing the script in Step 3, the write response includes the full absolute path. **Use this exact path** with `run_script` and add the `skillId` parameter.
315
316✅ **Correct approach**:
317```json
318// Step 3 write response: { "path": "D:\\ops-data\\create-presentation.js" }
319
320// Step 4: Run using the FULL PATH from response
321{
322 "action": "remote_tool_call",
323 "target": "filesystem",
324 "params": {
325 "agent_id": "local-dev",
326 "connector_id": "local-files", // Same connector you used for write
327 "command": "run_script",
328 "path": "D:\\ops-data\\create-presentation.js", // FULL path from write response
329 "skillId": "pptx", // Sets SKILL_DIR environment variable
330 "type": "node"
331 }
332}
333```
334
335**How skillId works**:
336- Sets environment variable: `SKILL_DIR=/path/to/agent-skills/pptx`
337- Your script uses `process.env.SKILL_DIR` to locate html2pptx.js
338- Works regardless of which connector the script is stored in
339
340**Check the result:**
341- ✅ `exitCode: 0` → Success! PPTX created
342- ❌ `exitCode: 1` → Read stdout/stderr error message, fix issues, retry
343- ❌ "Script not found" → Check connector_id matches where you wrote the file
344
345#### Step 5: Visual Validation
346
347Generate thumbnails and inspect:
348```bash
349python scripts/thumbnail.py output.pptx workspace/thumbnails --cols 4
350```
351
352Read thumbnail image and check for:
353- Text cutoff or overlap
354- Positioning issues
355- Contrast problems
356
357If issues found, adjust HTML and regenerate (return to Step 2).
358
359---
360
361**📋 Summary - Files to Create:**
3621. **HTML files**: `slide1.html`, `slide2.html`, `slide3.html`, ... (ONE file per slide)
3632. **JavaScript file**: `create-presentation.js` (ONE script that calls html2pptx() for each HTML file)
3643. **Run**: `node create-presentation.js` → `output.pptx` created
3654. **Validate**: Generate thumbnails to check layout
366
367### Troubleshooting Script Errors
368
369**If your script fails with exit code 1:**
370
371The error output from `run_script` should contain diagnostic information. Look for:
372
3731. **Module not found errors**:
374 ```
375 Error: Cannot find module './scripts/html2pptx'
376 ```
377 **Cause**: Not using `process.env.SKILL_DIR` for path resolution
378 **Fix**: Update require statement to use `path.join(process.env.SKILL_DIR, 'scripts', 'html2pptx')`
379
3802. **SKILL_DIR not set**:
381 ```
382 ERROR: SKILL_DIR environment variable not set
383 ```
384 **Cause**: Missing `skillId` parameter in run_script call
385 **Fix**: Add `"skillId": "pptx"` to the params when calling run_script. This sets the SKILL_DIR environment variable that your script needs to locate html2pptx.
386
3873. **HTML conversion errors**:
388 ```
389 Error: HTML file must have exactly one body element
390 ```
391 **Cause**: Invalid HTML structure
392 **Fix**: Verify HTML files have proper structure with single `<html>`, `<head>`, and `<body>` tags
393
3944. **Dimension errors**:
395 ```
396 Error: Body dimensions must be 720pt × 405pt for 16:9
397 ```
398 **Cause**: Incorrect HTML body dimensions
399 **Fix**: Ensure `body { width: 720pt; height: 405pt; }` in CSS
400
401**Debug template to add to your script:**
402
403```javascript
404// Add this at the start of your script for debugging
405console.log('=== Debug Info ===');
406console.log('Current directory:', __dirname);
407console.log('SKILL_DIR:', process.env.SKILL_DIR);
408console.log('html2pptx path:', path.join(process.env.SKILL_DIR, 'scripts', 'html2pptx'));
409console.log('==================');
410```
411
412**Common mistakes to avoid:**
413
414- ❌ Using relative paths: `require('./scripts/html2pptx')`
415- ❌ Using hard-coded paths: `require('D:\\ops-data\\agent-skills\\pptx\\scripts\\html2pptx')`
416- ❌ Missing error handler: Not catching errors means no diagnostic output
417- ❌ Not logging SKILL_DIR: Makes debugging path issues harder
418- ✅ Always use: `require(path.join(process.env.SKILL_DIR, 'scripts', 'html2pptx'))`
419- ✅ Always log: Console output helps diagnose failures
420- ✅ Always catch: Wrap main logic in try-catch with detailed error logging
421
422**Reference example:**
423
424See [`example-create-presentation.js`](example-create-presentation.js) for a complete working example with:
425- Correct SKILL_DIR usage
426- Comprehensive error handling
427- Detailed diagnostic output
428- Proper exit codes
429
430## Editing an existing PowerPoint presentation
431
432When edit slides in an existing PowerPoint presentation, you need to work with the raw Office Open XML (OOXML) format. This involves unpacking the .pptx file, editing the XML content, and repacking it.
433
434### Workflow
4351. **MANDATORY - READ ENTIRE FILE**: Read [`ooxml.md`](ooxml.md) (~500 lines) completely from start to finish. **NEVER set any range limits when reading this file.** Read the full file content for detailed guidance on OOXML structure and editing workflows before any presentation editing.
4362. Unpack the presentation: `python ooxml/scripts/unpack.py <office_file> <output_dir>`
4373. Edit the XML files (primarily `ppt/slides/slide{N}.xml` and related files)
4384. **CRITICAL**: Validate immediately after each edit and fix any validation errors before proceeding: `python ooxml/scripts/validate.py <dir> --original <file>`
4395. Pack the final presentation: `python ooxml/scripts/pack.py <input_directory> <office_file>`
440
441## Creating a new PowerPoint presentation **using a template**
442
443When you need to create a presentation that follows an existing template's design, you'll need to duplicate and re-arrange template slides before then replacing placeholder context.
444
445### Workflow
4461. **Extract template text AND create visual thumbnail grid**:
447 * Extract text: `python -m markitdown template.pptx > template-content.md`
448 * Read `template-content.md`: Read the entire file to understand the contents of the template presentation. **NEVER set any range limits when reading this file.**
449 * Create thumbnail grids: `python scripts/thumbnail.py template.pptx`
450 * See [Creating Thumbnail Grids](#creating-thumbnail-grids) section for more details
451
4522. **Analyze template and save inventory to a file**:
453 * **Visual Analysis**: Review thumbnail grid(s) to understand slide layouts, design patterns, and visual structure
454 * Create and save a template inventory file at `template-inventory.md` containing:
455 ```markdown
456 # Template Inventory Analysis
457 **Total Slides: [count]**
458 **IMPORTANT: Slides are 0-indexed (first slide = 0, last slide = count-1)**
459
460 ## [Category Name]
461 - Slide 0: [Layout code if available] - Description/purpose
462 - Slide 1: [Layout code] - Description/purpose
463 - Slide 2: [Layout code] - Description/purpose
464 [... EVERY slide must be listed individually with its index ...]
465 ```
466 * **Using the thumbnail grid**: Reference the visual thumbnails to identify:
467 - Layout patterns (title slides, content layouts, section dividers)
468 - Image placeholder locations and counts
469 - Design consistency across slide groups
470 - Visual hierarchy and structure
471 * This inventory file is REQUIRED for selecting appropriate templates in the next step
472
4733. **Create presentation outline based on template inventory**:
474 * Review available templates from step 2.
475 * Choose an intro or title template for the first slide. This should be one of the first templates.
476 * Choose safe, text-based layouts for the other slides.
477 * **CRITICAL: Match layout structure to actual content**:
478 - Single-column layouts: Use for unified narrative or single topic
479 - Two-column layouts: Use ONLY when you have exactly 2 distinct items/concepts
480 - Three-column layouts: Use ONLY when you have exactly 3 distinct items/concepts
481 - Image + text layouts: Use ONLY when you have actual images to insert
482 - Quote layouts: Use ONLY for actual quotes from people (with attribution), never for emphasis
483 - Never use layouts with more placeholders than you have content
484 - If you have 2 items, don't force them into a 3-column layout
485 - If you have 4+ items, consider breaking into multiple slides or using a list format
486 * Count your actual content pieces BEFORE selecting the layout
487 * Verify each placeholder in the chosen layout will be filled with meaningful content
488 * Select one option representing the **best** layout for each content section.
489 * Save `outline.md` with content AND template mapping that leverages available designs
490 * Example template mapping:
491 ```
492 # Template slides to use (0-based indexing)
493 # WARNING: Verify indices are within range! Template with 73 slides has indices 0-72
494 # Mapping: slide numbers from outline -> template slide indices
495 template_mapping = [
496 0, # Use slide 0 (Title/Cover)
497 34, # Use slide 34 (B1: Title and body)
498 34, # Use slide 34 again (duplicate for second B1)
499 50, # Use slide 50 (E1: Quote)
500 54, # Use slide 54 (F2: Closing + Text)
501 ]
502 ```
503
5044. **Duplicate, reorder, and delete slides using `rearrange.py`**:
505 * Use the `scripts/rearrange.py` script to create a new presentation with slides in the desired order:
506 ```bash
507 python scripts/rearrange.py template.pptx working.pptx 0,34,34,50,52
508 ```
509 * The script handles duplicating repeated slides, deleting unused slides, and reordering automatically
510 * Slide indices are 0-based (first slide is 0, second is 1, etc.)
511 * The same slide index can appear multiple times to duplicate that slide
512
5135. **Extract ALL text using the `inventory.py` script**:
514 * **Run inventory extraction**:
515 ```bash
516 python scripts/inventory.py working.pptx text-inventory.json
517 ```
518 * **Read text-inventory.json**: Read the entire text-inventory.json file to understand all shapes and their properties. **NEVER set any range limits when reading this file.**
519
520 * The inventory JSON structure:
521 ```json
522 {
523 "slide-0": {
524 "shape-0": {
525 "placeholder_type": "TITLE", // or null for non-placeholders
526 "left": 1.5, // position in inches
527 "top": 2.0,
528 "width": 7.5,
529 "height": 1.2,
530 "paragraphs": [
531 {
532 "text": "Paragraph text",
533 // Optional properties (only included when non-default):
534 "bullet": true, // explicit bullet detected
535 "level": 0, // only included when bullet is true
536 "alignment": "CENTER", // CENTER, RIGHT (not LEFT)
537 "space_before": 10.0, // space before paragraph in points
538 "space_after": 6.0, // space after paragraph in points
539 "line_spacing": 22.4, // line spacing in points
540 "font_name": "Arial", // from first run
541 "font_size": 14.0, // in points
542 "bold": true,
543 "italic": false,
544 "underline": false,
545 "color": "FF0000" // RGB color
546 }
547 ]
548 }
549 }
550 }
551 ```
552
553 * Key features:
554 - **Slides**: Named as "slide-0", "slide-1", etc.
555 - **Shapes**: Ordered by visual position (top-to-bottom, left-to-right) as "shape-0", "shape-1", etc.
556 - **Placeholder types**: TITLE, CENTER_TITLE, SUBTITLE, BODY, OBJECT, or null
557 - **Default font size**: `default_font_size` in points extracted from layout placeholders (when available)
558 - **Slide numbers are filtered**: Shapes with SLIDE_NUMBER placeholder type are automatically excluded from inventory
559 - **Bullets**: When `bullet: true`, `level` is always included (even if 0)
560 - **Spacing**: `space_before`, `space_after`, and `line_spacing` in points (only included when set)
561 - **Colors**: `color` for RGB (e.g., "FF0000"), `theme_color` for theme colors (e.g., "DARK_1")
562 - **Properties**: Only non-default values are included in the output
563
5646. **Generate replacement text and save the data to a JSON file**
565 Based on the text inventory from the previous step:
566 - **CRITICAL**: First verify which shapes exist in the inventory - only reference shapes that are actually present
567 - **VALIDATION**: The replace.py script will validate that all shapes in your replacement JSON exist in the inventory
568 - If you reference a non-existent shape, you'll get an error showing available shapes
569 - If you reference a non-existent slide, you'll get an error indicating the slide doesn't exist
570 - All validation errors are shown at once before the script exits
571 - **IMPORTANT**: The replace.py script uses inventory.py internally to identify ALL text shapes
572 - **AUTOMATIC CLEARING**: ALL text shapes from the inventory will be cleared unless you provide "paragraphs" for them
573 - Add a "paragraphs" field to shapes that need content (not "replacement_paragraphs")
574 - Shapes without "paragraphs" in the replacement JSON will have their text cleared automatically
575 - Paragraphs with bullets will be automatically left aligned. Don't set the `alignment` property on when `"bullet": true`
576 - Generate appropriate replacement content for placeholder text
577 - Use shape size to determine appropriate content length
578 - **CRITICAL**: Include paragraph properties from the original inventory - don't just provide text
579 - **IMPORTANT**: When bullet: true, do NOT include bullet symbols (•, -, *) in text - they're added automatically
580 - **ESSENTIAL FORMATTING RULES**:
581 - Headers/titles should typically have `"bold": true`
582 - List items should have `"bullet": true, "level": 0` (level is required when bullet is true)
583 - Preserve any alignment properties (e.g., `"alignment": "CENTER"` for centered text)
584 - Include font properties when different from default (e.g., `"font_size": 14.0`, `"font_name": "Lora"`)
585 - Colors: Use `"color": "FF0000"` for RGB or `"theme_color": "DARK_1"` for theme colors
586 - The replacement script expects **properly formatted paragraphs**, not just text strings
587 - **Overlapping shapes**: Prefer shapes with larger default_font_size or more appropriate placeholder_type
588 - Save the updated inventory with replacements to `replacement-text.json`
589 - **WARNING**: Different template layouts have different shape counts - always check the actual inventory before creating replacements
590
591 Example paragraphs field showing proper formatting:
592 ```json
593 "paragraphs": [
594 {
595 "text": "New presentation title text",
596 "alignment": "CENTER",
597 "bold": true
598 },
599 {
600 "text": "Section Header",
601 "bold": true
602 },
603 {
604 "text": "First bullet point without bullet symbol",
605 "bullet": true,
606 "level": 0
607 },
608 {
609 "text": "Red colored text",
610 "color": "FF0000"
611 },
612 {
613 "text": "Theme colored text",
614 "theme_color": "DARK_1"
615 },
616 {
617 "text": "Regular paragraph text without special formatting"
618 }
619 ]
620 ```
621
622 **Shapes not listed in the replacement JSON are automatically cleared**:
623 ```json
624 {
625 "slide-0": {
626 "shape-0": {
627 "paragraphs": [...] // This shape gets new text
628 }
629 // shape-1 and shape-2 from inventory will be cleared automatically
630 }
631 }
632 ```
633
634 **Common formatting patterns for presentations**:
635 - Title slides: Bold text, sometimes centered
636 - Section headers within slides: Bold text
637 - Bullet lists: Each item needs `"bullet": true, "level": 0`
638 - Body text: Usually no special properties needed
639 - Quotes: May have special alignment or font properties
640
6417. **Apply replacements using the `replace.py` script**
642 ```bash
643 python scripts/replace.py working.pptx replacement-text.json output.pptx
644 ```
645
646 The script will:
647 - First extract the inventory of ALL text shapes using functions from inventory.py
648 - Validate that all shapes in the replacement JSON exist in the inventory
649 - Clear text from ALL shapes identified in the inventory
650 - Apply new text only to shapes with "paragraphs" defined in the replacement JSON
651 - Preserve formatting by applying paragraph properties from the JSON
652 - Handle bullets, alignment, font properties, and colors automatically
653 - Save the updated presentation
654
655 Example validation errors:
656 ```
657 ERROR: Invalid shapes in replacement JSON:
658 - Shape 'shape-99' not found on 'slide-0'. Available shapes: shape-0, shape-1, shape-4
659 - Slide 'slide-999' not found in inventory
660 ```
661
662 ```
663 ERROR: Replacement text made overflow worse in these shapes:
664 - slide-0/shape-2: overflow worsened by 1.25" (was 0.00", now 1.25")
665 ```
666
667## Creating Thumbnail Grids
668
669To create visual thumbnail grids of PowerPoint slides for quick analysis and reference:
670
671```bash
672python scripts/thumbnail.py template.pptx [output_prefix]
673```
674
675**Features**:
676- Creates: `thumbnails.jpg` (or `thumbnails-1.jpg`, `thumbnails-2.jpg`, etc. for large decks)
677- Default: 5 columns, max 30 slides per grid (5×6)
678- Custom prefix: `python scripts/thumbnail.py template.pptx my-grid`
679 - Note: The output prefix should include the path if you want output in a specific directory (e.g., `workspace/my-grid`)
680- Adjust columns: `--cols 4` (range: 3-6, affects slides per grid)
681- Grid limits: 3 cols = 12 slides/grid, 4 cols = 20, 5 cols = 30, 6 cols = 42
682- Slides are zero-indexed (Slide 0, Slide 1, etc.)
683
684**Use cases**:
685- Template analysis: Quickly understand slide layouts and design patterns
686- Content review: Visual overview of entire presentation
687- Navigation reference: Find specific slides by their visual appearance
688- Quality check: Verify all slides are properly formatted
689
690**Examples**:
691```bash
692# Basic usage
693python scripts/thumbnail.py presentation.pptx
694
695# Combine options: custom name, columns
696python scripts/thumbnail.py template.pptx analysis --cols 4
697```
698
699## Converting Slides to Images
700
701To visually analyze PowerPoint slides, convert them to images using a two-step process:
702
7031. **Convert PPTX to PDF**:
704 ```bash
705 soffice --headless --convert-to pdf template.pptx
706 ```
707
7082. **Convert PDF pages to JPEG images**:
709 ```bash
710 pdftoppm -jpeg -r 150 template.pdf slide
711 ```
712 This creates files like `slide-1.jpg`, `slide-2.jpg`, etc.
713
714Options:
715- `-r 150`: Sets resolution to 150 DPI (adjust for quality/size balance)
716- `-jpeg`: Output JPEG format (use `-png` for PNG if preferred)
717- `-f N`: First page to convert (e.g., `-f 2` starts from page 2)
718- `-l N`: Last page to convert (e.g., `-l 5` stops at page 5)
719- `slide`: Prefix for output files
720
721Example for specific range:
722```bash
723pdftoppm -jpeg -r 150 -f 2 -l 5 template.pdf slide # Converts only pages 2-5
724```
725
726## Code Style Guidelines
727**IMPORTANT**: When generating code for PPTX operations:
728- Write concise code
729- Avoid verbose variable names and redundant operations
730- Avoid unnecessary print statements
731
732## Dependencies
733
734Required dependencies (should already be installed):
735
736- **markitdown**: `pip install "markitdown[pptx]"` (for text extraction from presentations)
737- **pptxgenjs**: `npm install -g pptxgenjs` (for creating presentations via html2pptx)
738- **playwright**: `npm install -g playwright` (for HTML rendering in html2pptx)
739- **react-icons**: `npm install -g react-icons react react-dom` (for icons)
740- **sharp**: `npm install -g sharp` (for SVG rasterization and image processing)
741- **LibreOffice**: `sudo apt-get install libreoffice` (for PDF conversion)
742- **Poppler**: `sudo apt-get install poppler-utils` (for pdftoppm to convert PDF to images)
743- **defusedxml**: `pip install defusedxml` (for secure XML parsing)