html2tex
Convert HTML files to LaTeX and PDF with the bundled html2tex.py script.
Run
From this skill directory, run:
python html2tex.py <input.html> [--output-dir DIR] [--no-compile]
Inputs
Assume the input HTML may contain:
- tables with
colspanandrowspan - embedded images
- bold, italic, underline, and strikethrough
- MathJax-style formulas
Use this skill especially for HTML produced by a Word-to-HTML workflow.
Outputs
Produce:
<input>.tex<input>.pdfwhen compilation is enabled
If --output-dir is omitted, write output next to the input file.
Workflow
- Confirm the input file exists.
- Run
html2tex.py. - If the user asked for PDF and compilation succeeds, report both
.texand.pdf. - If XeLaTeX is missing or compilation fails, rerun with
--no-compileonly when needed and report that only.texwas generated.
Dependencies
Require:
- Python 3
- XeLaTeX for PDF compilation
If XeLaTeX is unavailable, do not pretend PDF generation succeeded.
Notes
- Preserve tables and images where possible.
- Prefer reporting the generated file paths instead of pasting LaTeX output into the response.
- If the HTML appears broken, mention that the converter can only preserve structure that exists in the source HTML.