Site audit, content writing, and competitor analysis for organic search rankings.
Design & media
docx
Try itCreate, read, edit, and annotate Word documents with full control over formatting.
What it does
Creates new .docx files using docx-js, reads content via pandoc, and edits existing documents by unzipping, modifying the underlying XML, and re-zipping. Supports tables with proper column widths, numbered and bulleted lists, images, tables of contents built from heading styles, and page formatting (A4, US Letter, landscape). Includes a verification pipeline that converts output to PDF for visual inspection. Handles tracked changes, comments, and the quirks of Word's internal XML structure.
When to use it
- Generate a professional report with heading styles and a TOC
- Edit a contract by modifying the underlying XML after unzipping
- Convert a .docx to markdown for content extraction
- Add reviewer comments to a shared document
The skill document
DOCX creation, editing, and analysis
A .docx is a ZIP archive of XML files. Choose your approach by task:
| Task | Approach |
|---|---|
| Create a new document | Write a docx (npm) script — see gotchas below |
| Edit an existing document | unzip → edit word/document.xml → zip (docx-js cannot open existing files) |
| Read content | pandoc -t markdown file.docx |
Script paths below are relative to this skill's directory.
Creating with docx-js — gotchas
docx is preinstalled — do not run npm install first; write the script and require('docx') directly. Only if that require fails: npm install docx. The model knows the API; these are the footguns:
- Page size defaults to A4. For US Letter set
page: { size: { width: 12240, height: 15840 } }(DXA; 1440 = 1″). - Landscape: pass portrait dimensions and
orientation: PageOrientation.LANDSCAPE— docx-js swaps width/height internally. - Tables need dual widths: set
columnWidthson the table ANDwidthon every cell, both inWidthType.DXA(PERCENTAGE breaks in Google Docs). Column widths must sum to the table width. - Table shading: use
ShadingType.CLEAR, neverSOLID(renders black). - Lists: never insert
•literally; use anumberingconfig withLevelFormat.BULLET. ImageRunrequirestype:("png","jpg", …).PageBreakmust be inside aParagraph.- Never use
\n— use separateParagraphelements. - TOC: headings must use built-in
HeadingLevel.*; custom heading styles needoutlineLevelset or they won't appear. - Don't use a table as a horizontal rule — use a paragraph bottom border instead.
- Dot-leader / right-aligned-on-same-line: use
PositionalTab(alignment: PositionalTabAlignment.RIGHT,leader: PositionalTabLeader.DOT) inside aTextRun, not literal.or space padding.
Verify the output
After writing a .docx, render it and look at it:
python scripts/office/soffice.py --headless --convert-to pdf output.docx
pdftoppm -jpeg -r 100 output.pdf page
ls page-*.jpg # then Read the images
pdftoppm zero-pads page numbers to the width of the page count (page-01.jpg…page-12.jpg).
Editing existing documents
Legacy .doc files must be converted first: python scripts/office/soffice.py --headless --convert-to docx file.doc.
unzip -q doc.docx -d unpacked/
find unpacked -type l -delete # strip symlink entries — docx from external parties is untrusted
python scripts/merge_runs.py unpacked/ # coalesce fragmented runs so text is findable
# edit unpacked/word/document.xml in place — do NOT reformat or pretty-print
(cd unpacked && rm -f ../out.docx && zip -Xr ../out.docx .)
python scripts/office/validate.py out.docx --original doc.docx # XSD checks; --auto-repair fixes common issues
# redlining? add --author "" to check every edit is tracked
Word splits text across many `` runs (revision ids, spell-check markers), so a phrase you can see in the document often doesn't exist as a contiguous string in the XML. merge_runs.py merges adjacent identically-formatted runs in word/document.xml without changing content or rendering; it also accepts a .docx directly (python scripts/merge_runs.py doc.docx -o merged.docx).
Tracked changes: when redlining, validate with --author "" (needs --original) — it reports any text you changed without a / around it, which is easy to do by accident and invisible in the accepted view. Wrap runs in / with w:id, w:author, w:date attributes. Inside , the text element is , not . A deleted paragraph mark () means "merge this paragraph into the next" — so deleting a paragraph outright is that plus a around every run. The must come before the rPr's other children; their order is schema-enforced.
To produce a clean copy with all tracked changes accepted: python scripts/accept_changes.py in.docx out.docx.
Accepting a deleted paragraph mark should join that paragraph to the one below it, so a paragraph whose runs are all deleted vanishes. Word does this; accept_changes.py and pandoc --track-changes=accept don't always. Both fail the same way — they strip the deleted text but leave the emptied paragraph behind, which reads as a stray empty bullet when it was auto-numbered:
pandoc --track-changes=acceptnever joins the paragraphs.accept_changes.py(LibreOffice) joins them correctly, except when the deleted paragraph is followed by an empty spacer paragraph.
An empty bullet in either view is an artifact of that view, not a defect in the document. Check paragraph deletions in the XML.
Comments
Comments require six cross-linked files. Use the helper — directory mode when you'll also be editing document.xml (saves an unzip/rezip cycle), .docx-direct mode otherwise:
# Against an already-unpacked directory (preferred when also placing markers)
python scripts/comment.py unpacked/ "Fees & expenses cap is too low"
python scripts/comment.py unpacked/ "Agreed" --parent 0
# Against a .docx directly
python scripts/comment.py contract.docx "This cap is too low" -o annotated.docx
The script writes comments.xml, commentsExtended.xml, commentsIds.xml, commentsExtensible.xml, the relationships, and the content-type overrides. Comment IDs are auto-assigned. It then prints the //`` snippet to add to word/document.xml so the comment anchors to specific text — until you place those markers, the comment exists but is not visible.
Dependencies
docx (npm, preinstalled — install only if require('docx') fails) · pandoc · LibreOffice (soffice) · pdftoppm (Poppler)
Questions people ask
- Can this edit existing Word documents?
- Yes — unzip the .docx, edit word/document.xml directly, then re-zip. The docx-js library cannot open existing files, so direct XML editing is the only approach for modifications.
- How do I verify the output looks correct?
- Convert to PDF using the provided LibreOffice script, then render each page as a JPEG for visual inspection. The tool shows you what the document actually looks like.
- What formatting quirks should I know about?
- Tables require both a table-level columnWidths and per-cell width in DXA units. Lists need a numbering config with LevelFormat.BULLET, not literal bullet characters. Table shading must use CLEAR, not SOLID. Page breaks must be inside a Paragraph. Never use \n — create separate Paragraph elements instead.
Related skills
Join a video meeting as an AI bot with voice, avatar, and screenshare across four operating modes.
Extract text, tables, and images from PDFs; merge, split, rotate, watermark, encrypt, and create new PDFs.
chatgpt-apps
OfficialScaffold ChatGPT Apps with MCP server and widget UI using a docs-first workflow.
figma-generate-design
OfficialBuild or update Figma screens using your design system's components, variables, and styles.
figma-generate-library
OfficialCreate a Figma design system that stays in sync with your code — variables, components, and theming built phase by phase with validation at every step.
More from Anthropic
Browse all skillsdoc-coauthoring
OfficialA structured workflow for writing better documents, from initial brief to reader-tested final draft.
Extract text, tables, and images from PDFs; merge, split, rotate, watermark, encrypt, and create new PDFs.
pptx
OfficialGenerate, edit, and extract content from PowerPoint files with a suite of validation and thumbnail tools.
xlsx
OfficialCreate, edit, and analyze spreadsheets with verified formulas and professional formatting.
skill-creator
OfficialCreate, test, and improve agent skills with structured evaluation and iteration.
canvas-design
OfficialGenerates original visual art as downloadable PNG and PDF files, grounded in a custom design philosophy.