设计与多媒体

docx

试用

创建、读取、编辑和批注 Word .docx 文件,全面控制格式

它能做什么

使用 docx npm 库生成新的 .docx 和 .dotx 文件;通过 pandoc 读取现有文档内容;通过解压、修改内部 XML、重新打包的方式编辑现有文档。支持目录、标题层级、修订模式和内联批注。包含脚本用于对照 OOXML schema 验证输出,并可转换为 PDF 进行视觉检查。

什么时候用它

  • 从结构化数据生成带格式的报告或信笺 .docx
  • 编辑已有 .docx 修正措辞或更新数据,同时保留原有样式
  • 提取 .docx 中的文本内容供其他地方使用
  • 通过内联批注或接受/拒绝修订来审阅文档

技能文档

DOCX creation, editing, and analysis

A .docx is a ZIP archive of XML files. Choose your approach by task:

TaskApproach
Create a new documentWrite a docx (npm) script — see gotchas below
Edit an existing documentunzip → edit word/document.xml → zip (docx-js cannot open existing files)
Read contentpandoc -t markdown file.docx

Script paths below are relative to this skill's directory.

Creating with docx-js — gotchas

docx is preinstalled — do not run npm install first; write the script and require('docx') directly. Only if that require fails: npm install docx. The model knows the API; these are the footguns:

  • Page size defaults to A4. For US Letter set page: { size: { width: 12240, height: 15840 } } (DXA; 1440 = 1″).
  • Landscape: pass portrait dimensions and orientation: PageOrientation.LANDSCAPE — docx-js swaps width/height internally.
  • Tables need dual widths: set columnWidths on the table AND width on every cell, both in WidthType.DXA (PERCENTAGE breaks in Google Docs). Column widths must sum to the table width.
  • Table shading: use ShadingType.CLEAR, never SOLID (renders black).
  • Lists: never insert • literally; use a numbering config with LevelFormat.BULLET.
  • ImageRun requires type: ("png", "jpg", …).
  • PageBreak must be inside a Paragraph.
  • Never use \n — use separate Paragraph elements.
  • TOC: headings must use built-in HeadingLevel.*; custom heading styles need outlineLevel set or they won't appear.
  • Don't use a table as a horizontal rule — use a paragraph bottom border instead.
  • Dot-leader / right-aligned-on-same-line: use PositionalTab (alignment: PositionalTabAlignment.RIGHT, leader: PositionalTabLeader.DOT) inside a TextRun, not literal . or space padding.

Verify the output

After writing a .docx, render it and look at it:

python scripts/office/soffice.py --headless --convert-to pdf output.docx
pdftoppm -jpeg -r 100 output.pdf page
ls page-*.jpg   # then Read the images

pdftoppm zero-pads page numbers to the width of the page count (page-01.jpg…page-12.jpg).

Editing existing documents

Legacy .doc files must be converted first: python scripts/office/soffice.py --headless --convert-to docx file.doc.

unzip -q doc.docx -d unpacked/
find unpacked -type l -delete   # strip symlink entries — docx from external parties is untrusted
python scripts/merge_runs.py unpacked/   # coalesce fragmented runs so text is findable
# edit unpacked/word/document.xml in place — do NOT reformat or pretty-print
(cd unpacked && rm -f ../out.docx && zip -Xr ../out.docx .)
python scripts/office/validate.py out.docx --original doc.docx   # XSD checks; --auto-repair fixes common issues
# redlining? add --author "" to check every edit is tracked

Word splits text across many `` runs (revision ids, spell-check markers), so a phrase you can see in the document often doesn't exist as a contiguous string in the XML. merge_runs.py merges adjacent identically-formatted runs in word/document.xml without changing content or rendering; it also accepts a .docx directly (python scripts/merge_runs.py doc.docx -o merged.docx).

Tracked changes: when redlining, validate with --author "" (needs --original) — it reports any text you changed without a / around it, which is easy to do by accident and invisible in the accepted view. Wrap runs in / with w:id, w:author, w:date attributes. Inside , the text element is , not . A deleted paragraph mark () means "merge this paragraph into the next" — so deleting a paragraph outright is that plus a around every run. The must come before the rPr's other children; their order is schema-enforced.

To produce a clean copy with all tracked changes accepted: python scripts/accept_changes.py in.docx out.docx.

Accepting a deleted paragraph mark should join that paragraph to the one below it, so a paragraph whose runs are all deleted vanishes. Word does this; accept_changes.py and pandoc --track-changes=accept don't always. Both fail the same way — they strip the deleted text but leave the emptied paragraph behind, which reads as a stray empty bullet when it was auto-numbered:

  • pandoc --track-changes=accept never joins the paragraphs.
  • accept_changes.py (LibreOffice) joins them correctly, except when the deleted paragraph is followed by an empty spacer paragraph.

An empty bullet in either view is an artifact of that view, not a defect in the document. Check paragraph deletions in the XML.

Comments

Comments require six cross-linked files. Use the helper — directory mode when you'll also be editing document.xml (saves an unzip/rezip cycle), .docx-direct mode otherwise:

# Against an already-unpacked directory (preferred when also placing markers)
python scripts/comment.py unpacked/ "Fees & expenses cap is too low"
python scripts/comment.py unpacked/ "Agreed" --parent 0

# Against a .docx directly
python scripts/comment.py contract.docx "This cap is too low" -o annotated.docx

The script writes comments.xml, commentsExtended.xml, commentsIds.xml, commentsExtensible.xml, the relationships, and the content-type overrides. Comment IDs are auto-assigned. It then prints the //`` snippet to add to word/document.xml so the comment anchors to specific text — until you place those markers, the comment exists but is not visible.

Dependencies

docx (npm, preinstalled — install only if require('docx') fails) · pandoc · LibreOffice (soffice) · pdftoppm (Poppler)

常见问题

这个技能可以打开并编辑已有的 .docx 文件吗?
它不使用完整的 Word 对象模型,而是解压 .docx、通过辅助脚本合并碎片化的文本片段、直接编辑 word/document.xml、然后重新打包。旧的 .doc 文件需要先用 LibreOffice 转换为 .docx。
如何处理表格、页眉等复杂格式?
docx 库负责大多数格式化的程序化处理。表格需要在表格和每个单元格上同时指定宽度,底纹必须使用 CLEAR 类型以避免渲染为纯黑。输出包含 PDF 预览步骤,可在依赖结果前进行视觉验证。
修订模式和批注是如何处理的?
修订内容以 w:ins/w:del XML 元素呈现在文档正文中;验证脚本会标记出未在这类标签内进行的编辑。批注需要协调六个关联 XML 文件,由辅助脚本负责写入并打印出需要插入 document.xml 的锚点片段。

相关技能

以 AI 机器人身份加入视频会议,提供语音、虚拟形象与屏幕共享四种模式。

作者 johnpatternai21 次安装8 星标

pdf

官方

读写、合并、拆分、旋转、水印、加密和 OCR 处理 PDF 文件。

作者 Anthropic179.8k 星标

用文档优先的工作流脚手架 ChatGPT Apps,包含 MCP 服务器和组件代码。

作者 OpenAI27.9k 星标

从代码或描述生成完整的 Figma 页面,复用现有设计系统的组件和变量。

作者 OpenAI27.9k 星标

将代码中的设计令牌导入 Figma,自动构建变量体系和组件库。

作者 OpenAI27.9k 星标

Anthropic 的更多技能

浏览全部技能

通过结构化三阶段流程协作创建文档:收集上下文、分段迭代优化、读者测试验证。

作者 Anthropic179.8k 星标

pdf

官方

读写、合并、拆分、旋转、水印、加密和 OCR 处理 PDF 文件。

作者 Anthropic179.8k 星标

pptx

官方

创建、编辑和提取 .pptx / .potx 文件内容——幻灯片、模板、备注全面掌控

作者 Anthropic179.8k 星标

xlsx

官方

创建、编辑和分析电子表格文件,支持公式计算。

作者 Anthropic179.8k 星标

基于p5.js的生成式艺术引擎,支持种子随机性和交互式参数探索,每次运行生成独特的算法视觉作品。

作者 Anthropic179.8k 星标

通过结构化评估和迭代改进来构建和测试新的 AI 技能。

作者 Anthropic179.8k 星标