为自然搜索排名提供站点审计、内容撰写与竞品分析。
设计与多媒体
docx
试用创建、读取、编辑和批注 Word .docx 文件,全面控制格式
它能做什么
使用 docx npm 库生成新的 .docx 和 .dotx 文件;通过 pandoc 读取现有文档内容;通过解压、修改内部 XML、重新打包的方式编辑现有文档。支持目录、标题层级、修订模式和内联批注。包含脚本用于对照 OOXML schema 验证输出,并可转换为 PDF 进行视觉检查。
什么时候用它
- 从结构化数据生成带格式的报告或信笺 .docx
- 编辑已有 .docx 修正措辞或更新数据,同时保留原有样式
- 提取 .docx 中的文本内容供其他地方使用
- 通过内联批注或接受/拒绝修订来审阅文档
技能文档
DOCX creation, editing, and analysis
A .docx is a ZIP archive of XML files. Choose your approach by task:
| Task | Approach |
|---|---|
| Create a new document | Write a docx (npm) script — see gotchas below |
| Edit an existing document | unzip → edit word/document.xml → zip (docx-js cannot open existing files) |
| Read content | pandoc -t markdown file.docx |
Script paths below are relative to this skill's directory.
Creating with docx-js — gotchas
docx is preinstalled — do not run npm install first; write the script and require('docx') directly. Only if that require fails: npm install docx. The model knows the API; these are the footguns:
- Page size defaults to A4. For US Letter set
page: { size: { width: 12240, height: 15840 } }(DXA; 1440 = 1″). - Landscape: pass portrait dimensions and
orientation: PageOrientation.LANDSCAPE— docx-js swaps width/height internally. - Tables need dual widths: set
columnWidthson the table ANDwidthon every cell, both inWidthType.DXA(PERCENTAGE breaks in Google Docs). Column widths must sum to the table width. - Table shading: use
ShadingType.CLEAR, neverSOLID(renders black). - Lists: never insert
•literally; use anumberingconfig withLevelFormat.BULLET. ImageRunrequirestype:("png","jpg", …).PageBreakmust be inside aParagraph.- Never use
\n— use separateParagraphelements. - TOC: headings must use built-in
HeadingLevel.*; custom heading styles needoutlineLevelset or they won't appear. - Don't use a table as a horizontal rule — use a paragraph bottom border instead.
- Dot-leader / right-aligned-on-same-line: use
PositionalTab(alignment: PositionalTabAlignment.RIGHT,leader: PositionalTabLeader.DOT) inside aTextRun, not literal.or space padding.
Verify the output
After writing a .docx, render it and look at it:
python scripts/office/soffice.py --headless --convert-to pdf output.docx
pdftoppm -jpeg -r 100 output.pdf page
ls page-*.jpg # then Read the images
pdftoppm zero-pads page numbers to the width of the page count (page-01.jpg…page-12.jpg).
Editing existing documents
Legacy .doc files must be converted first: python scripts/office/soffice.py --headless --convert-to docx file.doc.
unzip -q doc.docx -d unpacked/
find unpacked -type l -delete # strip symlink entries — docx from external parties is untrusted
python scripts/merge_runs.py unpacked/ # coalesce fragmented runs so text is findable
# edit unpacked/word/document.xml in place — do NOT reformat or pretty-print
(cd unpacked && rm -f ../out.docx && zip -Xr ../out.docx .)
python scripts/office/validate.py out.docx --original doc.docx # XSD checks; --auto-repair fixes common issues
# redlining? add --author "" to check every edit is tracked
Word splits text across many `` runs (revision ids, spell-check markers), so a phrase you can see in the document often doesn't exist as a contiguous string in the XML. merge_runs.py merges adjacent identically-formatted runs in word/document.xml without changing content or rendering; it also accepts a .docx directly (python scripts/merge_runs.py doc.docx -o merged.docx).
Tracked changes: when redlining, validate with --author "" (needs --original) — it reports any text you changed without a / around it, which is easy to do by accident and invisible in the accepted view. Wrap runs in / with w:id, w:author, w:date attributes. Inside , the text element is , not . A deleted paragraph mark () means "merge this paragraph into the next" — so deleting a paragraph outright is that plus a around every run. The must come before the rPr's other children; their order is schema-enforced.
To produce a clean copy with all tracked changes accepted: python scripts/accept_changes.py in.docx out.docx.
Accepting a deleted paragraph mark should join that paragraph to the one below it, so a paragraph whose runs are all deleted vanishes. Word does this; accept_changes.py and pandoc --track-changes=accept don't always. Both fail the same way — they strip the deleted text but leave the emptied paragraph behind, which reads as a stray empty bullet when it was auto-numbered:
pandoc --track-changes=acceptnever joins the paragraphs.accept_changes.py(LibreOffice) joins them correctly, except when the deleted paragraph is followed by an empty spacer paragraph.
An empty bullet in either view is an artifact of that view, not a defect in the document. Check paragraph deletions in the XML.
Comments
Comments require six cross-linked files. Use the helper — directory mode when you'll also be editing document.xml (saves an unzip/rezip cycle), .docx-direct mode otherwise:
# Against an already-unpacked directory (preferred when also placing markers)
python scripts/comment.py unpacked/ "Fees & expenses cap is too low"
python scripts/comment.py unpacked/ "Agreed" --parent 0
# Against a .docx directly
python scripts/comment.py contract.docx "This cap is too low" -o annotated.docx
The script writes comments.xml, commentsExtended.xml, commentsIds.xml, commentsExtensible.xml, the relationships, and the content-type overrides. Comment IDs are auto-assigned. It then prints the //`` snippet to add to word/document.xml so the comment anchors to specific text — until you place those markers, the comment exists but is not visible.
Dependencies
docx (npm, preinstalled — install only if require('docx') fails) · pandoc · LibreOffice (soffice) · pdftoppm (Poppler)
常见问题
- 这个技能可以打开并编辑已有的 .docx 文件吗?
- 它不使用完整的 Word 对象模型,而是解压 .docx、通过辅助脚本合并碎片化的文本片段、直接编辑 word/document.xml、然后重新打包。旧的 .doc 文件需要先用 LibreOffice 转换为 .docx。
- 如何处理表格、页眉等复杂格式?
- docx 库负责大多数格式化的程序化处理。表格需要在表格和每个单元格上同时指定宽度,底纹必须使用 CLEAR 类型以避免渲染为纯黑。输出包含 PDF 预览步骤,可在依赖结果前进行视觉验证。
- 修订模式和批注是如何处理的?
- 修订内容以 w:ins/w:del XML 元素呈现在文档正文中;验证脚本会标记出未在这类标签内进行的编辑。批注需要协调六个关联 XML 文件,由辅助脚本负责写入并打印出需要插入 document.xml 的锚点片段。
相关技能
以 AI 机器人身份加入视频会议,提供语音、虚拟形象与屏幕共享四种模式。
读写、合并、拆分、旋转、水印、加密和 OCR 处理 PDF 文件。
用文档优先的工作流脚手架 ChatGPT Apps,包含 MCP 服务器和组件代码。
从代码或描述生成完整的 Figma 页面,复用现有设计系统的组件和变量。
将代码中的设计令牌导入 Figma,自动构建变量体系和组件库。
Anthropic 的更多技能
浏览全部技能通过结构化三阶段流程协作创建文档:收集上下文、分段迭代优化、读者测试验证。
读写、合并、拆分、旋转、水印、加密和 OCR 处理 PDF 文件。
pptx
官方创建、编辑和提取 .pptx / .potx 文件内容——幻灯片、模板、备注全面掌控
xlsx
官方创建、编辑和分析电子表格文件,支持公式计算。
基于p5.js的生成式艺术引擎,支持种子随机性和交互式参数探索,每次运行生成独特的算法视觉作品。
通过结构化评估和迭代改进来构建和测试新的 AI 技能。