用文档优先的工作流脚手架 ChatGPT Apps,包含 MCP 服务器和组件代码。
智能体与编排
define-goal
试用将模糊的意图转化为可验证、可执行的具体目标
它能做什么
在开始工作前,引导你将意图转化为清晰的目标——明确指定具体成果、验证方式、成功阈值和范围边界。当目标过于模糊或停留在「活动描述」层面时,会改写为可量化或二元判断的形式。它会检查当前是否已有活跃目标,仅在必要时才创建新的目标。需要注意的是,它不生成中间计划、决策日志或持久化的执行快照。
什么时候用它
- 用户明确要求创建目标或使用目标工具时
- 「让X变快」这类模糊需求需要转化为可量化的指标
- 需要澄清什么样的证据能证明目标已完成
- 存在冲突或重复的目标,需要先理清再开始
技能文档
Define Goal
Overview
Shape the user's intent into an objective an agent can pursue honestly. Prefer measurable outcomes, explicit evidence, and bounded scope over activity descriptions.
This skill covers goal definition and goal-tool creation only. Do not create intermediate planning artifacts, durable snapshots, ledgers, decision logs, or resume files from this skill.
Workflow
-
Confirm that goal definition is actually needed.
- Use this skill when the user asks for
$define-goal, asks to create or set a goal, asks for the goal tool, or wants help turning an intention into a clear objective. - If the user only asks for ordinary implementation work, do the work directly instead of forcing goal creation.
- Use this skill when the user asks for
-
Restate the likely goal in concrete terms. A usable goal names:
- the specific outcome that will be true
- the main artifact, system, repo, environment, or user-facing behavior involved
- how completion will be verified
- what is in scope
- what is out of scope when ambiguity would matter
- the stop condition for asking the user instead of grinding
-
Make it quantitative when the domain supports it. Prefer numbers that represent real success, not decorative precision:
- pass/fail validators: exact tests, checks, CI jobs, evals, commands, or acceptance criteria
- quality thresholds: latency, error rate, cost, accuracy, recall, precision, coverage, flake rate, bundle size, memory, uptime, completion rate, or manual review criteria
- artifact constraints: file paths, affected modules, allowed commands, output formats, target environments, deadlines, or maximum blast radius
- evidence counts: number of reproduced failures, successful reruns, reviewed examples, migrated records, addressed comments, or verified cases
-
Repair weak goals before setting them.
- Rewrite vague goals into measurable objectives when local context makes the rewrite safe.
- Ask one concise clarification question when the missing detail changes the intended outcome or validation.
- Reject pure activity goals such as "make progress," "keep investigating," "improve things," or "work on X" unless they are sharpened into a verifiable outcome.
-
Check active goal state before creating a goal.
- Call
get_goal. - If there is no active goal and the objective meets the quality bar, call
create_goal. - If there is an active goal that still matches the user's intent, continue using it instead of creating a duplicate.
- If there is an active goal that conflicts with the new request, ask whether to finish the current goal, mark it complete if done, or start a separate goal-backed thread.
- Call
-
Create the goal only after it passes the quality bar.
- Use a single concise objective string.
- Include the verification evidence in the objective itself.
- Include scope bounds when they constrain the work.
- Include a token budget only when the user explicitly requested one.
- Do not call
create_goalfor an ordinary multi-step task unless the user explicitly asked for goal-backed work.
Goal Quality Bar
Before create_goal, the objective should answer:
- What concrete thing will be true when this is done?
- What evidence will prove it?
- What quantitative or binary threshold defines success?
- What scope boundaries matter?
- What should cause the agent to stop and ask?
Good:
Reduce checkout API p95 latency below 250 ms for the documented slow path by making the smallest safe server-side change, then verify with
npm run test:checkoutand the existing local latency benchmark showing p95 under 250 ms across 3 consecutive runs.
Good:
Resolve the open review comments on PR 123 that request code changes, update only the affected auth files and tests, and verify with the targeted auth test command plus
gh pr view 123showing no unresolved change-request threads.
Weak:
Make checkout faster.
Weak:
Keep investigating the PR comments.
Quantification Heuristics
- For bugs, define success as reproduction first, fix second, and a failing-then-passing validator when possible.
- For tests, name the exact command and required pass condition.
- For performance, name the metric, target threshold, measurement method, and number of runs.
- For quality work, define an observable acceptance bar such as reviewed examples, lint/typecheck/test pass, or user-approved artifact.
- For research, define the decision the research must enable, the sources or systems in scope, and the evidence standard.
- For operations, define healthy state, monitoring window, failure threshold, and rollback or escalation trigger.
Clarifying Questions
Ask only when a reasonable rewrite would risk pursuing the wrong outcome. Keep the question short and oriented around the missing validator or scope boundary.
Useful question shapes:
- "What metric should define success here: latency, cost, accuracy, or user-visible behavior?"
- "Which environment should I verify against: local, staging, or production?"
- "What is the minimum evidence you want before I mark this goal complete?"
If the user cannot provide a metric, propose the most honest binary validator available and ask for confirmation.
常见问题
- 这个工具具体做什么?
- 帮助定义和创建一个结构清晰的目标,包含具体目标描述、验证证据和范围边界。在创建目标前,它会将模糊的目标改写为可量化的形式。
- 它会跟踪进度或在时间跨度内管理目标吗?
- 不会。这个工具只负责目标定义和创建,不生成中间计划、决策日志或任何持久化的执行快照。
- 什么时候应该用这个工具,什么时候直接开始工作?
- 当你明确要求创建目标,或者想把一个模糊的意图转化为清晰的目标时使用。对于普通的实现任务,直接开始工作即可,不需要强制创建目标。
OpenAI 的更多技能
浏览全部技能从代码或描述生成完整的 Figma 页面,复用现有设计系统的组件和变量。
将代码中的设计令牌导入 Figma,自动构建变量体系和组件库。
通过 Figma Plugin API 执行 JavaScript,创建、编辑或自动化 Figma 文件中的元素。
从概念、品牌、图像参考生成 Codex 兼容的动画宠物精灵图集。
imagegen
官方根据提示词或参考图生成、编辑位图图像——照片、插画、UI 模型、精灵图、纹理等。