docx
官方用脚本创建、读取和编辑 Word .docx 与 .dotx 文件。
设计与多媒体
通过内置 CLI 调用 OpenAI 音频接口,把文本转成语音。
使用 OpenAI 的 GPT-4o mini TTS 模型和内置语音,从文本生成单条语音片段或批量音频文件。统一通过打包好的 CLI(`scripts/text_to_speech.py`)运行,保证结果可复现;支持通过 instruction 控制语调、节奏、强调等表达方式。中间文件写入 `tmp/speech/`,成品输出到 `output/speech/`。依赖 `openai` Python 包和 `OPENAI_API_KEY` 环境变量;不支持自定义声音克隆。
Generate spoken audio for the current project (narration, product demo voiceover, IVR prompts, accessibility reads). Defaults to gpt-4o-mini-tts-2025-12-15 and built-in voices, and prefers the bundled CLI for deterministic, reproducible runs.
scripts/text_to_speech.py) with sensible defaults (see references/cli.md).tmp/speech/ for intermediate files (for example JSONL batches); delete when done.output/speech/ when working in this repo.--out or --out-dir to control output paths; keep filenames stable and descriptive.Prefer uv for dependency management.
Python packages:
uv pip install openai
If uv is unavailable:
python3 -m pip install openai
OPENAI_API_KEY must be set for live API calls.If the key is missing, give the user these steps:
OPENAI_API_KEY as an environment variable in their system.If installation isn't possible in this environment, tell the user which dependency is missing and how to install it locally.
gpt-4o-mini-tts-2025-12-15 unless the user requests another model.cedar. If the user wants a brighter tone, prefer marin.instructions are supported for GPT-4o mini TTS models, but not for tts-1 or tts-1-hd.--rpm at 50.OPENAI_API_KEY before any live API call.openai package) for all API calls; do not use raw HTTP.scripts/text_to_speech.py) over writing new one-off scripts.scripts/text_to_speech.py. If something is missing, ask the user before doing anything else.Reformat user direction into a short, labeled spec. Only make implicit details explicit; do not invent new requirements.
Quick clarification (augmentation vs invention):
Template (include only relevant lines):
Voice Affect:
Tone:
Pacing:
Emotion:
Pronunciation:
Pauses:
Emphasis:
Delivery:
Augmentation rules:
Input text: "Welcome to the demo. Today we'll show how it works."
Instructions:
Voice Affect: Warm and composed.
Tone: Friendly and confident.
Pacing: Steady and moderate.
Emphasis: Stress "demo" and "show".
{"input":"Thank you for calling. Please hold.","voice":"cedar","response_format":"mp3","out":"hold.mp3"}
{"input":"For sales, press 1. For support, press 2.","voice":"marin","instructions":"Tone: Clear and neutral. Pacing: Slow.","response_format":"wav"}
More principles: references/prompting.md. Copy/paste specs: references/sample-prompts.md.
Use these modules when the request is for a specific delivery style. They provide targeted defaults and templates.
references/narration.mdreferences/voiceover.mdreferences/ivr.mdreferences/accessibility.mdreferences/cli.mdreferences/audio-api.mdreferences/voice-directions.mdreferences/codex-network.mdreferences/cli.md: how to run speech generation/batches via scripts/text_to_speech.py (commands, flags, recipes).references/audio-api.md: API parameters, limits, voice list.references/voice-directions.md: instruction patterns and examples.references/prompting.md: instruction best practices (structure, constraints, iteration patterns).references/sample-prompts.md: copy/paste instruction recipes (examples only; no extra theory).references/narration.md: templates + defaults for narration and explainers.references/voiceover.md: templates + defaults for product demo voiceovers.references/ivr.md: templates + defaults for IVR/phone prompts.references/accessibility.md: templates + defaults for accessibility reads.references/codex-network.md: environment/sandbox/network-approval troubleshooting.用脚本创建、读取和编辑 Word .docx 与 .dotx 文件。
用文档优先的流程搭建 ChatGPT Apps SDK 项目,产出工具规划、MCP 服务端与 Widget 脚手架。
按正确顺序在 Figma 中搭建与代码对齐的完整设计系统,覆盖变量、组件与主题。
从文字描述、参考图或品牌线索生成 Codex 兼容的动画宠物与 8x9 雪碧图集。
用 C# 和 Windows App SDK 引导、搭建并验证 WinUI 3 桌面应用。
把 Figma 设计组件与代码组件通过 Code Connect 映射关联起来。
用文档优先的流程搭建 ChatGPT Apps SDK 项目,产出工具规划、MCP 服务端与 Widget 脚手架。
按正确顺序在 Figma 中搭建与代码对齐的完整设计系统,覆盖变量、组件与主题。
通过智能体在 Figma 文件中安全、增量地执行 Plugin API JavaScript。
从文字描述、参考图或品牌线索生成 Codex 兼容的动画宠物与 8x9 雪碧图集。
通过内置 imagegen 工具生成或编辑位图图像,CLI 兜底模式仅在用户明确要求时启用。
从 OpenAI 开发者文档获取带引用和来源路径的权威、实时答案。