docx
官方用脚本创建、读取和编辑 Word .docx 与 .dotx 文件。
设计与多媒体
Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in intervi
Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.
Transcribe audio using OpenAI, with optional speaker diarization when requested. Prefer the bundled CLI for deterministic, repeatable runs.
OPENAI_API_KEY is set. If missing, ask the user to set it locally (do not ask them to paste the key).transcribe_diarize.py CLI with sensible defaults (fast text transcription).output/transcribe/ when working in this repo.gpt-4o-mini-transcribe with --response-format text for fast transcription.--model gpt-4o-transcribe-diarize --response-format diarized_json.--chunking-strategy auto.gpt-4o-transcribe-diarize.output/transcribe// for evaluation runs.--out-dir for multiple files to avoid overwriting.Prefer uv for dependency management.
uv pip install openai
If uv is unavailable:
python3 -m pip install openai
OPENAI_API_KEY must be set for live API calls.export CODEX_HOME="${CODEX_HOME:-$HOME/.codex}"
export TRANSCRIBE_CLI="$CODEX_HOME/skills/transcribe/scripts/transcribe_diarize.py"
User-scoped skills install under $CODEX_HOME/skills (default: ~/.codex/skills).
Single file (fast text default):
python3 "$TRANSCRIBE_CLI" \
path/to/audio.wav \
--out transcript.txt
Diarization with known speakers (up to 4):
python3 "$TRANSCRIBE_CLI" \
meeting.m4a \
--model gpt-4o-transcribe-diarize \
--known-speaker "Alice=refs/alice.wav" \
--known-speaker "Bob=refs/bob.wav" \
--response-format diarized_json \
--out-dir output/transcribe/meeting
Plain text output (explicit):
python3 "$TRANSCRIBE_CLI" \
interview.mp3 \
--response-format text \
--out interview.txt
references/api.md: supported formats, limits, response formats, and known-speaker notes.用脚本创建、读取和编辑 Word .docx 与 .dotx 文件。
用文档优先的流程搭建 ChatGPT Apps SDK 项目,产出工具规划、MCP 服务端与 Widget 脚手架。
按正确顺序在 Figma 中搭建与代码对齐的完整设计系统,覆盖变量、组件与主题。
从文字描述、参考图或品牌线索生成 Codex 兼容的动画宠物与 8x9 雪碧图集。
用 C# 和 Windows App SDK 引导、搭建并验证 WinUI 3 桌面应用。
把 Figma 设计组件与代码组件通过 Code Connect 映射关联起来。
用文档优先的流程搭建 ChatGPT Apps SDK 项目,产出工具规划、MCP 服务端与 Widget 脚手架。
按正确顺序在 Figma 中搭建与代码对齐的完整设计系统,覆盖变量、组件与主题。
通过智能体在 Figma 文件中安全、增量地执行 Plugin API JavaScript。
从文字描述、参考图或品牌线索生成 Codex 兼容的动画宠物与 8x9 雪碧图集。
通过内置 imagegen 工具生成或编辑位图图像,CLI 兜底模式仅在用户明确要求时启用。
从 OpenAI 开发者文档获取带引用和来源路径的权威、实时答案。