chatgpt-apps
OfficialScaffold ChatGPT Apps SDK projects from docs-first plans into MCP server + widget code.
Design & media
Generate spoken audio from text using OpenAI's gpt-4o-mini-tts model with built-in voices.
Produces narration, voiceover, IVR prompts, or accessibility audio from text via the OpenAI Audio API. Ships with a bundled CLI (`scripts/text_to_speech.py`) that handles single clips and JSONL-driven batches with capped rate limits. Defaults to model `gpt-4o-mini-tts-2025-12-15` and built-in voices (e.g., `cedar`, `marin`); custom voice cloning is not supported. Requires `OPENAI_API_KEY`; text is capped at 4096 characters per request.
Generate spoken audio for the current project (narration, product demo voiceover, IVR prompts, accessibility reads). Defaults to gpt-4o-mini-tts-2025-12-15 and built-in voices, and prefers the bundled CLI for deterministic, reproducible runs.
scripts/text_to_speech.py) with sensible defaults (see references/cli.md).tmp/speech/ for intermediate files (for example JSONL batches); delete when done.output/speech/ when working in this repo.--out or --out-dir to control output paths; keep filenames stable and descriptive.Prefer uv for dependency management.
Python packages:
uv pip install openai
If uv is unavailable:
python3 -m pip install openai
OPENAI_API_KEY must be set for live API calls.If the key is missing, give the user these steps:
OPENAI_API_KEY as an environment variable in their system.If installation isn't possible in this environment, tell the user which dependency is missing and how to install it locally.
gpt-4o-mini-tts-2025-12-15 unless the user requests another model.cedar. If the user wants a brighter tone, prefer marin.instructions are supported for GPT-4o mini TTS models, but not for tts-1 or tts-1-hd.--rpm at 50.OPENAI_API_KEY before any live API call.openai package) for all API calls; do not use raw HTTP.scripts/text_to_speech.py) over writing new one-off scripts.scripts/text_to_speech.py. If something is missing, ask the user before doing anything else.Reformat user direction into a short, labeled spec. Only make implicit details explicit; do not invent new requirements.
Quick clarification (augmentation vs invention):
Template (include only relevant lines):
Voice Affect:
Tone:
Pacing:
Emotion:
Pronunciation:
Pauses:
Emphasis:
Delivery:
Augmentation rules:
Input text: "Welcome to the demo. Today we'll show how it works."
Instructions:
Voice Affect: Warm and composed.
Tone: Friendly and confident.
Pacing: Steady and moderate.
Emphasis: Stress "demo" and "show".
{"input":"Thank you for calling. Please hold.","voice":"cedar","response_format":"mp3","out":"hold.mp3"}
{"input":"For sales, press 1. For support, press 2.","voice":"marin","instructions":"Tone: Clear and neutral. Pacing: Slow.","response_format":"wav"}
More principles: references/prompting.md. Copy/paste specs: references/sample-prompts.md.
Use these modules when the request is for a specific delivery style. They provide targeted defaults and templates.
references/narration.mdreferences/voiceover.mdreferences/ivr.mdreferences/accessibility.mdreferences/cli.mdreferences/audio-api.mdreferences/voice-directions.mdreferences/codex-network.mdreferences/cli.md: how to run speech generation/batches via scripts/text_to_speech.py (commands, flags, recipes).references/audio-api.md: API parameters, limits, voice list.references/voice-directions.md: instruction patterns and examples.references/prompting.md: instruction best practices (structure, constraints, iteration patterns).references/sample-prompts.md: copy/paste instruction recipes (examples only; no extra theory).references/narration.md: templates + defaults for narration and explainers.references/voiceover.md: templates + defaults for product demo voiceovers.references/ivr.md: templates + defaults for IVR/phone prompts.references/accessibility.md: templates + defaults for accessibility reads.references/codex-network.md: environment/sandbox/network-approval troubleshooting.Scaffold ChatGPT Apps SDK projects from docs-first plans into MCP server + widget code.
Build full Figma screens from code or descriptions using the target file's published design system.
Create Codex-compatible animated pets and pet spritesheets from a concept, brand, or reference art.
Get a repository-specific AppSec threat model with concrete abuse paths and mitigations, written to a Markdown file.
Generate project-specific design system rules for AI coding agents implementing Figma designs.
Build and evaluate MCP servers that expose external APIs to LLMs through well-designed tools.
Scaffold ChatGPT Apps SDK projects from docs-first plans into MCP server + widget code.
Build full Figma screens from code or descriptions using the target file's published design system.
Reference skill for executing JavaScript in Figma files via the `use_figma` MCP tool, defining API rules and patterns for plugin code.
Create Codex-compatible animated pets and pet spritesheets from a concept, brand, or reference art.
Generate or edit bitmap images directly inside a Codex project, with workspace-bound save paths.
Authoritative OpenAI developer-doc answers with citations for coding and API questions.