Important Feedback/improvements for gpt-live-1 API

Hi akarsh, until append is reliable, here’s how I’d make sure a missed instruction can’t go unnoticed on a live call:

  1. Verify instead of assuming. After each appended “say X” instruction, check the output transcript of the next turn or two for a few key words of the disclosure. If they’re missing, you get a logged, countable failure instead of a silent one, and you can re-send or escalate.
  2. Take must-say content away from the model. For the greeting and legal disclosures, play fixed audio from your telephony layer before you connect the caller to the live session. It sounds the same every time and can’t be paraphrased or skipped.
  3. Measure the follow rate. Script 30–50 synthetic calls per scenario (greeting, mid-call instruction, disclosure after a tool result) and track how often the instruction was actually spoken. OpenAI’s voice agents guide recommends the same kind of fixed synthetic-speech runs. That number turns “not reliable” into something you can show OpenAI and re-test after each model update.

Adam