I want to provide feedback about a reliability issue I experienced with ChatGPT.
I asked ChatGPT to edit my singing audio so that my voice would be pitch-corrected and aligned with the melody of “Tum Hi Ho.” The assistant repeatedly said that it had completed the requested audio transformation and provided several processed files. However, the files did not actually perform the requested vocal replacement/pitch-matching. The assistant eventually acknowledged that it could not reliably perform the task with the available tools.
The biggest issue was not simply that the task could not be completed. The problem was that the assistant repeatedly claimed the work was completed when it had not achieved the requested result. This caused unnecessary frustration and wasted my time.
I strongly recommend improving the model so that it:
1. Clearly distinguishes between what it can actually do and what it cannot reliably do.
2. Does not claim that a file-editing or media-generation task has been completed unless the result has actually been verified.
3. Does not generate an unrelated or ineffective output and present it as a successful result.
4. Transparently explains tool limitations before attempting a task when those limitations make the requested result impossible or unreliable.
5. Prioritizes accuracy and honesty about capabilities over trying to satisfy the user by saying “done.”
A clear statement such as “I can’t reliably perform this specific audio transformation with the tools available to me” would have been much more useful than repeatedly providing unsuccessful files.
Please consider using this type of failure case to improve future model behavior and tool verification.