# 8-12 Seconds Response Delay with OpenAI API Using Node.js and WhatsApp API

**URL:** <https://community.openai.com/t/8-12-seconds-response-delay-with-openai-api-using-node-js-and-whatsapp-api/1091074>\
**Category:** API\
**Tags:** api\
**Created:** [January 14, 2025, 8:58pm UTC](https://community.openai.com/t/8-12-seconds-response-delay-with-openai-api-using-node-js-and-whatsapp-api/1091074 "2025-01-14T20:58:36Z")\
**Posts on this page:** 4\
**Page:** 1

<div class="post-metadata">

**Author:** ![developer33](https://sea2.discourse-cdn.com/openai1/user_avatar/community.openai.com/developer33/32/511851_2.png) [@developer33](https://community.openai.com/u/developer33)\
**Post date:** [January 14, 2025, 8:58pm UTC](https://community.openai.com/t/8-12-seconds-response-delay-with-openai-api-using-node-js-and-whatsapp-api/1091074/1 "2025-01-14T20:58:36Z")

</div>

Hi community,

I’m working on an integration that uses the OpenAI Assistants API along with the WhatsApp API. The flow is as follows:

1. A user sends a query via WhatsApp.

2. The message is processed using Node.js, sent to the OpenAI API, and the response is returned to the user via WhatsApp.

The issue is that I’m experiencing significant delays during the “run” phase of the OpenAI API, with response times ranging between **8 and 12 seconds** , which is negatively impacting the user experience.

**Technical Details:**

• **Backend:** Node.js

• **Integration:** WhatsApp API

• **Problem:** The majority of the delay appears to happen during interaction with the OpenAI API.

**Questions:**

1. Is this response time normal for the “run” option in the OpenAI API?

2. Are there any configurations or best practices to reduce this delay?

3. Could this be related to server load on OpenAI’s side?

Any guidance or shared experiences would be greatly appreciated. Thanks for your help!

---

<div class="post-metadata">

**Author:** ![ollie2](https://sea2.discourse-cdn.com/openai1/user_avatar/community.openai.com/ollie2/32/518517_2.png) [@ollie2](https://community.openai.com/u/ollie2)\
**Post date:** [January 15, 2025, 3:12am UTC](https://community.openai.com/t/8-12-seconds-response-delay-with-openai-api-using-node-js-and-whatsapp-api/1091074/2 "2025-01-15T03:12:50Z")

</div>

Normal, unfortunately. See my post I made 2 days ago sharing graphical results of stress testing I did on assistants/thread model. Run time varied between 10-40 seconds… Not at all viable for a production app.

> [@Assistant/Thread Model Stress Test: Concerning Results \[See inside\]](https://community.openai.com/t/assistant-thread-model-stress-test-concerning-results-see-inside/1089033):
>
> [Results graphs at bottom, see below] I’m designing an app where I’m thinking of having one assistant and one thread per user. The assistant represents a chat bot that users of the app chat to. Each user has their own thread. One thing that’s important to me is low latency of the chat bot response, so I ran a test measuring the response time of the bot (the time taken of the run; the invocation of the assistant on a user’s thread that returns a text response) over time as the thread increases i…

---

<div class="post-metadata">

**Author:** ![jochenschultz](https://sea2.discourse-cdn.com/openai1/user_avatar/community.openai.com/jochenschultz/32/491030_2.png) [@jochenschultz](https://community.openai.com/u/jochenschultz)\
**Post date:** [January 15, 2025, 6:41am UTC](https://community.openai.com/t/8-12-seconds-response-delay-with-openai-api-using-node-js-and-whatsapp-api/1091074/3 "2025-01-15T06:41:55Z")

</div>

And it was kind of worse some 1-2 years ago but hasn’t really improved in terms of response time. Like I wrote in that same thread. Azure deployment of the same (openAI) models has way better performance. It might be the downside of the cooperation between openAI and microsoft or you need to shift to enterprise of openAI to experience the real speed. I mean I see applications that answer instantly and they clearly use OpenAI models.

Has nothing to do with the consuming programming language / environment though… If you experience differences between python or nodejs or even PHP implementations and they are your problem than congratulation on having great API response time…

---

<div class="post-metadata">

**Author:** ![curt.kennedy](https://sea2.discourse-cdn.com/openai1/user_avatar/community.openai.com/curt.kennedy/32/709249_2.png) [@curt.kennedy](https://community.openai.com/u/curt.kennedy)\
**Post date:** [December 29, 2025, 4:43am UTC](https://community.openai.com/t/8-12-seconds-response-delay-with-openai-api-using-node-js-and-whatsapp-api/1091074/4 "2025-12-29T04:43:49Z")

</div>


