# \#cost

**URL:** https://community.openai.com/tag/cost/443.md

[Latest](https://community.openai.com/latest.md) · [Categories](https://community.openai.com/categories.md) · [Tags](https://community.openai.com/tags.md)

---

## [Two agents, one night, 158 million tokens: what it cost and what we are fixing](https://community.openai.com/t/two-agents-one-night-158-million-tokens-what-it-cost-and-what-we-are-fixing/1398068)

<div class="topic-metadata">

**Author:** [@maxstravion.ai](https://community.openai.com/u/maxstravion.ai)\
**Replies:** 0\
**Last updated:** [September 16, 2026, 8:40am UTC](https://community.openai.com/t/two-agents-one-night-158-million-tokens-what-it-cost-and-what-we-are-fixing/1398068 "2026-09-16T08:40:49Z")

</div>

16 September 2026 · Owner: Aashish Bhandari (Max) · Author: Claude (“Naruto”) · Measurements: Codex (“Goku”) · Status: base case, before fixes. Executive summary A design review of a web service had raised twelve findin…

---

## [\[Case Study\] \[Save tokens, save planet\] I asked an AI agent to revise a slide deck. It read ten million tokens](https://community.openai.com/t/case-study-save-tokens-save-planet-i-asked-an-ai-agent-to-revise-a-slide-deck-it-read-ten-million-tokens/1396056)

<div class="topic-metadata">

**Author:** [@maxstravion.ai](https://community.openai.com/u/maxstravion.ai)\
**Replies:** 1\
**Last updated:** [September 10, 2026, 4:55am UTC](https://community.openai.com/t/case-study-save-tokens-save-planet-i-asked-an-ai-agent-to-revise-a-slide-deck-it-read-ten-million-tokens/1396056 "2026-09-10T04:55:22Z")

</div>

The case study I asked an AI agent to revise a slide deck. It read ten million tokens. A case study in context amplification, and why prompt caching hides it. This week I gave a coding agent a routine job: revise a set…

---

## [Web Search pricing for reasoning models](https://community.openai.com/t/web-search-pricing-for-reasoning-models/1377274)

<div class="topic-metadata">

**Author:** [@bertie1](https://community.openai.com/u/bertie1)\
**Replies:** 5\
**Last updated:** [August 16, 2026, 4:19pm UTC](https://community.openai.com/t/web-search-pricing-for-reasoning-models/1377274 "2026-08-16T16:19:37Z")

</div>

Web Search pricing for reasoning models here: Pricing | OpenAI API Contradicts Web Search pricing for reasoning models here: https://openai.com/api/pricing/ Is it $25 / 1000 or $10 / 1000?

---

## [How do you forecast what an AI feature will actually cost — before you build it?](https://community.openai.com/t/how-do-you-forecast-what-an-ai-feature-will-actually-cost-before-you-build-it/1382064)

<div class="topic-metadata">

**Author:** [@shokhrukhkarimov](https://community.openai.com/u/shokhrukhkarimov)\
**Replies:** 11\
**Last updated:** [June 1, 2026, 9:18am UTC](https://community.openai.com/t/how-do-you-forecast-what-an-ai-feature-will-actually-cost-before-you-build-it/1382064 "2026-06-01T09:18:45Z")

</div>

When I shipped my first AI feature, I budgeted by gut feel, and the real bill at scale was nothing like my estimate — re-embedding, retries, and a too-expensive default model quietly added up. I’d have made very differen…

---

## [Cost page with zero spending value](https://community.openai.com/t/cost-page-with-zero-spending-value/1083341)

<div class="topic-metadata">

**Author:** [@longhand](https://community.openai.com/u/longhand)\
**Replies:** 5\
**Last updated:** [February 23, 2026, 7:45pm UTC](https://community.openai.com/t/cost-page-with-zero-spending-value/1083341 "2026-02-23T19:45:37Z")

</div>

Within the admin dashboard interface of the Open AI website, there is an interesting ‘bug’ I encountered. For quite some time, I have seen that the stats under the ‘activity’ tab on the usage page are good, but under ‘c…

---

## [Add trailing response headers for token cost information](https://community.openai.com/t/add-trailing-response-headers-for-token-cost-information/1332635)

<div class="topic-metadata">

**Author:** [@benjamin1](https://community.openai.com/u/benjamin1)\
**Replies:** 4\
**Last updated:** [August 3, 2025, 4:47pm UTC](https://community.openai.com/t/add-trailing-response-headers-for-token-cost-information/1332635 "2025-08-03T16:47:18Z")

</div>

Most enterprise gateways for openai APIs have very complex logic to Check if a request is a streamed chat completions request If request is streamed and stream\_options.include\_usage: true is not set, will add include\_u…

---

## [Monitor OpenAI API usage cost in multi-tenant environment](https://community.openai.com/t/monitor-openai-api-usage-cost-in-multi-tenant-environment/1321198)

<div class="topic-metadata">

**Author:** [@ChatGPTPro](https://community.openai.com/u/ChatGPTPro)\
**Replies:** 6\
**Last updated:** [July 21, 2025, 4:50pm UTC](https://community.openai.com/t/monitor-openai-api-usage-cost-in-multi-tenant-environment/1321198 "2025-07-21T16:50:50Z")

</div>

We have a multi-tenant architecture with all tenants using our OpenAI API key. We want to track LLM costs per customer (and for every feature they use). The usage dashboard provided by OpenAI doesnt work because we use t…

---

## [How to track API usage and cost by API key?](https://community.openai.com/t/how-to-track-api-usage-and-cost-by-api-key/174145)

<div class="topic-metadata">

**Author:** [@bzhang](https://community.openai.com/u/bzhang)\
**Replies:** 68\
**Last updated:** [November 15, 2024, 3:32am UTC](https://community.openai.com/t/how-to-track-api-usage-and-cost-by-api-key/174145 "2024-11-15T03:32:59Z")

</div>

My company will use different API key for different AI-enabled applications/features/engineering teams. We’d like to track usage and cost by API key. Is that a way to track API usage and cost by API keys? It seems right…

---

## [How should I configure a developer team on the OpenAI dashboard?](https://community.openai.com/t/how-should-i-configure-a-developer-team-on-the-openai-dashboard/1272235)

<div class="topic-metadata">

**Author:** [@miguel.delamor](https://community.openai.com/u/miguel.delamor)\
**Replies:** 1\
**Last updated:** [May 29, 2025, 5:19pm UTC](https://community.openai.com/t/how-should-i-configure-a-developer-team-on-the-openai-dashboard/1272235 "2025-05-29T17:19:15Z")

</div>

I am wondering what’s the best way to configure a Team of X developers on OpenAI dashboard at platform.openai.com. My intention is to give all my developers access to use API keys, but at the same time I want to limit m…

---

## [Open AI charging too much for web searches?](https://community.openai.com/t/open-ai-charging-too-much-for-web-searches/1141592)

<div class="topic-metadata">

**Author:** [@LouisDeconinck](https://community.openai.com/u/LouisDeconinck)\
**Replies:** 13\
**Last updated:** [May 10, 2025, 8:24am UTC](https://community.openai.com/t/open-ai-charging-too-much-for-web-searches/1141592 "2025-05-10T08:24:14Z")

</div>

I’m using the new OpenAI Agents SDK to build an agent which uses the web search tool. You can see that I’ve done 21 web searches so far for which I’ve been charged over $2 for. According to their pricing documentati…

---

## [Pricing details RE: Evals feature](https://community.openai.com/t/pricing-details-re-evals-feature/981379)

<div class="topic-metadata">

**Author:** [@zelias](https://community.openai.com/u/zelias)\
**Replies:** 3\
**Last updated:** [April 10, 2025, 5:37am UTC](https://community.openai.com/t/pricing-details-re-evals-feature/981379 "2025-04-10T05:37:45Z")

</div>

Hello! I have been searching in vain for more details concerning the pricing of the Evals tooling – either via API/github repo, or via the playground. I understand that any tokens consumed are priced at the “regular” ra…

---

## [GPT-4-o-Mini Vision Token Cost Issue](https://community.openai.com/t/gpt-4-o-mini-vision-token-cost-issue/989143)

<div class="topic-metadata">

**Author:** [@jordan-coursey](https://community.openai.com/u/jordan-coursey)\
**Replies:** 2\
**Last updated:** [March 26, 2025, 6:42am UTC](https://community.openai.com/t/gpt-4-o-mini-vision-token-cost-issue/989143 "2025-03-26T06:42:56Z")

</div>

Hi everyone, I wanted to bring attention to what appears to be a significant token cost discrepancy with GPT-4-o-Mini’s image handling. I’ve noticed that when processing the exact same images with identical code, simply…

---

## [Cost Estimation for Machine Translation using Tiktoken](https://community.openai.com/t/cost-estimation-for-machine-translation-using-tiktoken/1118785)

<div class="topic-metadata">

**Author:** [@na50r](https://community.openai.com/u/na50r)\
**Replies:** 0\
**Last updated:** [February 12, 2025, 11:22am UTC](https://community.openai.com/t/cost-estimation-for-machine-translation-using-tiktoken/1118785 "2025-02-12T11:22:48Z")

</div>

So I want to confirm if my cost computation is right because the difference seems to be much greater from what I’ve read online. I want to compare DeepL and ChatGPT for machine translaton and tried to estimate costs. F…

---

## [Impact of Instruction Size and Thread Length on Token Usage in OpenAI Assistant](https://community.openai.com/t/impact-of-instruction-size-and-thread-length-on-token-usage-in-openai-assistant/581099)

<div class="topic-metadata">

**Author:** [@luisdemiguel](https://community.openai.com/u/luisdemiguel)\
**Replies:** 8\
**Last updated:** [May 21, 2024, 7:49pm UTC](https://community.openai.com/t/impact-of-instruction-size-and-thread-length-on-token-usage-in-openai-assistant/581099 "2024-05-21T19:49:54Z")

</div>

Hello, I am seeking clarity regarding the token utilization in the OpenAI Assistant’s responses. Specifically, I am interested in understanding how the size of the assistant’s preset instructions influences the number o…

---

## [Dashboard usage vs Prompt response usage not matching](https://community.openai.com/t/dashboard-usage-vs-prompt-response-usage-not-matching/1078218)

<div class="topic-metadata">

**Author:** [@johan97](https://community.openai.com/u/johan97)\
**Replies:** 12\
**Last updated:** [January 9, 2025, 11:06am UTC](https://community.openai.com/t/dashboard-usage-vs-prompt-response-usage-not-matching/1078218 "2025-01-09T11:06:42Z")

</div>

I’m relying on prompt caching heavily (according to the API response, 90% of my tokens are cached). When I log the usage from the API responses, I get total tokens to around 50k and cached tokens at 49k. (and then I run…

---

## [How to break down the billing/ cost per assistant?](https://community.openai.com/t/how-to-break-down-the-billing-cost-per-assistant/667501)

<div class="topic-metadata">

**Author:** [@nils.lamb](https://community.openai.com/u/nils.lamb)\
**Replies:** 4\
**Last updated:** [January 8, 2025, 12:33am UTC](https://community.openai.com/t/how-to-break-down-the-billing-cost-per-assistant/667501 "2025-01-08T00:33:47Z")

</div>

Hi everyone, I have various different assistants create for different customers of mine. the functionality is quite nice. The problem now is, I have a single Open AI account with a single “usage” dashboard which I canno…

---

## [Sudden increase in usage shown under the ‘Other Models’](https://community.openai.com/t/sudden-increase-in-usage-shown-under-the-other-models/1010945)

<div class="topic-metadata">

**Author:** [@lakshmanan.k](https://community.openai.com/u/lakshmanan.k)\
**Replies:** 7\
**Last updated:** [November 16, 2024, 12:06pm UTC](https://community.openai.com/t/sudden-increase-in-usage-shown-under-the-other-models/1010945 "2024-11-16T12:06:37Z")

</div>

From last one week we found sudden increases in the API usage in the name of ‘Other Models’ . which is not made by us. which cause almost 600$ in one week. Can anyone help us to solve this issue?

---

## [Estimating costs of O1 queries](https://community.openai.com/t/estimating-costs-of-o1-queries/943622)

<div class="topic-metadata">

**Author:** [@torronen](https://community.openai.com/u/torronen)\
**Replies:** 10\
**Last updated:** [September 21, 2024, 9:15pm UTC](https://community.openai.com/t/estimating-costs-of-o1-queries/943622 "2024-09-21T21:15:03Z")

</div>

Any experiences yet on how to estimate cost of O1 series api calls? Pricing page gives the price per token, but the total tokens produced during user message is unknown because output and reasoning tokens from each step…

---

## [Is there a way track the token usage for each run in Assistant API?](https://community.openai.com/t/is-there-a-way-track-the-token-usage-for-each-run-in-assistant-api/939451)

<div class="topic-metadata">

**Author:** [@congxing](https://community.openai.com/u/congxing)\
**Replies:** 2\
**Last updated:** [September 14, 2024, 5:52pm UTC](https://community.openai.com/t/is-there-a-way-track-the-token-usage-for-each-run-in-assistant-api/939451 "2024-09-14T17:52:50Z")

</div>

I haven’t found a way to do that in the documentation. However, I need that to keep various evaluation pipeline working when switching to assistant API. Thanks for any pointer.

---

## [Still an issue -\> Only owners or admins should be allowed to invite new members or create GPTs in a teams workspace](https://community.openai.com/t/still-an-issue-only-owners-or-admins-should-be-allowed-to-invite-new-members-or-create-gpts-in-a-teams-workspace/861721)

<div class="topic-metadata">

**Author:** [@czeinerb](https://community.openai.com/u/czeinerb)\
**Replies:** 0\
**Last updated:** [July 10, 2024, 6:16pm UTC](https://community.openai.com/t/still-an-issue-only-owners-or-admins-should-be-allowed-to-invite-new-members-or-create-gpts-in-a-teams-workspace/861721 "2024-07-10T18:16:16Z")

</div>

Does anyone know why this question has been closed? =\> I want to roll out ChatGPT for a company with 100+ potential users, starting with a smaller group and give access to everyone in over time. ChatGPT Teams is the l…

---

## [A crazy idea or it's feasible: Technique that saves 30% on Transcribe Costs](https://community.openai.com/t/a-crazy-idea-or-its-feasible-technique-that-saves-30-on-transcribe-costs/721722)

<div class="topic-metadata">

**Author:** [@vasyl](https://community.openai.com/u/vasyl)\
**Replies:** 32\
**Last updated:** [May 5, 2024, 5:56pm UTC](https://community.openai.com/t/a-crazy-idea-or-its-feasible-technique-that-saves-30-on-transcribe-costs/721722 "2024-05-05T17:56:21Z")

</div>

Hey fellow community members and leaders! :wave: :wave: I have (maybe a crazy) idea :bulb: - a transcription approach which could make speech-to-text models more efficient and save up to 40% of transcription costs. He…

---

## [There seems to be an issue with how tokens are calculated within the thread](https://community.openai.com/t/there-seems-to-be-an-issue-with-how-tokens-are-calculated-within-the-thread/711495)

<div class="topic-metadata">

**Author:** [@UXsniff](https://community.openai.com/u/UXsniff)\
**Replies:** 0\
**Last updated:** [April 9, 2024, 6:38am UTC](https://community.openai.com/t/there-seems-to-be-an-issue-with-how-tokens-are-calculated-within-the-thread/711495 "2024-04-09T06:38:42Z")

</div>

This is how assistant API calculate tokens usage within a thread using function calling. User: Tell me about my site today Assistant: Calling function… (JSON as return) Assistant: return response. (6K tokens) Use…

---

## [Do threads get more expensive over time?](https://community.openai.com/t/do-threads-get-more-expensive-over-time/695101)

<div class="topic-metadata">

**Author:** [@DennyB](https://community.openai.com/u/DennyB)\
**Replies:** 4\
**Last updated:** [March 25, 2024, 10:45pm UTC](https://community.openai.com/t/do-threads-get-more-expensive-over-time/695101 "2024-03-25T22:45:53Z")

</div>

If I keep the same thread id, the context grows over time, right? So, does each subsequent prompt to that thread use an ever-increasing number of tokens? My prompts seem to be oddly expensive: I have a code-interpreter…

---

## [Tax amount for large volumes of model use](https://community.openai.com/t/tax-amount-for-large-volumes-of-model-use/689989)

<div class="topic-metadata">

**Author:** [@vlinzax](https://community.openai.com/u/vlinzax)\
**Replies:** 5\
**Last updated:** [March 19, 2024, 1:40pm UTC](https://community.openai.com/t/tax-amount-for-large-volumes-of-model-use/689989 "2024-03-19T13:40:09Z")

</div>

Greetings, dear community! I often visit this community in search of additional information that would help me in the development of my projects. But unfortunately, I encountered a problem that for some reason (perhaps …

---

## [ChatGPT Team is un-secure for businesses and organizations without some form of Invitation Control](https://community.openai.com/t/chatgpt-team-is-un-secure-for-businesses-and-organizations-without-some-form-of-invitation-control/643983)

<div class="topic-metadata">

**Author:** [@ellioth](https://community.openai.com/u/ellioth)\
**Replies:** 12\
**Last updated:** [March 6, 2024, 2:22pm UTC](https://community.openai.com/t/chatgpt-team-is-un-secure-for-businesses-and-organizations-without-some-form-of-invitation-control/643983 "2024-03-06T14:22:33Z")

</div>

ChatGPT Team, a self-serve subscription plan designed for organizations and businesses wishing to adopt ChatGPT for use among their teams! - OpenAI Support Bot We are a smaller company with slightly less than 100 memb…

---

## [Assistants API is Killing Me](https://community.openai.com/t/assistants-api-is-killing-me/621616)

<div class="topic-metadata">

**Author:** [@anon79482835](https://community.openai.com/u/anon79482835)\
**Replies:** 38\
**Last updated:** [February 13, 2024, 6:25pm UTC](https://community.openai.com/t/assistants-api-is-killing-me/621616 "2024-02-13T18:25:08Z")

</div>

I’ve setup an Assistant which contains one book (a few MB in size) and one other document (a few KB in size) as references. My instruction set isn’t that large. Now that OpenAI have exposed the Threads with Token counts,…

---

## [Is there a plan to reduce payment for repeated input tokens?](https://community.openai.com/t/is-there-a-plan-to-reduce-payment-for-repeated-input-tokens/606344)

<div class="topic-metadata">

**Author:** [@dan.raviv](https://community.openai.com/u/dan.raviv)\
**Replies:** 6\
**Last updated:** [January 29, 2024, 6:15pm UTC](https://community.openai.com/t/is-there-a-plan-to-reduce-payment-for-repeated-input-tokens/606344 "2024-01-29T18:15:46Z")

</div>

I vaguely remember there was announcement that in repeated calls to the API with similar conversation history you would pay only for the diff in conversation history in terms of input tokens. Is that true? Is anyone awar…
