Direct answer: the AI cost per task is the number of tokens one job uses, multiplied by the token price of that model. The formula: (input tokens × input price + output tokens × output price) ÷ 1,000,000. One customer service reply of 1,200 input tokens and 250 output tokens costs Rp 3.86 to Rp 430, depending on the model you pick.
Updated 11 September 2026. Every price here comes from the provider's own page on that date. The exchange rate I use: 1 USD = Rp 17,561.81, from open.er-api.com at 00:02 UTC on 11 September 2026. For comparison, the European Central Bank reference rate on 10 September 2026 was 1 USD = Rp 17,574.
Limit: model prices change without a long notice, and the rate moves daily. Recalculate before you sign a contract that uses these numbers.
The AI cost per task formula
An AI model charges per token, not per question. A token is a piece of a word. One English word is roughly 1 to 2 tokens, so 1,000 tokens is around 700 words.
Prices are always written per 1 million tokens, and the input price differs from the output price. Output is almost always dearer, usually 4 to 6 times the input price.
Cost of 1 task = (input tokens × input price + output tokens × output price) ÷ 1,000,000
Monthly cost = cost of 1 task × task count × retry factor + fixed cost
An example with Claude Haiku 4.5, priced at 1 USD input and 5 USD output per 1 million tokens:
(1,200 × 1 + 250 × 5) ÷ 1,000,000 = 0.00245 USD
0.00245 × 17,561.81 = Rp 43.03 per task

API prices per 1 million tokens, 11 September 2026
This table holds the standard price with no discount, short context. The rupiah column is my own calculation at the rate above, not an official rupiah price.
| Model | Input (USD) | Output (USD) | Cost of 1 chat reply |
|---|---|---|---|
| Gemini 2.5 Flash-Lite | 0.10 | 0.40 | Rp 3.86 |
| GPT-5.6 Luna | 0.20 | 1.20 | Rp 9.48 |
| Gemini 3.8 Flash | 0.75 | 3.75 | Rp 32.27 |
| Claude Haiku 4.5 | 1.00 | 5.00 | Rp 43.03 |
| Claude Sonnet 5 | 2.00 | 10.00 | Rp 86.05 |
| GPT-5.6 Terra | 2.00 | 12.00 | Rp 94.83 |
| GPT-5.6 Sol | 4.00 | 20.00 | Rp 172.11 |
| Claude Opus 5 | 5.00 | 25.00 | Rp 215.13 |
| Claude Fable 5.1 | 10.00 | 50.00 | Rp 430.26 |
| GPT-6 Astra | 10.00 | 50.00 | Rp 430.26 |
The last column uses the same task: 1,200 input tokens and 250 output tokens. The gap between the first row and the last row is 111 times. Model choice, not price negotiation, decides your bill.
Three notes that come straight from the documentation:
- The Gemini 3.8 Flash price of 0.75 USD and 3.75 USD runs through 31 December 2026. From 1 January 2027 it becomes 1.50 USD and 7.50 USD, which is double.
- OpenAI charges long context at a different rate. GPT-5.6 Sol moves to 8 USD input and 30 USD output once your conversation passes the short-context limit.
- Claude 4.7 and later use a new tokenizer that produces around 30% more tokens for the same text. The price per token dropped, but the token count rises, so recalculate with your own text.
AI subscription prices in Indonesia, 11 September 2026
A subscription charges per person per month, not per task. The ChatGPT and Google prices below are the rupiah prices they show for Indonesia themselves. Claude publishes in dollars, so its rupiah column is my calculation.
| Plan | Price per month | Note |
|---|---|---|
| ChatGPT Go | Rp 75,000 | More messages, this plan may include ads |
| ChatGPT Plus | Rp 349,000 | Reasoning models, Codex, deep research |
| ChatGPT Pro | From Rp 1,889,000 | 5 times the Plus usage |
| Google AI Plus | Rp 75,000 | 2 times more Gemini access, 400 GB storage |
| Google AI Pro | Rp 309,000 | 4 times more access, 5 TB storage |
| Google AI Ultra | From Rp 1,579,000 | Up to 20 times more access, 20 TB storage |
| Claude Pro | 17 USD annual, 20 USD monthly | Around Rp 298,551 to Rp 351,236 |
| Claude Max | From 100 USD | Around Rp 1,756,181, choose 5 times or 20 times Pro |
The annual Claude Pro price is billed up front at 200 USD. None of these plans include applicable tax.
Worked example: 1,500 chat replies per month
This example is a simulation with sample data, not a client result. Use it as a shape and put your own numbers in.
Starting condition. An online shop receives 1,500 chats per month. One reply uses 1,200 input tokens, holding the system instruction, the last 5 messages, and the product data. The reply is 250 tokens.
Step 1. Read the cost of 1 task from the table above.
Step 2. Multiply by 1,500.
| Model | Cost of 1 task | 1,500 tasks per month |
|---|---|---|
| Gemini 2.5 Flash-Lite | Rp 3.86 | Rp 5,795 |
| GPT-5.6 Luna | Rp 9.48 | Rp 14,225 |
| Gemini 3.8 Flash | Rp 32.27 | Rp 48,405 |
| Claude Haiku 4.5 | Rp 43.03 | Rp 64,540 |
| Claude Sonnet 5 | Rp 86.05 | Rp 129,079 |
| Claude Opus 5 | Rp 215.13 | Rp 322,698 |
Step 3. Add the retries. Assume 8% of tasks fail and run again. The Haiku 4.5 bill moves from Rp 64,540 to Rp 69,703.
Step 4. Add the fixed cost. A small server at Rp 120,000 per month brings the total to Rp 189,703.
Observable output. The real cost per task becomes Rp 126.47, almost 3 times the raw model number. That is the figure worth deciding with, not Rp 43.03.

Subscription or API, which costs less
This question often lands in the wrong room. A subscription and an APIAPIThe official door 2 systems use to exchange data, without anybody copying it by hand.Open the glossary sell different things.
A subscription gives 1 person access to a chat app. That person types, reads, and pastes the result. You cannot run a 24-hour automation on a subscription, because the terms restrict it and the quota follows 1 person.
An API gives your program access. The program runs with nobody watching, and the bill follows usage.
Compare the 2 only when a person could do the work either way, for example a staff member rewriting product descriptions:
| Option | Monthly cost | Equivalent task count |
|---|---|---|
| ChatGPT Plus, 1 person | Rp 349,000 | 3,680 tasks on GPT-5.6 Terra, or 36,801 tasks on GPT-5.6 Luna |
| Claude Pro monthly, 1 person | Around Rp 351,236 | 8,163 tasks on Claude Haiku 4.5 |
My working rule: under 3,000 tasks per month with a person doing the work, a subscription is simpler. Above that, or when the work must run without a person, use the API.
5 costs teams forget
- Retries. A failed task is still billed. Record your failure rate, then multiply the cost by it.
- Thinking tokens. On reasoning models, thinking tokens are billed as output. Google states this openly on its pricing page.
- A cache write costs more than a cache read. On Claude Opus 5 a 5-minute cache write costs 6.25 USD while reading it costs 0.50 USD. Caching pays only when the same prompt repeats.
- The batch discount is not automatic. OpenAI halves the price on the batch and flex paths, for example GPT-5.6 Sol drops from 4 USD to 2 USD input. You must change how you call it, not merely wait.
- Long context raises the rate. A conversation that keeps growing makes every later task dearer, even before the long-context rate applies.
I cover how to cut operational AI cost as a whole in cost per task for operational AI tools.
Checklist for calculating AI cost per task
- Define 1 task clearly. One reply, one summary, one draft. Owner: the process owner.
- Measure input and output tokens from 20 real examples, not from a guess. Evidence: 1 note sheet.
- Take the model price from the official page on that day. Write the date down.
- Write the exchange rate you used and its date.
- Add the failure rate and the retries.
- Add the fixed cost, then divide by the task count.
- Compare with the cost of a person doing the same task. Stop when the gap is thin.
Questions I get often
What does one AI chat reply cost? At 1,200 input tokens and 250 output tokens, it costs Rp 3.86 on Gemini 2.5 Flash-Lite up to Rp 215.13 on Claude Opus 5, at the 11 September 2026 prices.
Why is the output price higher than the input price? A model produces output one token at a time, and that work uses more computation than reading the input. The gap is usually 4 to 6 times.
Is 1 token the same as 1 word? No. One English word is roughly 1 to 2 tokens, and Indonesian text often needs more. Measure with your own text, because language and style change the number.
Can I run an automation on a subscription? No. A subscription gives 1 person access to a chat app, while an automation needs an API billed by usage.
Which model is cheapest for simple work? On the 11 September 2026 table, Gemini 2.5 Flash-Lite and GPT-5.6 Luna are the cheapest. Test the quality on your own work before you move the whole load.
Do the rupiah numbers here stay fixed? No. The rate moves every day and a provider can change a price. Recalculate with the formula above before you decide.
Limits and your next step
This calculation measures the model cost. It does not measure the cost of building the workflow, cleaning the data, or the time your team spends reviewing the output.
I promise no saving either. I promise numbers you can check, so your decision stands on today's data.
If you want to know which tasks are worth moving to AI and what they would cost in your business, that is the work we do in AI Workflow Audit. The result is a priority list with an effort and impact estimate.
Sources
- Anthropic, Pricing, read 11 September 2026. Per-model prices, cache prices, and the new tokenizer note.
- OpenAI, Pricing, read 11 September 2026. Standard, batch, and flex prices, plus the long-context rate.
- Google, Gemini API pricing, read 11 September 2026. Per-model prices and the end date of the promotional price.
- Anthropic, Plans and pricing, read 11 September 2026. Claude Pro and Max prices.
- OpenAI, ChatGPT pricing, read 11 September 2026. Rupiah prices for Indonesia.
- Google One, Google AI plans, read 11 September 2026. Rupiah prices for Indonesia.
- Exchange Rate API, 11 September 2026 at 00:02 UTC. Rate 1 USD = Rp 17,561.81.




