उपयोग दर विवरण
यह page बताता है कि Neotask usage-based charges को ठीक-ठीक कैसे calculate करता है। Individual plan पर अधिकांश उपयोगकर्ता कभी ये charges नहीं देखेंगे। यह breakdown उन enterprise power users के लिए है जिनके agents scale पर चलते हैं।
Plans के overview और billing कैसे काम करती है इसके लिए, Billing & Plans देखें।
उपयोग की गणना कैसे होती है
हर बार जब आपका agent कोई message process करता है, यह AI tokens (input और output) consume करता है। उन tokens की cost इस बात पर निर्भर करती है कि आपका agent कौन सा AI model उपयोग करता है। System चार प्रकार के tokens को अलग-अलग track करता है, प्रत्येक की अपनी rate है:
- Input tokens: model को भेजा गया prompt और context
- Output tokens: model द्वारा generate किया गया response
- Cache read tokens: prompt cache से serve किए गए tokens (input से सस्ते)
- Cache write tokens: भविष्य में reuse के लिए prompt cache में written tokens
Token Cost Formula
लागत = (input tokens / 1,000,000) x input rate
+ (output tokens / 1,000,000) x output rate
+ (cache read tokens / 1,000,000) x cache read rate
+ (cache write tokens / 1,000,000) x cache write rate
स्रोत: server/src/config/modelPricing.ts
AI Model Token Rates
ये प्रत्येक supported AI model के लिए per-token rates हैं। सभी rates USD प्रति 1 million tokens में हैं।
Anthropic Claude
| मॉडल | इनपुट | आउटपुट | कैश रीड | कैश राइट |
|---|---|---|---|---|
| Claude Opus 4.6 | $15.00 | $75.00 | $1.50 | $18.75 |
| Claude Sonnet 4.5 | $3.00 | $15.00 | $0.30 | $3.75 |
| Claude Haiku 4.5 | $0.80 | $4.00 | $0.08 | $1.00 |
| Claude 3.5 Sonnet | $3.00 | $15.00 | $0.30 | $3.75 |
| Claude 3.5 Haiku | $0.80 | $4.00 | $0.08 | $1.00 |
| Claude 3 Opus | $15.00 | $75.00 | $1.50 | $18.75 |
| Claude 3 Sonnet | $3.00 | $15.00 | $0.30 | $3.75 |
| Claude 3 Haiku | $0.25 | $1.25 | $0.03 | $0.30 |
OpenAI
| मॉडल | इनपुट | आउटपुट | कैश रीड | कैश राइट |
|---|---|---|---|---|
| GPT-4o | $2.50 | $10.00 | $1.25 | $2.50 |
| GPT-4o Mini | $0.15 | $0.60 | $0.075 | $0.15 |
| GPT-4 Turbo | $10.00 | $30.00 | $5.00 | $10.00 |
| GPT-4 | $30.00 | $60.00 | $15.00 | $30.00 |
| o1 | $15.00 | $60.00 | $7.50 | $15.00 |
| o1-mini | $3.00 | $12.00 | $1.50 | $3.00 |
| o3-mini | $1.10 | $4.40 | $0.55 | $1.10 |
Google Gemini
| मॉडल | इनपुट | आउटपुट | कैश रीड | कैश राइट |
|---|---|---|---|---|
| Gemini 2.0 Flash | $0.10 | $0.40 | $0.025 | $0.10 |
| Gemini 1.5 Pro | $1.25 | $5.00 | $0.3125 | $1.25 |
| Gemini 1.5 Flash | $0.075 | $0.30 | $0.01875 | $0.075 |
स्रोत: server/src/config/modelPricing.ts lines 8-42
Platform Fees और Markups
कई अलग-अलग fees हैं जो आपके agent के उपयोग के तरीके के आधार पर लागू हो सकती हैं। ये एकल flat rate नहीं हैं; प्रत्येक अलग-अलग situations में लागू होती है।
1. Credit Mode Markup (System Key): 20%
Neotask के managed API keys (System Key mode) का उपयोग करते समय, raw token cost के ऊपर 20% markup लागू होता है। यह API key management, providers के बीच automatic failover, model routing और infrastructure को cover करता है।
यदि आप BYOK mode (Bring Your Own Keys) का उपयोग करते हैं, तो Token costs सीधे आपके AI provider को जाती हैं, लेकिन आपके credits से 20% platform fee लागू होती है।
स्रोत: server/src/config/modelPricing.ts line 53, CREDIT_MODE_MARKUP_PCT = 0.20
2. Overage Platform Fee (परिवर्तनीय)
जब आपकी usage आपके included credit pool से अधिक हो जाती है, तो शेष cost overage बन जाती है। overage amount के ऊपर एक platform fee लागू होती है। Rate आपके plan tier पर निर्भर करती है:
| स्तर | Overage शुल्क |
|---|---|
| Individual | 25% |
| Business | 15% |
| Enterprise | 10% |
Overage fee केवल उस amount पर charged होती है जो आपके credit pool से अधिक हो। आपके pool के भीतर usage पर कोई अतिरिक्त fee नहीं।
Overage automatically आपके card पर charge होती है जब unsettled amount $10 तक पहुंचती है (Stripe minimum charge: $0.50)।
स्रोत: server/src/config/planConfig.ts, overageFeePct और overageChargeThreshold
3. Automation Markup: Progressive 50% से 33% तक
जब आपके agents automated jobs (scheduled tasks, cron jobs, recurring automations) चलाते हैं, token cost के ऊपर एक additional automation markup लागू होती है। यह इसलिए है क्योंकि automated agents manual work को replace करते हैं। वे unattended चलते हैं, schedule पर, उन tasks को handle करते हुए जिनके लिए अन्यथा employees की आवश्यकता होती।
Automation markup progressive brackets का उपयोग करती है (income tax की तरह)। आप automated usage के अपने पहले dollars पर higher rate pay करते हैं और आपका automated spending बढ़ने पर lower rate। कोई cliff effects नहीं हैं; प्रत्येक dollar केवल अपनी bracket rate पर charged होता है।
| संचयी स्वचालित व्यय (प्रति billing cycle) | सीमांत Markup दर |
|---|---|
| $0 - $10 | 50% |
| $10.01 - $25 | 45% |
| $25.01 - $50 | 40% |
| $50.01 - $100 | 37% |
| $100.01+ | 33% |
स्रोत: server/src/config/automationMarkup.ts lines 38-44, DEFAULT_AUTOMATION_BRACKETS
4. Coding Task Markup: Flat 50%
Automated jobs जिनमें coding tools (file reads, writes, code execution, bash commands) शामिल हैं, flat 50% markup पर charged होते हैं। Progressive reduction coding tasks पर लागू नहीं होती; यह cumulative spend की परवाह किए बिना हमेशा 50% है।
Coding tools: exec, read, write, bash, code
स्रोत: server/src/config/automationMarkup.ts line 49, CODING_TASK_MARKUP_RATE = 0.50
लागतें कैसे Stack होती हैं
ये fees situation के आधार पर stack हो सकती हैं:
Interactive chat (System Key mode):
- Raw token cost + 20% credit mode markup
- कोई automation markup नहीं (automated नहीं)
Interactive chat (BYOK mode):
- Raw token cost + credits से 20% platform fee
Automated cron job (System Key mode):
- Raw token cost + 20% credit mode markup + automation markup (50%-33%)
Automated coding task (System Key mode):
- Raw token cost + 20% credit mode markup + 50% coding task markup
Automation Billing उदाहरण
एक scheduled cron job daily चलता है। AI agent को अपना काम complete करने में लगभग 2-3 घंटे लगते हैं। Job raw tokens में $10 consume करता है (AI provider को actual cost)।
विवरण:
- Token cost: $10 (AI provider को pass-through; यह compute की cost है)
- Automation markup: $10 x 50% = $5 (unattended job चलाने के लिए platform fee)
- कुल शुल्क: $15
$10 actual AI compute को cover करता है जो आपके agent ने उपभोग किया। $5 automation fee है। Platform ने job को schedule पर चलाया, execution monitor किया, retries handle किए और results deliver किए, सभी बिना किसी के computer पर होने की आवश्यकता के।
Automation ज़्यादा क्यों costs करती है: Automated agents human labor replace करते हैं। एक cron job जो हर सुबह आपके analytics check करता है, reports draft करता है, inventory monitor करता है, या incoming orders process करता है वह काम कर रहा है जिसके लिए अन्यथा किसी के समय की आवश्यकता होती। Automation markup उस unattended execution के value को reflect करता है, और यह जितना अधिक आप automate करते हैं उतना कम होता है, scale को reward करता है। $100+/cycle automated jobs पर खर्च करने वाला tenant 50% के बजाय केवल 33% markup pay करता है।
Top-Up Credits
आप किसी भी समय usage के लिए prepay करने के लिए additional credits खरीद सकते हैं:
- न्यूनतम top-up: $5
- अधिकतम top-up: $10,000
- Credits कभी expire नहीं होते। वे आपके account में उपयोग होने तक रहते हैं।
- Auto top-up: वैकल्पिक रूप से automatic top-ups configure करें जब आपका balance एक threshold से नीचे गिरे (default: $2 balance एक $10 top-up trigger करता है)
Top-up credits आपके credit pool exhausted होने के बाद और overage accrued होने से पहले consumed होते हैं। इसका अर्थ है top-ups आपके included credits और overage charges के बीच buffer के रूप में act करते हैं।
स्रोत: server/src/services/overageCharger.ts, server/src/services/balanceService.ts
Credit Deduction क्रम
जब आपका agent एक task complete करता है, cost इस order में deducted होती है:
- Signup credits (एक बार $10 signup credits, reset नहीं होते)
- Top-up balance (purchased credits, कभी expire नहीं होते)
- Overage (platform fee लागू होने के साथ आपके card पर charged)
इसका अर्थ है आपके included credits हमेशा पहले उपयोग होते हैं, फिर कोई purchased top-ups, और केवल दोनों exhausted होने के बाद overage billing शुरू होती है।
स्रोत: server/src/services/balanceService.ts lines 240-286
Budget Controls
Enterprise users के पास built-in budget enforcement है:
- Global daily budget: Default $500/day (configurable)। Exceed होने पर, agent gateway automatically runaway costs रोकने के लिए shut down होता है।
- Per-agent spend limits: प्रत्येक individual agent के लिए daily cap सेट करें। जब agent अपनी limit hit करता है, इसके sessions automatically paused हो जाते हैं।
- Auto-shutoff: Configurable। Budget exceed होने पर automatic gateway shutdown enable या disable करें।
- Real-time enforcement: Budget हर 30 seconds में checked होता है।
स्रोत: Neotask-Electron/src/main/budgetEnforcer.ts, Neotask-Electron/src/main/agentSpendLimitStore.ts
अपनी Usage देखें
Desktop app में Usage page से real time में यह सब track करें:
- Model, provider, agent, channel और date के अनुसार Token-level breakdown
- Cost breakdown: input, output, cache read, cache write costs अलग-अलग दिखाई जाती हैं
- Time granularity: today, last 24h, 7 days, 30 days, 365 days, या custom range
- Per-agent attribution: देखें कि कौन सा agent क्या consume कर रहा है
- Per-channel attribution: messaging channel के अनुसार cost breakdown
- Per-provider attribution: AI provider के अनुसार cost (Anthropic, OpenAI, Google)
- बाहरी analysis के लिए CSV/JSON export
अपना subscription manage करने के बारे में अधिक के लिए Billing & Plans देखें।