Cost per request
$0.01627Step 1 Provider and Model
Select pricing rows from the active snapshot.Step 2 Quick Mode
Set the key pricing and behavior assumptions first.Steady usage, moderate context, and balanced quality-cost tradeoffs.
Step 3 Advanced Assumptions
Tune retrieval, reranking, embeddings, vector, cache, and infra.Show advanced inputs
Scenario Share URL
Share this link to load these exact assumptions. Pricing uses the published snapshot shown on the page.Paste into ChatGPT or Claude to discuss this scenario.
Headline metric
Price likely supports target marginRequired price at 80.0% margin: $9.76 per user/month.
Current price: $49.00. Price gap to target: $-39.24.
Cost / user / month
$1.95Break-even price (0% margin)
$1.95Required monthly revenue
$6,346Implied monthly gross profit
$5,077Top Cost Drivers
Largest sensitivity shifts when each variable is increased by 10%.Totals
Pricing and margin summary under current assumptions.Cost per user/month
$1.95Current gross margin
96.0%Required price at target margin
$9.76| Cost per request | $0.01627 |
| Cost per user/month | $1.95 |
| Current gross margin % | 96.0% |
| Required price at target margin | $9.76 |
Component Breakdown (USD/user/month)
Each component is computed independently then summed.Generation
$0.11Retrieval
$0.04Reranking
$2.40Embeddings Ingestion
$0.00Vector Db
$0.00Cache
$-0.65Infra
$0.05| Generation | $0.11 |
| Retrieval | $0.04 |
| Reranking | $2.40 |
| Embeddings Ingestion | $0.00 |
| Vector Db | $0.00 |
| Cache | $-0.65 |
| Infra | $0.05 |
Sensitivity RankingDelta in total cost if one variable increases by 10%.
| Variable | Delta cost % |
|---|---|
| Requests Per User MonthUser activity level per month. | 10.0% |
| Rerank DocsDocs reranked per request. | 9.2% |
| Cache Hit RateFraction of requests served by cache. | -3.3% |
| Output TokensGenerated tokens per request. | 0.3% |
| Retrieved ChunksRetrieved chunk count per request. | 0.2% |
| Tokens Per ChunkAverage chunk size in tokens. | 0.2% |
| Input TokensPrompt-side tokens per request. | 0.1% |
| Vector Queries Per RequestVector query count per request. | 0.0% |
| Monthly Active UsersActive-user estimate used to amortize fixed monthly embedding refresh. | -0.0% |
Assumptions and Units
Explicit assumptions keep this decision reproducible.- CurrencyUSD
- Token unittoken
- Pricing snapshot2026-03-15
- Monthly users650 active users for fixed-term amortization and business totals
- Target margin formularequired_price = cost / (1 - margin)
Recommended Next Step
Use this section to turn break-even math into a pricing plan.If retrieval or hosting is driving cost, check infra options first, then come back to re-check price and margin.
Pricing and Packaging Plan
RAG Break-even Price per SeatHow to Calculate the Break-even Point for AI WorkflowsCompare infra providers
View Infra RecommendationsSources and Snapshot
Pricing comes from a daily snapshot updated by batch workflows.Active Pricing Row
Selected model
OpenAI / GPT-5 Mini
- Input tokens$0.25 / 1M
- Output tokens$2 / 1M
Shared retrieval defaults
- Embedding input$0.02 / 1M
- Rerank docs$1 / 1K
- Snapshot date: 2026-03-15
- Source links and update notes: Pricing Snapshot Reference
Continue Analysis
Move to the next tool or guide without losing your current scenario.Switch tools
- AI Workflow Cost Calculator
- LLM Model Cost Comparison
- RAG Retrieval Cost Calculator
- Reranking Cost Calculator
- Cache Savings Simulator
- Context Window Cost Calculator
- RAG vs Long-Context Calculator
- Embedding Ingestion Cost Calculator
- Browse all tools
Read guides