Articles and pages on what an AI product actually costs to run, per request, per active user, per subscriber, and the FinOps practice built around that number.
-
LLM Cost Per Request: What One AI Request Actually Costs
Read the piece: LLM Cost Per Request: What One AI Request Actually CostsNine cost drivers behind one LLM request (input, output, cached tokens, retries, tool calls) and the ranges each landed at on production AI products.
-
LLM Cost Calculator
Read the piece: LLM Cost CalculatorFree LLM cost calculator: model, token split, cache hit rate in; cost per request, per user, per subscriber out, built from production AI spend.
-
FinOps for AI: What Changes When the Workload Is Inference
Read the piece: FinOps for AI: What Changes When the Workload Is InferenceCloud FinOps rules break once the workload is inference. Four assumptions that fail on AI products, the fix for each, and what changed once they shipped.
-
FinOps Consulting for AI Products
Read the piece: FinOps Consulting for AI ProductsFinOps consulting for AI products. Cost per request, per user, and per subscriber, a model that runs before the build, and a forecast Finance runs alone.