Ai finops
What an AI Agent Costs Per Conversation on AgentCore
- Pratik Kulkarni
- AI FinOps , AWS
- 28 Aug, 2026
You can read AgentCore's per-service rates straight off the AWS pricing page. What that page can't tell you — and what you actually need before you build a business on agents — is what one of your u
Read more
When Is Self-Hosting an LLM Cheaper Than Bedrock?
- Pratik Kulkarni
- AI FinOps , AWS
- 26 Aug, 2026
Two questions send teams down this path: "our Bedrock bill is growing, should we run this on our own GPU?" and "we fine-tuned a Llama, where does it go?" For most teams the answer to both is no, a
Read more
Model Evals: How to Know If You Can Use a Cheaper Model
- Pratik Kulkarni
- AI FinOps , AWS
- 11 Jun, 2026
An eval, in the AI FinOps context, is a structured comparison: run a representative sample of real production inputs through your current model and a cheaper candidate, score both against a defined qu
Read more
What Is AI FinOps?
- Pratik Kulkarni
- AI FinOps , AWS
- 09 Jun, 2026
AI FinOps is the practice of making AI workload costs visible, attributable, and optimizable — applied to the specific economics of model inference, where the unit of cost is the token, not the instan
Read more
Stretch Your Claude Code Budget with Bedrock Prompt Caching
- Pratik Kulkarni
- AWS , AI FinOps
- 12 May, 2026
Anthropic recently tightened usage limits on Claude Code — and if you're doing serious development work, you feel it. Long refactoring sessions, codebase-wide architecture questions, iterative debuggi
Read more