Research
Notes for engineering and finance teams building AI spend governance.
Start here
Plain-English explainers if you're new to LLM cost.
- How do LLM providers charge?
- Input, output, cache, and reasoning tokens explained
- What is LLM cost attribution?
- Why does my LLM bill spike?
- How to budget for AI spend
All research
- How to audit your LLM spend in 30 minutes
- Caching strategies compared
- Token budget implementation guide
- AI coding plan comparison - evidence-first benchmark
- Agent economics: how much does a coding agent actually cost?
- MCP server cost impact
- LLM cost trends 2025–2026
- AI observability for production teams
- FinOps for LLM: a practical framework
- AI cost optimization
- RAG cost optimization
- LLM token tracking
- Anthropic cost attribution
- LLM cost monitoring
- LLM cost attribution
- Model routing
- Provider arbitrage
- LLM cost anomaly detection
- AI FinOps for production teams
- LLM chargeback and showback
- GenAI cost management
- OpenAI cost attribution
- LLM budget governance
- What is LLM FinOps?
- Agent spend attribution
- Prompt cache attribution
- Reasoning token attribution
- Batch API FinOps
- How to cap inference costs and prevent runaway spending
- Semantic cache economics
- Eval cost allocation
- Multimodal cost allocation
- Cost per request as a product KPI
- Invoice reconciliation for AI bills
- Token budget implementation
- Agent spend guardrails
- Prompt caching explained
- OpenTelemetry GenAI conventions for cost attribution
- On-prem and self-hosted LLM FinOps
- LLM FinOps standards
- LLM cost dashboard: what to put on it
- How much does GPT-5 cost? Complete pricing guide
- Cheapest way to run AI code generation
- LLM cost calculator - estimate your monthly AI spend
- LLM API pricing tracker - per-token pricing across 12+ providers
- Reasoning model cost guide - hidden thinking tokens and budgeting
- Open-Source vs Closed-Source cost comparison - self-hosting break-even analysis
- Hidden LLM costs - beyond per-token pricing
- Azure OpenAI vs direct OpenAI cost
- LLM cost per user benchmarks
- LLM cost benchmarks by industry
- 2026 AI coding economics: GLM 5.2 vs DeepSeek vs MiMo vs Claude
- Cursor vs Copilot vs Claude Code cost
- Claude Sonnet 5 intro pricing ends August 31 - your bill rises 50%
- GPT-5.6 pricing tier guide - Luna, Terra, Sol routing by workload
- OpenAI fine-tuning sunset - the migration economics
- Agent total cost of ownership - the 70% that isn't tokens
- AI cost recovery playbook - when the bill already exploded
- Long context vs RAG - where the cost breakeven actually is
- Model deprecation migration - budgeting for forced moves
- Chargeback when your provider won't bill per team
Want this applied to your own LLM spend? FinOps LLM runs a free audit of your AI costs and shows where the savings are. Book free audit →