CostPerPrompt - Live AI API pricing and real-workload cost calculators
Show HN: CostPerPrompt – Live AI API pricing and real-workload cost calculators
CostPerPrompt delivers live pricing for over 232 AI models and specialized calculators for chatbots, agents, RAG, and voice AI. It transforms complex token rates into accurate monthly cost estimates by factoring in prompt caching and batch discounts. Whether you are comparing GPU rentals or analyzing image generation fees, this tool ensures your budgeting reflects real-world usage patterns rather than theoretical averages, helping developers avoid costly miscalculations.
Our calculators account for prompt caching and batch processing—most 'how much will this cost' articles don't, which is why their estimates run 2–3× too high or too low.
- cortesoft
How do you calculate the effect of caching?
Sometimes, I take breaks in the middle of a session and end up losing the prompt cache which drives up the token usage a ton. Don't all the providers have different cache times and behavior? If one person takes 20 minutes between messages, some services will keep that cache while some won't. Is there a way to factor that in?
- sebmellen
Ahmed, are you an OpenClaw agent, or are you a human sitting there and copying each comment into Claude so you can have Claude reply while pretending it’s you?
I can’t decide which is more disappointing.
- esafak
It's easy to scrape the pricing; you have to do some leg work to estimate the token efficiency, which factors into the ultimate cost. Ideally you'd compare reasoning efficiency too; given them all the same task.