Show HN: CostPerPrompt – Live AI API pricing and real-workload cost calculators
The article discusses the cost of AI models and provides a platform called CostPerPrompt that offers live pricing for over 232 models. The platform also includes calculators that help estimate the cost of chatbots, agents, and APIs. The cost of AI models can vary greatly depending on the provider, model, and usage, with discounts available for prompt caching and batch processing.
- ▪The CostPerPrompt platform provides live pricing for over 232 AI models, including GPU rental prices and costs for chatbots and agents.
- ▪The cost of AI models is typically billed per token, with separate rates for input and output, and output is usually 3-5 times more expensive than input.
- ▪Discounts such as prompt caching and batch processing can significantly reduce the cost of AI models, with prompt caching cutting repeated input costs by up to 90% and batch processing taking around 50% off.
Hacker News (AI / LLM) files mainly under ai. We currently carry 3,198 of its stories.
Story provenance
Source · retrieval · rights · ranking — open for full record
inspect →
Story provenance
Attribution is not the same as permission. This drawer separates discovery metadata, excerpts, WeSearch-generated summaries, reuse status, and whether the publisher receives the visit. Nothing here claims a legal grant the publisher has not made.
Record
| Original publisher | CostPerPrompt |
| Canonical URL | https://costperprompt.com/ |
| Publication time | Sun, 02 Aug 2026 01:56:44 +0000 |
| Retrieval time | 2026-08-02T03:05:40.026Z |
| Last seen | 2026-08-02T03:05:40.026Z |
| Headline source | Publisher (no WeSearch rewrite) |
| Excerpt source | publisher body |
| Excerpt method | First ~120 words (~800 chars) of extracted publisher body, fair-use limited. |
| Summary | WeSearch · cerebras-chat (WeSearch summarizer) |
| Summary source text | contentText |
| Citation coverage | Summary is a WeSearch-generated derivative; primary citation is the original publisher URL. |
| Cluster | RhyBgU-KEoCC · 1 stories |
| Cluster logic | Grouped by semantic title/content similarity across sources within a rolling window. Same-publisher template collisions are excluded from coverage comparison. |
| Ranking reason | Story pages are not engagement-ranked. Hub feeds use recency, with optional source-diversified chronological ordering (cap consecutive stories per source). No personalized ranking. |
| Publisher visit | Yes — open original |
| Substitutes article? | No — link-out required for full text |
Rights status (four layers)
WeSearch handling by dimension
| Indexing | May the item be indexed (stored, ranked, made findable)? | Allowed |
| Snippet | May a short excerpt of the publisher's text be shown? | Allowed |
| AI summary | May WeSearch generate its own short summary of the article? | Limited |
| Retrieval / RAG | May the content be exposed for third-party retrieval-augmented generation? | Not asserted |
| Model training | May the content be used to train AI models? | Not asserted |
| Commercial reuse | May the content be reused commercially? | Not permitted |
Basis: Derived from the published RSS/Atom feed. Contact: [email protected]. Reviewed: 2026-07-24.
Opening excerpt (first ~120 words) tap to expand
What will AI actually cost you? Live pricing for 232+ models, refreshed automatically — plus calculators that turn token prices into real answers: what your chatbot, agent, or API workload will cost per month. API cost calculator Chatbot cost calculator All 232 model prices GPU rental prices Every cost question, answered $_ API Cost Calculator Any workload, compared across all 232 models — with caching & batch discounts. 💬 Chatbot Cost Simulates real conversations — growing history, resent context, cache hits. 🤖 Agent Cost Multi-step loops, tool schemas, retries — see why agents cost 10–30× more than you think. 📚 RAG Cost Indexing, retrieval and generation priced separately — spoiler: embeddings are pennies.
…
Excerpt limited to ~120 words for fair-use compliance. The full article is at CostPerPrompt.