
CostPerPrompt — Live AI API Pricing & LLM Cost Calculators
Calculate costs across 230+ AI model APIs and simulate budget impacts.
ahmed_hassan7 · HN
CostPerPrompt – Live AI API pricing and real-workload cost calculators
The full gallery
Tech stack
60 projects

Calculate costs across 230+ AI model APIs and simulate budget impacts.
ahmed_hassan7 · HN
CostPerPrompt – Live AI API pricing and real-workload cost calculators

Compare LLM API pricing and calculate your monthly costs instantly.
u/ahmedk2002 · Reddit
I built a real-time LLM API pricing comparator — because I was tired of not knowing the actual cost difference between models I use LLMs daily at work and kept running into the same frustration: provider pricing pages give you raw numbers per million tokens, but no way to understand what that actually means for your specific use case. Is GPT-4o really that much more expensive than Claude Sonnet for 10k requests per day? What about DeepSeek vs Gemini Flash for high-volume summarization? I

Analytics dashboard for LLM API spending by model and environment with optimization suggestions.
ATsimbalistov · HN
Show HN: Tracking GenAI cost and endpoint fragility so app teams don't have to

Compress LLM prompts and documents to reduce token usage and API costs.
@marcusyul · X
THEY JUST GAVE AWAY 100 MILLION FREE TOKENS SO YOU CAN STOP BURNING THROUGH YOUR CLAUDE CODE BUDGET. if you code with AI you already know: the session fills up, starts failing, and on top of that you're overpaying there's a tool that fixes this: it shrinks the context before the model even sees it same model, same response, a fraction of the cost in a real session: from $154 to $43. a 72% drop and right now: → extend your Fable sessions in Claude Code → 100M free tokens to try it out you don't switch models you don't touch your code you just stop paying to repeat yourself link below ⬇️

Use one API to access and switch between LLM providers while optimizing inference costs.
justin2025 · Product Hunt
Auriko Trading desk for LLM calls

Automatically route each prompt to the cheapest capable model to cut API costs.
u/ASDKING100 · Reddit
Launched an AI API router tonight, and found a bug hours in that would've taken real payments without ever upgrading the account Built LLMLite over the past few weeks — it classifies each prompt and routes it to the cheapest model that can actually handle it, instead of hitting GPT-4o for everything. Free tier, no card needed to try it. Tonight, right as I was about to launch, ran a real transaction to test the payment flow end to end. Paddle processed it, webhook fired, signature verified

Free Etsy profit calculator that includes fees, offsite ads, materials, packaging and your labor. Compare pricing scenarios and export PDF/CSV — no signup.
@5150_MadMan · X
Check out what I just built with Lovable!

Find AI models optimized for your hardware with performance and pricing estimates.
cdnsteve · HN
Tokenstead, find AI models for your hardware

Compress prompts before LLM API calls to reduce token usage and costs.
@asgujjuasitgets · X

Calculate how much you'll spend when your app stack exceeds free tier limits.
@K_dev001 · X
Vibe coding makes launching an app almost free. Running it is another story. I built to calculate your full AI + app stack costs and show which free tier breaks first. Sourced pricing. No signup. No guessing. ->

Compress prompts and reduce LLM token costs by detecting duplicate tool calls.
@DeveloperL92487 · X
I built my first app in 60min And now I got $500 MRR in one month Check here if you are interested It’s a tool to reduce agent token consumption, speed up agent response, and clean up memory cache

An LLM gateway for OpenAI, Anthropic, Google and Azure. Every request logged, priced to the token, and audited for waste you can actually recover.
@razdagan3 · X