
TokenSpend
Track AI coding costs attributed to pull requests, teams, and organizations.
haseebejaz · HN
TokenSpend, the AI ROI Solution
The full gallery
Tech stack
18 projects

Track AI coding costs attributed to pull requests, teams, and organizations.
haseebejaz · HN
TokenSpend, the AI ROI Solution

API providing token-level citations for LLM output grounded in attention analysis.
apoorvumang · HN
TokenPath – token-level citations for LLM output, read from attention

Simulate SaaS pricing models, token costs, and profit margins across 30+ currencies.
@HelloCalcaas · X

Compare LLM API pricing across 450+ routes and calculate real monthly costs with caching and batch pricing factors.
u/Greywolff06 · Reddit
I built LLMPrice — a free calculator for comparing LLM API costs across 450+ pricing routes I kept running into the same problem when comparing LLM APIs: the headline token price doesn't always tell you what your actual workload will cost. Caching, batch pricing, reasoning tokens, retries, different endpoints, and OpenRouter routes can change the result quite a bit. So I built LLMPrice.com. You enter your workload once — requests, input/output tokens, caching, retries, etc. — and it com

Reduce unnecessary tokens in prompts to lower API costs for Claude, ChatGPT, Gemini, and Grok.
u/HourRevolutionary666 · Reddit
Solo founder, first SaaS. Honestly not sure how to get from “it works” to “people use it” Okay so here’s where I actually am right now, not the polished version. Spent months building this on my own. It’s an AI/SaaS tool called Token Optimiser that trims unnecessary tokens out of prompts before they hit the model, so you pay less per call without losing what the prompt actually needs. It’s live at https://www.tokenoptimiser.com , I ran it through a proper benchmark to make sure the numbers

Monitor and compare your spending across AI providers like OpenAI, Claude, and Gemini.
@J0nasDav1d · X
Celebrate burning tokens

Analytics dashboard for LLM API spending by model and environment with optimization suggestions.
ATsimbalistov · HN
Show HN: Tracking GenAI cost and endpoint fragility so app teams don't have to

Semantic caching reduces LLM token costs and latency for AI queries.
u/ornymo_official · Reddit
how to reduce ai costs there are lots of way to reduce costs but there all complex to setup i know this cause i tried one in production so i built ornymo we cache meaning not the exact string allowing us to give same awnsers thus reducing llm costs and latency check it out at ornymo.com free for a limited time and let me know your feedback submitted by /u/ornymo_official to r/buildinpublic [link] [comments]

Access 207+ AI models from different providers through a single unified API.
TaylorM492 · HN
InferAll – One API for OpenAI, Anthropic, Google, Nvidia Nim

Access thousands of AI models through a single OpenAI-compatible API.
@mageofweb3 · X

Scores AI-generated ad creative and returns verdicts (run/fix/kill) via MCP and REST API.
ds246 · HN
Spendict – a performance marketer's verdict for AI agents, over MCP

Which AI should you use? 82 models, assistants and tools, each rated on what it is actually for. One vote per person, per week.
@AIModelRanking · X