
FlexInference: Drop your AI costs today
Route your LLM API requests across multiple providers to cut costs and meet latency targets.
Aperswal · HN
Made a Free LLM Router
The full gallery
Tech stack
26 projects

Route your LLM API requests across multiple providers to cut costs and meet latency targets.
Aperswal · HN
Made a Free LLM Router

Open-source prompt compression that reduces LLM token costs by compressing input before API calls.
@asgujjuasitgets · X

API providing token-level citations for LLM output grounded in attention analysis.
apoorvumang · HN
TokenPath – token-level citations for LLM output, read from attention

Semantic caching reduces LLM token costs and latency for AI queries.
u/ornymo_official · Reddit
how to reduce ai costs there are lots of way to reduce costs but there all complex to setup i know this cause i tried one in production so i built ornymo we cache meaning not the exact string allowing us to give same awnsers thus reducing llm costs and latency check it out at ornymo.com free for a limited time and let me know your feedback submitted by /u/ornymo_official to r/buildinpublic [link] [comments]

Monitor AI API costs from OpenAI, Anthropic, and xAI by connecting provider keys.
@ylin_ · X
So my launch of – the simple, proxy-less AI model cost monitoring software – didn’t generate much response. I said I would build this in public, so here is an update. I am very happy with the app for my own use. It’s everything I want it to be. So even without users, I’m glad I built this just for myself. It’s probably not accurate to say I “launched” this. All I did was publish the app, make a single post on X, LinkedIn and Show HN. That’s not really any sort of distribution effort needed for an app. So I will continue to get the word out regarding Costbase on Reddit, and other places, as well as engaging with X users who have problem with AI cost. (If you guys have ideas, lmk) It’s just a small app, a side project, so I won’t put too much time marketing this before calling it quit. Pre-AI, I’ve always spent so much effort building an app that I don’t have any time left for marketing (I also run an ERP consultancy). Now, building is pretty much automated, so th

Billing infrastructure for AI startups, handling payments, usage limits, and credits.
@coleywoleyyy · X
Yup, it’s buns. I vibe coded my own tool at Helicone, might make it easier

Use one API to access and switch between LLM providers while optimizing inference costs.
justin2025 · Product Hunt
Auriko Trading desk for LLM calls

Play as an AI, gobbling tokens while dodging prompt injections and rate limits in this browser arcade game.
dschwede · HN
Token Gobbler – the goofy game where you're the LLM

Decentralized freelance marketplace where you can earn Solana tokens for skilled work.
@aka_gaurang · X
There are bounties for creators to grow

Compare latency and throughput across multiple LLM API providers.
@QAInsights · X

Trade SpaceX stock tokens on Robinhood Chain with automatic treasury accumulation and supply reduction.
@Stock_Works · X
We built the SpaceX stock token accumulation engine A true RWA flywheel that can only exist on Robinhood chain The treasury nearly owns 200 $SPCX stock tokens. With 5 ETH still liquid. Find me another platform doing this?

Monitor AI costs by feature, forecast spending, and set guardrails for your team.
@eastbase_studio · X