
FlexInference: Drop your AI costs today
Route your LLM API requests across multiple providers to cut costs and meet latency targets.
Aperswal · HN
Made a Free LLM Router
The full gallery
Tech stack
60 projects

Route your LLM API requests across multiple providers to cut costs and meet latency targets.
Aperswal · HN
Made a Free LLM Router

Automate load testing for modern teams without complex setup or manual scripting.
@GorodkovVi85373 · X
- load testing made easy even without enginnering team. Faster, cheaper, distributional

Semantic caching reduces LLM token costs and latency for AI queries.
u/ornymo_official · Reddit
how to reduce ai costs there are lots of way to reduce costs but there all complex to setup i know this cause i tried one in production so i built ornymo we cache meaning not the exact string allowing us to give same awnsers thus reducing llm costs and latency check it out at ornymo.com free for a limited time and let me know your feedback submitted by /u/ornymo_official to r/buildinpublic [link] [comments]

Get instant browser notifications when Claude or GPT usage limits reset.
@HYBBB2002 · X
🔔 Never miss an AI limit reset event again! presentAI tracks Claude & GPT reset events in real-time and sends you an instant alert the exact moment your limits are back.

Every token your team spends on AI coding agents, attributed to the builder who spent it and the pull request it shipped. Harness-agnostic, local-first capture. Open source.
@RuffinelliMarco · X
Would love to see you on the board.

Memecoin arena: bid SOL or USDC from Phantom to rank your token.
@PasthiDev · X
memecoin version of outbid! ahahah

Manage API keys and usage budgets for coding-agent workflows with request routing.
u/Zyron_X · Reddit
I built a service for people to use Codex API without 5-hour limit disruption I built a small service for people who use the OpenAI Codex API regularly and want more predictable usage without the 5-hour or weekly limits. It currently provides: Frontier OpenAI models (GPT 5.6 family included) Managed API key Monthly usage budgets depending to plan No 5-hour limit No weekly limit Under the hood, it is built on top of an open-source project and proxies requests to

Compare 50+ perp DEXs in one place — live volume, open interest, funding rates, fees, price arbitrage and best execution routes. Hyperliquid, Lighter, Aster, Nado, dYdX and more.
@cryppimagic · X
中国的朋友们,大家好。我很喜欢看你们在 X 上发的内容。 想跟大家分享一下,我做了一个完全免费(无任何限制)的套利扫描工具,覆盖 45 家 Perp DEX 和 CEX

Popular agents in the cloud, one call away. Claude Code, Codex, Hermes, pi & more — one API key, billed per use.
@agentsky_dev · X
Built for agents, not just models. AgentSky is live: one API to run Claude Code, Codex, DeepSeek, Kimi, OpenCode, and more in the cloud. Connect your real tools, then compare time, cost, and tokens side by side in Agent Playground. Try it:

Freelancer invoicing without manual entries
@Cshsasvsis · X
SaaS tool for freelancer invoicing. Do check it out.

Play as an AI, gobbling tokens while dodging prompt injections and rate limits in this browser arcade game.
dschwede · HN
Token Gobbler – the goofy game where you're the LLM

Investigate wallets, tokens, transactions, approvals, and smart contracts with evidence-led Web3 security tools, blockchain research, and connected intelligence.
@ceotokentoolhub · X
Scaling @TokenToolHub