
FlexInference: Drop your AI costs today
Route your LLM API requests across multiple providers to cut costs and meet latency targets.
Aperswal · HN
Made a Free LLM Router
The full gallery
Tech stack
17 projects

Route your LLM API requests across multiple providers to cut costs and meet latency targets.
Aperswal · HN
Made a Free LLM Router

ReRune is software localization management for product teams. Manage translation keys, AI-assisted translations, CLI/API sync, and OTA app text updates in one workspace.
@reruneio · X

Paste your URL, get a shipshape score from 0 to 100, and remove the AI defaults with a FIX.md your coding agent applies. We score the defaults, not your taste.
@zllhff · X
Paste your URL, get a design score out of 100 in about a minute, free. shipshape flags the generic AI look your tool shipped by default, then writes a FIX.md you can use with your tools to remove it. The result looks designed, not vibe-coded.

Permanent webhook endpoints and instant HTTPS tunnels for developers and AI agents.
@BinaryScriptar · X
Introducing OtterKit. Always-on webhook endpoints and instant tunnels, built for developers and their AI agents. It started with a problem I kept hitting: every time my coding agent needed to test a Stripe or GitHub webhook, it had to open a tunnel. What OtterKit does • Permanent webhook URLs that answer senders 24/7 with the status, body and headers you choose, verify signatures on arrival (Stripe, GitHub, Shopify, Slack and 12 more), store full history, and forward matching events to your app, signed and retried • Instant HTTPS tunnels to any local port in one command, stable subdomains, auto stop, basic auth • Replay any captured request, edit the payload, re sign it, fire it at your local handler • Pulses: scheduled HTTP calls and dead man's switches with email alerts • Payload drift detection: baseline the JSON shape per event type, get emailed when a provider changes it • Custom domains ( and an email inbox on the same endpoint • A

Your OpenAI client, a different base URL, a much smaller invoice. Frontier open-source models on a decentralized GPU network.
@runnoclip · X
AI inference service that cuts your token bills by 50-90%