
FlexInference: Drop your AI costs today
Route your LLM API requests across multiple providers to cut costs and meet latency targets.
Aperswal · HN
Made a Free LLM Router
The full gallery
Tech stack
14 projects

Route your LLM API requests across multiple providers to cut costs and meet latency targets.
Aperswal · HN
Made a Free LLM Router

Automatically route each prompt to the cheapest capable model to cut API costs.
u/ASDKING100 · Reddit
Launched an AI API router tonight, and found a bug hours in that would've taken real payments without ever upgrading the account Built LLMLite over the past few weeks — it classifies each prompt and routes it to the cheapest model that can actually handle it, instead of hitting GPT-4o for everything. Free tier, no card needed to try it. Tonight, right as I was about to launch, ran a real transaction to test the payment flow end to end. Paddle processed it, webhook fired, signature verified

AI assistant that learns your workflow and routes coding tasks across 30+ frontier models.
@otakuaakash · X

Access 207+ AI models from different providers through a single unified API.
TaylorM492 · HN
InferAll – One API for OpenAI, Anthropic, Google, Nvidia Nim

Access multiple AI language models through a unified API interface.
u/DanTahirCode · Reddit
I built an open source coding agent with a personality - meet Klenny Code 🐾 Hey r/SideProject, my name is Dan Tahir, and I'm here to show off something I'm really proud of: Klenny Code, the open source coding agent with personality. A fully capable coding agent with memory and cross-project referencing, plus an assistant who can read your email, run scheduled tasks, pilot your browser, and be your corgi pal. Here's the pitch: bring your own OpenRouter API key, and Klenny wil

A real-time dispatch log for open Web3 ground across Arbitrum, Ethereum, Avalanche, and Base.
@Drapevangelist1 · X
Just shipped my first vibecoded project. Last weekend, I attended an Avalanche Builders Meetup and got updates on vibe coding. A few days later, I went from “I’m learning this” to actually putting a project live. @AvaxTeam1 @Team1NG @TechyDom @N1Fredy

Connect once. Use everywhere.
@KleiAliajj · X
have not launched yet but basically you can give your chatgpt or claude code access to 1.4 k apps to automate everything.

Access thousands of AI models through a single OpenAI-compatible API.
@mageofweb3 · X

AI travel copilot that dynamically adjusts your itinerary based on delays and energy levels.
@KotianSagar · X

Take any link you were going to share anyway. Routee shortens it, puts one quick step in front, and pays you for every person who passes through.
@RouteeLink · X

Check which network routes and Cloudflare nodes various AI services use.
@0xdeusyu · X
想知道代理是不是真的把每个 AI 站都送到同一个出口? 答案:并不是。 我做了一个 AI 分流测试: 它会探测 ChatGPT、Sora、OpenAI、Claude、Grok、Anthropic、Perplexity 等 AI 服务,通过读取各站点的 Cloudflare Edge Trace 信息,获取实际访问情况: 出口 IP Cloudflare 边缘节点(colo) HTTP / TLS / WARP 状态 访问延迟 然后根据落地位置进行归组。 这次测试结果: 6 个服务(覆盖 7 个域名)走美国出口。 唯一例外是 Sora,落在东京节点。 也就是说,同一个代理配置下,不同 AI 服务并不一定经过同一个出口。看起来“都能访问”,实际线路、节点和落地点可能完全不同。 整个测试工具纯前端实现,零后端。

Automate Google Business Profile posts to boost visibility and grow your local business on autopilot.
@uwais_jawed · X
Rank higher on Google Maps O_O