
I am speed. A fast.com-style benchmarking tool for LLM APIs | OpenAI, Anthropic,
Compare latency and throughput performance across LLM API providers.
@QAInsights · X
The full gallery
Tech stack
60 projects

Compare latency and throughput performance across LLM API providers.
@QAInsights · X

OpenAI-compatible API for running open-weight LLMs and video models.
bingus-bongo · HN
Use GLM-5.3 in Cursor today via tokengo API

An LLM agent that tracks goals and plans across sessions while showing exactly what it retrieves, verifies, and fails on.
u/OGMYT · Reddit
I built LOLM, a lower-cost LLM agent that shows what it actually did — looking for blunt feedback I’m one of the founders/builders behind LOLM. Most AI products show an answer but hide whether the system retrieved anything useful, verified the result, switched models, hit a limit, or simply stopped. LOLM exposes those parts through controller events and run receipts. It includes: - Live agent - CLI - Coding and small app-building workflows - Memory and self-hosting options - Control decis

Deploy autonomous coding agents to your repositories with multi-model routing.
@LeeLeepenkman · X
nice im working on lots of AI stuff so right now :)

Manage API keys and usage budgets for coding-agent workflows with request routing.
u/Zyron_X · Reddit
I built a service for people to use Codex API without 5-hour limit disruption I built a small service for people who use the OpenAI Codex API regularly and want more predictable usage without the 5-hour or weekly limits. It currently provides: Frontier OpenAI models (GPT 5.6 family included) Managed API key Monthly usage budgets depending to plan No 5-hour limit No weekly limit Under the hood, it is built on top of an open-source project and proxies requests to

Visualize and simulate the BBRv3 congestion control algorithm in your browser.
dilyevsky · HN
BBRv3 for gVisor's netstack, visualized in the browser using WASM

Tests LLM endpoints with adversarial cases and provides OWASP-mapped security audit reports.
@aryaan_sheth · X
- LLM security for small teams

Host a dedicated LLM instance in the EU with flat-rate pricing and no usage limits.
CodingPanda42 · HN
Virtual Private LLM, fixed fee with no usage or token limits

Mask sensitive data before sending to any LLM, then restore it in the reply.
@velumprivacy · X

AI travel copilot that dynamically adjusts your itinerary based on delays and energy levels.
@KotianSagar · X

Build and host AI-powered applications with data storage and persistent URLs.
@akhileshrangani · X
i built codex micro and used it inside of claude to control codex AND claude code it uses a herdr bridge that is running on my mac talks it through a ngrok proxy uses to render inside of claude

Run LLM inference on consumer GPUs with NVIDIA TensorRT-LLM optimization.
brianhabana123 · HN
TensorRT-LLM running natively on Windows (no WSL)