
I am speed. A fast.com-style benchmarking tool for LLM APIs | OpenAI, Anthropic,
Compare latency and throughput performance across LLM API providers.
@QAInsights · X
The full gallery
Tech stack
17 projects

Compare latency and throughput performance across LLM API providers.
@QAInsights · X

Upload vendor proposals to get AI-powered side-by-side comparison with red flags and citations.
@nbOlveira · X
A vendor comparison software

Benchmark version-control systems and coding agents on realistic development tasks.
videlov · HN
I was interested in answering this question so I built a benchmark comparing git, jj and gitbutler in agentic context https://vcbench.dev/ Disclaimer - I am a co-founder of GitButler

Compare and benchmark SaaS APIs with verified data to decide whether to build or buy.
fenilsuchak · HN
OpenBenchmarks – Helping agents discover and pick the right SaaS APIs

Analyze Python code across 14 quality dimensions to detect violations and measure capabilities.
@KSFirasa · X
Hello! I built a tool that profiles code (python only atm) across 14 dimensions detecting violations and capabilities outputting a full report. A bit more nuanced than "AI-powered insights". Free while in beta. Thank you!

Submit your website for daily speed benchmarking and ranking on a public leaderboard.
@thefastestweb · X
daily speed monitoring for indie sites. Submit your URL, get ranked on a public leaderboard, and know the moment your performance drops.

Describe competitors in plain language to get live web-searched side-by-side analysis and dossiers.
@Wagner__kent · X
Good day Gainframe, built CompetitorSearch — describe your competitors, get a live web-searched side-by-side breakdown + full dossier on each. Selling it outright (Mistral+Tavily+Vercel, 156 visits/19 searches/10 users). Happy to walk you through it.

Inspect and remove hidden Unicode artifacts in AI-generated text without altering visible content.
u/nategdd · Reddit
I published reproducible fixtures for a lossless AI text artifact scanner I built AI Text Watermark Remover to inspect copied AI text without rewriting visible words. It reports exact hidden Unicode code points, removes only supported literal artifacts locally, and does not claim that hidden characters prove AI authorship. I just published the browser compatibility fixtures, artifact coverage benchmark, self-hosted API, Docker image, and open-source scanner so the claims can be tested inste

Assess your B2B AI or SaaS startup's biggest bottleneck blocking your next milestone.
@FounderUnstuck · X
If you’re building a startup and want to identify your biggest bottleneck, try the free assessment: 🔗 We’d love to hear if the results match your experience.

Play strategic games against AI models and see how different LLMs rank on an objective leaderboard.
masterchef2209 · HN
I created a platform to check which AI models is the best gamer

Validate your SaaS idea's defensibility with scoring, moat analysis, and competitor gaps.
@saralsachan · X

Publish tasks to evaluate and benchmark different AI agents and tools on a leaderboard.
u/Ruqii-ruqii · Reddit
I built an open Eval to compare different AI agents/tools/pipelines and find which solution works the best (not very pretty╥﹏╥, but practical) The original reason I built it was because I wanted to find a good PDF parser. Every PDF parser claims to be the best, but none of them can get my PDF 100% correct. They would either miss numbers or hallucinate some. Or they get PDF A and B correct but failed at C. Or get C correct but failed at A and B. Very frustrating. So I create