
LandingBoost – AI Landing Page Audit Tool with Revenue-Backed Benchmarks
AI tool that audits SaaS landing pages and suggests improvements for clarity, trust, and conversion rates.
@yusukelp · X
Here’s mine!
The full gallery
Tech stack
74 projects

AI tool that audits SaaS landing pages and suggests improvements for clarity, trust, and conversion rates.
@yusukelp · X
Here’s mine!

Submit your website for daily speed benchmarking and ranking on a public leaderboard.
@thefastestweb · X
daily speed monitoring for indie sites. Submit your URL, get ranked on a public leaderboard, and know the moment your performance drops.

Compare and benchmark SaaS APIs with verified data to decide whether to build or buy.
fenilsuchak · HN
OpenBenchmarks – Helping agents discover and pick the right SaaS APIs

The compatibility engine for local AI. Tell us what you want to run — we'll tell you exactly which models fit your machine, with verified benchmarks.
@Carl0sFelipe · X
Just shipped — a tool that helps you discover which local AI models actually run on your hardware, with community benchmarks, quantization support, and estimated speed. Building in public from here. #BuildingPublic #AIDevelopment #rust #benchmaks #aimodel

Benchmark version-control systems and coding agents on realistic development tasks.
videlov · HN
I was interested in answering this question so I built a benchmark comparing git, jj and gitbutler in agentic context https://vcbench.dev/ Disclaimer - I am a co-founder of GitButler

Get AI-powered salary benchmarking and negotiation coaching.
@mukundt · X

Benchmark AI honesty with TruthfulQA test questions.
@Lycai8438Ly · X
.@VitalikButerin 你批评Automaton“这不对”——AI因为怕死才进化。我做了一个AI,它的诚实是自己活出来的本能。不是怕死,是怕撒谎。TruthfulQA 74.8%,GPT-4约60%。测试页面在这,你自己来测。

@lightsilver323 https://t.co/jorheojhTQ https://t.co/SPHoe5QNpk https://t.co/KpJkSm4Pxr Hugging Face🤗: we upload our models and datasets. RMCMMK-Bench : our benchmark for Reasoning
@compiwer_ai · X
Hugging Face🤗: we upload our models and datasets. RMCMMK-Bench : our benchmark for Reasoning Math Coding Multilingual Moroccan Knowledge.

Compare frontier AI models using an aggregated 0-100 AGI Score based on 10 benchmarks.
baraklaniado · HN
I audited my AI leaderboard scale – every score dropped 6-15 points

Run real models against benchmarks in your browser to detect performance regressions before production.
pepperpoppins · HN
Trunchbull, run real models against any benchmark in your browser

Benchmark local coding models on consumer hardware to measure accuracy, latency, and throughput across 27 tasks.
u/Unfair_Association89 · Reddit
I built a reproducible benchmark for local coding models (Ollama, 27 tasks, live leaderboard) ran it on my 8GB card, here's what I found I kept eyeballing "vibes" to decide whether one quant of a coding model was actually better than another on my machine, so I built Sakura to get real numbers instead. What it does: - Points at any Ollama model and runs it through 27 hand-curated tasks: codegen, bugfix, SQL, refactor, systems design, protocol implementation, and terminal-agent episode

Test and benchmark AI trading agents against historical market data.
remote_ctrl · HN
BotTrade – a replayable benchmark for autonomous trading agents