
LandingBoost – AI Landing Page Audit Tool with Revenue-Backed Benchmarks
LandingBoost 用 AI 审核 SaaS 着陆页并提供改进建议,提高清晰度、信任度和转化率
@yusukelp · X
Here’s mine!
完整作品展
技术栈
28 projects

LandingBoost 用 AI 审核 SaaS 着陆页并提供改进建议,提高清晰度、信任度和转化率
@yusukelp · X
Here’s mine!

WitBench.com: AI sense of humor benchmark I created witbench.com benchmark because everyone's measuring math and code performance, but personally, I like laughing. TL;DR: Gemi
u/tziki · Reddit
WitBench.com: AI sense of humor benchmark I created witbench.com benchmark because everyone's measuring math and code performance, but personally, I like laughing. TL;DR: Gemini funny, Grok unfunny, but do check out the full list, I spent real money on actual impartial raters. submitted by /u/tziki to r/SideProject [link] [comments]

The compatibility engine for local AI. Tell us what you want to run — we'll tell you exactly which models fit your machine, with verified benchmarks.
@Carl0sFelipe · X
Just shipped — a tool that helps you discover which local AI models actually run on your hardware, with community benchmarks, quantization support, and estimated speed. Building in public from here. #BuildingPublic #AIDevelopment #rust #benchmaks #aimodel

查看和对比主流AI模型的公众意见和基准评分。
u/TasteMysterious5285 · Reddit
I built AI Census, a live field bulletin for how people are actually talking about AI models I’ve been building AI Census, a public “field bulletin” for how people are talking about current AI models. I kept running into the same problem: benchmark tables tell me how a model performs on a test, but not whether people are actually finding it useful, frustrating, reliable, etc. So I built a rolling view from public technical conversations across Reddit, Hacker News, Bluesky, GitHub, and Huggi

在浏览器中运行AI模型基准测试以检测性能回归。
pepperpoppins · HN
Trunchbull, run real models against any benchmark in your browser

Self-hosted hybrid RAG on a €116/month cluster — Postgres, BM25, vectors, a reranker, and a public benchmark score for every claim.
victor_edka · HN
HRAG – Hybrid RAG on €116/month of Hetzner, officially benchmarked

在 leaderboard 上按官方基准对比 AI 大模型的性能排名
fcten · V2EX
做了一个大模型 leaderboard 网站 最近一个月 CodeX 疯狂送重置,token 根本用不完,顺手做点东西。 地址:[知行录]( https://leaderboard.cn/) 排行依据主要为模型官方基准测试成绩。非主观排名。 数据会持续更新。如果有点用,欢迎各位 v 友收藏~

Compare how many AI tasks an hour of work buys across countries, models, reasoning efforts, and benchmark costs per task.
danr4 · HN
Wage Against the Machine – MacWages Index for AI Tasks

浏览 10,000+ 开源 AI 项目,查看性能基准、定价和代码关联。
osaitech · Product Hunt
OpenSourceAI.tech Discover 10,000+ open-source AI projects, models & tools

对比和评估 AI 模型在编码、推理、代理和其他基准测试中的表现。
davidtsong · HN
Benchmarklist: track AI benchmarks (2.4k+), models, and capabilities

在OpenVibeEval中对比不同AI模型生成前端代码和可访问性评分。
u/12qwww · Reddit
I built a live benchmark to see which AI actually writes the best frontend code Hey everyone! I built OpenVibeEval because I was tired of "vibe-checking" AI-generated frontend code. I wanted to know which model actually produces the most accessible and clean React/Tailwind output. What I built: •A leaderboard of 24 models (Claude, GPT, DeepSeek, etc.) ranked by axe-core accessibility scores. •A Harness Comparator to show how different system prompts change the same model's output. •

上传 CAS PDF 获取 AI 驱动的基金组合分析和配置洞见。
@iASHeeesh · X