
Trunchbull
在浏览器中运行AI模型基准测试以检测性能回归。
pepperpoppins · HN
Trunchbull, run real models against any benchmark in your browser
完整作品展
技术栈
18 projects

在浏览器中运行AI模型基准测试以检测性能回归。
pepperpoppins · HN
Trunchbull, run real models against any benchmark in your browser

Echo – Fable-level results at 1/3 the cost using open-weight models
adam_rida · HN
Echo – Fable-level results at 1/3 the cost using open-weight models

Delta Runtime reduces repeated AI computation by preserving validated work and recomputing only what changes.
@MaureenSeaberg · X
Delta makes AI up to 14× faster—by computing only what changed: live demo.

一个 AI 工具,用于分析您每周的时间使用情况,并在 10 分钟内展示您的生产力提升机会。
@BlackLedgerSig · X

计算应用栈何时超出AI、托管、数据库等免费层。
@K_dev001 · X
Vibe coding makes launching an app almost free. Running it is another story. I built to calculate your full AI + app stack costs and show which free tier breaks first. Sourced pricing. No signup. No guessing. ->

在OpenVibeEval中对比不同AI模型生成前端代码和可访问性评分。
u/12qwww · Reddit
I built a live benchmark to see which AI actually writes the best frontend code Hey everyone! I built OpenVibeEval because I was tired of "vibe-checking" AI-generated frontend code. I wanted to know which model actually produces the most accessible and clean React/Tailwind output. What I built: •A leaderboard of 24 models (Claude, GPT, DeepSeek, etc.) ranked by axe-core accessibility scores. •A Harness Comparator to show how different system prompts change the same model's output. •