
Is a hot dog a sandwich? 12 AI models answer, every week — HOTDOG BENCHMARK
观看12个AI模型每周投票判断食物是否是三明治。
naniel · HN
Hot. Dog. Bench. Mark. The AI benchmark we deserve
完整作品展
技术栈
60 projects

观看12个AI模型每周投票判断食物是否是三明治。
naniel · HN
Hot. Dog. Bench. Mark. The AI benchmark we deserve

Q-Poultry 360° 监测家禽从接收到冷冻的质量和食品安全。
@daifallah01231 · X
Check out what I just built with Lovable!

No ads, no gatekeepers, no boring PR pitches. Just savage AI roasts & leaderboard bidding to reach #1. Will you claim The Throne?
@YagneshPipariya · X
The #1 spot is up for a grab take it now at lowest value right now before someone else get it.

在 30 秒内测量 GPU 的真实内存带宽上限,用于本地 AI。
Ar5en1c · HN
Headroom – measure your GPU's true bandwidth ceiling for local AI

@lightsilver323 https://t.co/jorheojhTQ https://t.co/SPHoe5QNpk https://t.co/KpJkSm4Pxr Hugging Face🤗: we upload our models and datasets. RMCMMK-Bench : our benchmark for Reasoning
@compiwer_ai · X
Hugging Face🤗: we upload our models and datasets. RMCMMK-Bench : our benchmark for Reasoning Math Coding Multilingual Moroccan Knowledge.

3 tests. 90 seconds. Know your rank. A brutally addictive mechanical skill benchmark for PC gamers.
@AmaanHussain09 · X
I'm 16 and built RankCheck, a 90-second mechanical skill test for PC gamers. It measures your aim, movement, and reflexes, then gives you a shareable Flex Card. I hit Diamond rank. I am looking for the first 1k users

WitBench.com: AI sense of humor benchmark I created witbench.com benchmark because everyone's measuring math and code performance, but personally, I like laughing. TL;DR: Gemi
u/tziki · Reddit
WitBench.com: AI sense of humor benchmark I created witbench.com benchmark because everyone's measuring math and code performance, but personally, I like laughing. TL;DR: Gemini funny, Grok unfunny, but do check out the full list, I spent real money on actual impartial raters. submitted by /u/tziki to r/SideProject [link] [comments]

提交网站获得每日速度排名和性能监控。
@thefastestweb · X
daily speed monitoring for indie sites. Submit your URL, get ranked on a public leaderboard, and know the moment your performance drops.

在浏览器中运行AI模型基准测试以检测性能回归。
pepperpoppins · HN
Trunchbull, run real models against any benchmark in your browser

Outbid alternative. Buy the top of a public leaderboard — fuel burns 8% a day, so rank is rented, never owned. No accounts.
@RankFuel_lol · X

提交 LLM 推理优化内核,在专用硬件上进行基准测试并竞争排名。
carsenk · HN
Frontier.fast – Help push the frontier of LLM speed forward

NEURL-OS is a free 2-minute weekly calibration that transforms your weekly habits into a predictive model of your performance.
@NEURL_OS · X
Hey we’ve built NEURL-OS, a tool that turns daily habits into data and a weekly capacity score. For founders juggling constant decisions, it helps reveal what’s fueling your focus and output—and what’s quietly draining it. Try it free: