
Referee.chat — An AI Panel Works Your Goal, a Referee Rules
设定目标和标准,由AI小组评估,裁判用引用来源做出判决。
nadermx · HN
Referee.Chat - Set the goal. An AI panel works. Referee clears it done
完整作品展
技术栈
60 projects

设定目标和标准,由AI小组评估,裁判用引用来源做出判决。
nadermx · HN
Referee.Chat - Set the goal. An AI panel works. Referee clears it done

验证AI代理的决策,然后尝试篡改验证记录。
foh_quarters · HN
Verify what an AI agent did, then tamper with the record (no signup)

对比和评估 AI 模型在编码、推理、代理和其他基准测试中的表现。
davidtsong · HN
Benchmarklist: track AI benchmarks (2.4k+), models, and capabilities

向多个AI模型提问,比较答案,观察它们辩论至共识。
u/trekhleb · Reddit
I kept pasting the same question into ChatGPT, Claude, and Gemini in three tabs; so I built a Yes-Brainer — a council of AI models, that answer your question in parallel, debate to consensus, or get judged to a verdict. submitted by /u/trekhleb to r/SideProject [link] [comments]

查看和对比主流AI模型的公众意见和基准评分。
u/TasteMysterious5285 · Reddit
I built AI Census, a live field bulletin for how people are actually talking about AI models I’ve been building AI Census, a public “field bulletin” for how people are talking about current AI models. I kept running into the same problem: benchmark tables tell me how a model performs on a test, but not whether people are actually finding it useful, frustrating, reliable, etc. So I built a rolling view from public technical conversations across Reddit, Hacker News, Bluesky, GitHub, and Huggi

通过偏好学习让你的AI产品理解用户品味
@marcellafjacob · X

AI商业智能平台,具备代理分析和准备度评估功能。
@ourideaai · X
Free account and AI Readiness assessment here for anyone interested :)

社区追踪AI模型体验指数,实时收集用户意见每小时更新。
schafberg · HN
Is AI Dumber Today? An index of AI model experience from user's opinion

Spendict – 为AI生成的广告创意评分,返回投放/修改/停止建议。
ds246 · HN
Spendict – a performance marketer's verdict for AI agents, over MCP

Etch: 追踪、重放和验证AI代理的决策,提供签名审计线索。
u/Funky_Chicken_22 · Reddit
OSS to SaaS positioning problem: when the user persona and the buyer persona are completely disjoint Founder here. Sharing a positioning problem I think a lot of OSS-to-SaaS founders hit and don't talk about publicly. Context: I have been running an OSS project (world-model-mcp) with ~2,500 monthly PyPI installs. Two weeks ago I opened up the hosted companion, Etch, at etch.systems. Launched publicly on Product Hunt at 12:00 PDT yesterday. The positioning problem: OSS user persona: in

帮助企业员工保存和获取机构知识与决策的AI助手。
@romanbodnarchuk · X
Check out what I just built with Lovable!

Aurora - 本地AI代理验证系统,提供透明的推理过程和MCP集成。
brandon_grutkowski · Product Hunt
Aurora Glass-box Quantitative AI for Humans and Agents