
完整作品展
技术栈
60 projects


通过单个预付账户为AI代理提供100多个API访问。
interface1860 · HN
TaskFuel – agents discover and pay per call for 100 APIs

检查 AI 推理踪迹,评估模型真实性。
malik_dixon1 · Product Hunt
TraceLogicAI: AI Architecture Evaluation Compare AI architectures with evidence, not guesswork

Whetstone 将 AI 候选方案与基准对比,否决回归并返回可审计的决策。
@JustinGarr90748 · X
We're building Cyberelf labs because a better score doesn't mean a better model.

用友好界面管理AI代理团队。
jackcollinshq · Product Hunt
YAGNI Proactive agent teams you manage like humans

Orca是用于构建和部署AI代理的平台,提供隔离执行和成本计量。
@okiktech · X

发送真实AI代理测试网站,找出它们卡住的地方并获得修复方案。
@louiswharmby · X
It's funny how that happens! My dog vibe coded and managed to achieve 500 users in the first month! 🤣

Monitor, govern, and optimize your AI agents at scale.
@AiShivam · X
AgentStacKPro is an OS for AI agents designed to make them production-ready, and it holds incredible potential. You can also check out the project I currently developed

测试您的网站对AI代理的准备情况并获得改进建议。
andrewqu · HN
Is-agentic – Score how agentic your product and site is

验证AI代理的决策,然后尝试篡改验证记录。
foh_quarters · HN
Verify what an AI agent did, then tamper with the record (no signup)

部署基于您业务数据训练的AI代理,用于回答客户问题和筛选潜在客户。
u/george_owen123 · Reddit
What I learned building an AI chatbot: the ecom tools are everywhere, but service businesses are completely underserved Spent the last few months building Chirpy, an AI receptionist for websites, with my co-founder. The biggest thing I've learned is a positioning one, and I think it applies well beyond my niche. When we started, we assumed we'd compete in the ecommerce chatbot space, because that's where all the money and attention is. But it's extremely saturated. Gorgias, Tidio, Intercom,

对比和评估 AI 模型在编码、推理、代理和其他基准测试中的表现。
davidtsong · HN
Benchmarklist: track AI benchmarks (2.4k+), models, and capabilities