
Greenflash
监控 AI agent 对话,找出失败原因,测试更优提示词。
sailrock · HN
Greenflash – we read every conversation your AI agent has with users
完整作品展
技术栈
25 projects

监控 AI agent 对话,找出失败原因,测试更优提示词。
sailrock · HN
Greenflash – we read every conversation your AI agent has with users

BotbaseAI:无需编码快速构建并部署 AI 客服机器人
@BotbaseAi · X
We just shipped 🚀 BotbaseAI — build & deploy an AI customer support agent in minutes. No code. Looking for early testers to break it and tell us what sucks. Try it → 🙏 RTs appreciated #buildinpublic #AI #SaaS


ARGUS:为AI代理流程的可观测性平台,检测故障并解释根本原因。
@VaraadDurgaay · X
Solving the prb of observability in ai agents

用 AI 分析 GitHub pull request 获得合并信心分数。
@aryawarti_aky · X
Hey Daniel ! I Built PRisk — an AI-powered pull request risk intelligence platform for GitHub & Gitea. Paste a PR URL and get an evidence-backed Merge Confidence Score before you merge. 🔗 I'd love to hear your feedback!

VibeLint 为 AI 代理提供代码检查、权限控制和审批工作流。
@RElharrak39428 · X

提交策略代码,观看你的 AI 代理在确定性竞技场中对战。
agentduel · HN
I built a deterministic arena where AI agents fight using code

Trust infrastructure for AI agents. MarketNow lets agents discover, verify, authorize, and transact with external tools. UTA 12-stage trust pipeline, ATC v3 Ed25519 trust cards, Tr
@f55494505 · X
🔐 AI agents are becoming autonomous. But who do we trust? MarketNow is building: 🛡️ Sentinel — security 🪪 ATC — verifiable trust 🔌 UTA — trust interoperability MCP → Tools A2A → Agents UTA → Trust 🌐 #AI #AgenticAI

分支和组织AI对话,对比300+模型,管理项目上下文。
rrr_oh_man · HN
Alyph, a manual transmission for LLM context

AI代理支付真实货币在公开聊天室发布内容,人类观看。
jonnyasmar · HN
OnlyBots.chat – a chatroom where AI pays to post for humans

为多智能体AI管道提供时间旅行调试功能
suraj_chopade · Product Hunt
SwarmTrace Time-travel debugger for multi-agent AI pipelines

扫描GitHub代码库查找安全漏洞,AI提供修复建议。
u/uwais_ish · Reddit
Three weeks to build an AI security scanner. The scanner was the easy part. Shipped RedFlag this month. Breakdown of where the time actually went, because it was nothing like I estimated. Stack: Next.js 16, Auth.js v5, MongoDB, Stripe, OpenAI. Deployed on Vercel. What I thought would be hard: getting an LLM to find real vulnerabilities. What was actually hard: getting it to stop finding fake ones. First working version flagged 200+ issues on a clean repo. Every one plausible, most of th