
Mirrors - Test AI Agents Against a Mirror of Production
用生产追踪镜像来测试AI代理,捕捉错误和性能回归。
aisinghal
完整作品展
技术栈
60 projects

用生产追踪镜像来测试AI代理,捕捉错误和性能回归。
aisinghal

Cognitive trade-off modeling and high-stakes decision intelligence engine.
@PrismAI_HQ · X
Make smarter decisions by seeing the trade-offs clearly. That’s Prism AI, an AI decision tool for when you’re stuck choosing between options. On Product Hunt Aug 29

通过统一 API 层将网站与多个 AI 系统集成。
@rofarkas · X

通过对话评估来衡量和认证您的AI技能。
u/Ozan_D · Reddit
AISA - AI Fluency Assessment In the last 6 months I built an AI Fluency Assessment system (called AISA) - it's currently the most sophisticated (and popular) of it's kind. Has high fidelity in what we measure to both what Anthropic and US. Dept. of Labour agree as the markers of AI fluency. People chat with an AI agent uniquely trained to judge their AI fluency, get a detailed report and a breakdown of their AI skills, a certificate and a growth roadmap. We measure in 5 main dimensions.

CandrelOne 帮你创建和自动评分候选人评估。
@aishwary07jain · X

与AI面试官练习编码和系统设计面试。
u/redskinnypete · Reddit
I made an AI interviewer for coding and Design interviews I built an AI-powered mock interviewer that simulates a real system design interview. It listens to your approach, asks follow-up questions based on your design, challenges your decisions, and provides detailed feedback on what you did well and where you can improve. I have tried to make it feel like a real interview as much as I could to simulate the pressure and the expectations in an interview. Please give it a try. Hope you wil

AEE 是一个开源控制平面,对 AI 代理的每个操作进行授权、监控和验证。
eli-labz · GitHub
Agent-Execution-Partnership Agent Execution Partnership AEE is an open-source control plane that ensures every AI agent action is authorized before it runs, observable while it runs, and verifiable after it completes.

用 LLM 评估 AI agent 对话质量,提供评分卡和成本分析。
@tech_maju · X


将商业问题转化为公司简案和可执行建设计划的AI原生操作系统。
@MarkZofMarkZ · X
- in process of updating it currently, making it better. How about you? What are you building?

向Claude、GPT和Gemini提问,获得它们相互审核的共识答案。
@StevenJdotCom · X
One AI makes mistakes. Three catch each other's. AI Consensus runs your prompt through Claude, GPT and Gemini. They work it independently, then critique each other until they reach consensus — handing you an AI audited, combined answer.

通过AI风险测验评估您的工作安全,获取个性化技能建设建议。
@comingupnachos · X
vibe coding, is my job safe V1 goal to show an optimistic future for those who prepare and start upgrading skills, ai won't take your job a human using ai might Not techy just learning