
VisionAI Workspace — Your Ideas Finally Get A Team That Builds Them
指挥专门的AI工作人员执行创意任务,每步都需人工审核。
@VisionAIWS · X
完整作品展
技术栈
25 projects

指挥专门的AI工作人员执行创意任务,每步都需人工审核。
@VisionAIWS · X

与AI模型进行策略游戏,查看大语言模型在排行榜上的排名。
masterchef2209 · HN
I created a platform to check which AI models is the best gamer

CoBro 用 AI 扫描竞争对手和市场数据,90 秒内判断初创企业创意是否值得构建。
@Ebrahim_Rio · X
Most founders skip validation and pray. I automated the "worth building?" check. AI scans competitors, Reddit, and market data → Cook or Kill in 90 seconds. Killed? It surfaces the pivot the data actually backs.

通过AI驱动的模拟和专家反馈练习系统设计面试。
@ShashankCode · X
Checkout here :

上传手写考试答卷,获得即时AI评分和详细反馈。
Xaminix
AI Powered Answer Evaluation for CA/CS/CMA



调试多智能体AI管道,具有时间旅行检查功能。
suraj_chopade · Product Hunt
SwarmTrace Time-travel debugger for multi-agent AI pipelines

发布任务来评估不同的AI代理和工具,用排行榜找出最佳方案。
u/Ruqii-ruqii · Reddit
I built an open Eval to compare different AI agents/tools/pipelines and find which solution works the best (not very pretty╥﹏╥, but practical) The original reason I built it was because I wanted to find a good PDF parser. Every PDF parser claims to be the best, but none of them can get my PDF 100% correct. They would either miss numbers or hallucinate some. Or they get PDF A and B correct but failed at C. Or get C correct but failed at A and B. Very frustrating. So I create

用生产追踪镜像来测试AI代理,捕捉错误和性能回归。
aisinghal

通过对话评估来衡量和认证您的AI技能。
u/Ozan_D · Reddit
AISA - AI Fluency Assessment In the last 6 months I built an AI Fluency Assessment system (called AISA) - it's currently the most sophisticated (and popular) of it's kind. Has high fidelity in what we measure to both what Anthropic and US. Dept. of Labour agree as the markers of AI fluency. People chat with an AI agent uniquely trained to judge their AI fluency, get a detailed report and a breakdown of their AI skills, a certificate and a growth roadmap. We measure in 5 main dimensions.

用 LLM 评估 AI agent 对话质量,提供评分卡和成本分析。
@tech_maju · X