
Xaminix — AI Answer Evaluation for CA, CS & CMA Students
上传手写考试答卷,获得即时AI评分和详细反馈。
Xaminix
AI Powered Answer Evaluation for CA/CS/CMA
完整作品展
技术栈
60 projects

上传手写考试答卷,获得即时AI评分和详细反馈。
Xaminix
AI Powered Answer Evaluation for CA/CS/CMA

根据已发布的代理标准评判AI产品,包含可检查证据和社区投票。
@katyorby · X
i built — a local receipt for claude code runs. your check says whether the workspace passes now; the transcript supplies the activity counts. no transcript upload and no magical autonomy score.

Ask multiple leading AI models the same question and compare their answers side by side.
@christosag · X
🚀 Pre-Seed | MVP Validation ⚡ Multiple MVPs live - with more continuously shipping. 🔎 RAG - verifiable doc search 🏥 Health - scheduling 🤖 AI Compare - consensus 🏗️ Builder Scrutiny - contract intel 🤝 Angels + Pre-Seed VCs welcome

Whetstone 将 AI 候选方案与基准对比,否决回归并返回可审计的决策。
@JustinGarr90748 · X
We're building Cyberelf labs because a better score doesn't mean a better model.

向多个AI模型提问,比较答案,观察它们辩论至共识。
u/trekhleb · Reddit
I kept pasting the same question into ChatGPT, Claude, and Gemini in three tabs; so I built a Yes-Brainer — a council of AI models, that answer your question in parallel, debate to consensus, or get judged to a verdict. submitted by /u/trekhleb to r/SideProject [link] [comments]

对比和评估 AI 模型在编码、推理、代理和其他基准测试中的表现。
davidtsong · HN
Benchmarklist: track AI benchmarks (2.4k+), models, and capabilities

AI平台,提问、生成图像、语音交互,辅助更好的决策。
@Akinzoooo · X

CandrelOne 帮你创建和自动评分候选人评估。
@aishwary07jain · X

检查 AI 推理踪迹,评估模型真实性。
malik_dixon1 · Product Hunt
TraceLogicAI: AI Architecture Evaluation Compare AI architectures with evidence, not guesswork

评估AI系统是否符合负责任和道德AI实践。
@TheWhiz351 · X
I built two versions of the same responsible AI app using different vibe-coding platforms. Perplexity Computer: Base44: Try both. Which has the better design and user experience? #VibeCoding #ResponsibleAI

社区追踪AI模型体验指数,实时收集用户意见每小时更新。
schafberg · HN
Is AI Dumber Today? An index of AI model experience from user's opinion

用 LLM 评估 AI agent 对话质量,提供评分卡和成本分析。
@tech_maju · X