
Cyberelf Labs — Whetstone, a Promotion Gate for AI Systems
Whetstone 将 AI 候选方案与基准对比,否决回归并返回可审计的决策。
@JustinGarr90748 · X
We're building Cyberelf labs because a better score doesn't mean a better model.
完整作品展
技术栈
60 projects

Whetstone 将 AI 候选方案与基准对比,否决回归并返回可审计的决策。
@JustinGarr90748 · X
We're building Cyberelf labs because a better score doesn't mean a better model.

与AI面试官一起练习编程和系统设计面试。
u/redskinnypete · Reddit
I made an AI interviewer for coding and Design interviews I built an AI-powered mock interviewer that simulates a real system design interview. It listens to your approach, asks follow-up questions based on your design, challenges your decisions, and provides detailed feedback on what you did well and where you can improve. I have tried to make it feel like a real interview as much as I could to simulate the pressure and the expectations in an interview. Please give it a try. Hope you wil

对比和评估 AI 模型在编码、推理、代理和其他基准测试中的表现。
davidtsong · HN
Benchmarklist: track AI benchmarks (2.4k+), models, and capabilities

verified-3d-mesh-intersection Formally verified 3D mesh intersection - trust 93 lines of spec, not 1000+ lines of AI-written code
schildep · GitHub
verified-3d-mesh-intersection Formally verified 3D mesh intersection - trust 93 lines of spec, not 1000+ lines of AI-written code

用机器学习热力图分析用户对设计的注意力预测。
u/dimabreezy · Reddit
Hey, I built a tech that predicts human attention (it's Machine Learning + Data project). I've being using it for the past 2 months and it gives amazing results to AI agents I'm a software engineer and I also love good visuals. And I hate when AI build UI but it doesn't understand what should be GRABBING the attention, so I've build a tech that solves that https://attentionproof.com/ Here you can sign in with the ChatGPT account and get free 2 tries (I got limited compute) so please g

通过评分、护城河分析和竞争对手差距来验证 SaaS 想法的防守性。
@saralsachan · X

用AI根据具体性、证据质量和风险清晰度为创业想法评分。
@nellaiorgs · X
Building NELL Labs -- an AI-native startup validation platform. Instead of generic LLM answers, it scores your idea across specificity, evidence quality, risk clarity, and next-step usefulness, so you know if it's actually validated or just sounds good.

使用五个独立数值引擎通过REST API和MCP可靠地验证数学表达式
@UtevR57878 · X
Not vibe-coded, rather AI assisted dev:

一键生成多个 AI 模型的图片,方便对比效果。
ShotAI — 一键生成多个 AI 模型的图片,方便对比效果

使用Misata为Python生成精确声明的合成测试数据。
@RasinMuhammedX · X
Declarative Synthetic Data Generation.

ColabWize 审计学术引用并验证作者身份,证明学术工作的真实性。
@clawncore · X

AppTruth AI — 通过AI验证发现应用中的隐藏漏洞
@uriel_bitton · X
As you add more features to your vibe coded app, you increase the chances for bad behaviour. The worst part is you dont know they exist Every beta user i've had try out the app told me they found a few issues they had no idea existed Find yours using