
Maina Voice — Speech to Text & Model Benchmarking
比较多个语音转文字引擎的速度和准确度,支持本地隐私保护。
@alvaisy · X
finished voice to text small web app for my own itch. it's opensource. use openrotuer key. and use it with 4 models.
完整作品展
技术栈
28 projects

比较多个语音转文字引擎的速度和准确度,支持本地隐私保护。
@alvaisy · X
finished voice to text small web app for my own itch. it's opensource. use openrotuer key. and use it with 4 models.

对比 AI 模型在编码任务上的表现,支持成本追踪和 ELO 排名。
@intheworldofai · X
On the World of AI Bench (vibe-coding composite): Claude Fable 5 → 85.2 GPT-5.6-sol → 82.4 kimi-k3 → 81.5 Moonshot’s K3 just walked in and claimed bronze on one of the toughest coding-focused leaderboards out there.

让AI模型通过3D动画展现香蕉植物的完整生命周期来比较性能。
fran-mora · HN
I gave 5 AI coding agents one prompt: grow a banana plant through its whole life in three.js: sprout, leaves, flower, fruit, rot, then pups that restart the loop. It's deceptively simple and yet very hard to get right from procedural code: you have to write working three.js and understand how the plant is actually built; how it hangs, ages and decays. Get the biology wrong and the code renders something weird. These are agents, not bare models (Claude Code and Codex for now). They can use tools, including playwright to check their work and improve it.

在历史市场数据中测试和评估 AI 交易代理。
remote_ctrl · HN
BotTrade – a replayable benchmark for autonomous trading agents

PRcade 通过团队排行榜和分析可视化GitHub代码审查性能
u/SnooStrawberries827 · Reddit
my team had 47 open PRs and nobody was reviewing them, so I gamified it our team hit 47 open PRs at one point last month and nobody was reviewing them. tried slack reminders, deadlines, rotating reviewers, none of it really stuck. might be related to the fact that everyone's hyped about how fast AI can write code now, copilot cranking out entire features in hours, but none of that matters if the PR just sits there for a week. feels like writing code stopped being the bottleneck a while back

CraftScore 使用Git分析和代码质量指标衡量工程技能。
ricardo_santis2 · Product Hunt
CraftScore Measure engineering craft, not AI output.

One bundle. Three independent judges. Reproducible rankings.
@Imranmohsin18 · X
sites I’m working on: Please check out and give me feedback all are public showcasing my skills :D ( not selling anything atm )


审计网站性能、SEO 问题和 Core Web Vitals。
@_ariesblaze · X
1. SEO & Websites Audits. 2. Invoice mgt. 3. Event and Ticket Mgt

实时可视化硬件在运行LLM推理时的性能指标
dev_dan_2 · HN
WatchMachineGo – A visualizer to show hardware performing LLM inference

收集产品反馈并附带截图、音频和会话回放,便于团队分类和追踪。
u/Zealousideal_North83 · Reddit
Fidibeki — Product feedback that finishes the job Building products is what drives me. While working on my own products, I noticed the same problem: feedback comes in, loses context, and users rarely hear what happened next. So I built Fidibeki. Users can report issues directly inside a product with screenshots, audio, or session replay. I can triage everything in one inbox, merge duplicates, track progress, and keep reporters updated. Privacy is built in: session replay

用AI根据具体性、证据质量和风险清晰度为创业想法评分。
@nellaiorgs · X
Building NELL Labs -- an AI-native startup validation platform. Instead of generic LLM answers, it scores your idea across specificity, evidence quality, risk clarity, and next-step usefulness, so you know if it's actually validated or just sounds good.