
Trunchbull
在浏览器中运行AI模型基准测试以检测性能回归。
pepperpoppins · HN
Trunchbull, run real models against any benchmark in your browser
完整作品展
技术栈
60 projects

在浏览器中运行AI模型基准测试以检测性能回归。
pepperpoppins · HN
Trunchbull, run real models against any benchmark in your browser

查看LLM模型在10个基准问题上的评分和排名。
fristovic · HN
She watched me look at model rankings and asked what do the numbers mean... I literally had no good way of explaining it to her so I just came up with something that is approximately in the same ballpark as some of the benchmarks out there lol

一次一个句子阅读全本书籍,包含经典著作和翻译支持。
@ReadSentence · X

从文本生成带情感的AI语音,支持59种语言。
@prosodyai · X
ProsodyAI — AI text-to-speech that performs, not just pronounces. Tell it how to sound ("warm, slow, like a bedtime story") and it does. Three engines in one app: ours, OpenAI and Gemini. 100,000 characters free every month, voice cloning included.


用0-100 AGI分数对标前沿AI模型的基准性能。
baraklaniado · HN
I audited my AI leaderboard scale – every score dropped 6-15 points

用自然语言查询电子表格和数据集,实时生成答案和报告。
u/maybeImakemoney · Reddit
I built the thing. Now I am not sure the base use case is one people will pay for. Founder here. This started as a side learning project to see whether an LLM could answer questions about Excel data, back when they could not do it well. I built the first version on n8n, with workflows that ingested files, generated metadata with an LLM, and answered questions against the converted data plus that metadata. Then I started using it for my own analysis and report generation, saw that the time sav

通过统计语法推理探索 Voynich 手稿的候选翻译和词类。
@geeky_gamer25 · X
Check out what I just built with Lovable!

免费AI转录,支持22种印地语和19种全球语言及本地文字。
@opjhabuilds · X
Transcribe any and every audio or video you have, and make it part of your AI brain and knowledge memory.

录制多语言会议,自动生成记录、摘要和行动项
@aunahmedm · X
Samjha — AI meeting notes for 600M+ multilingual professionals. Urdu. Hindi. Punjabi. Bengali. English. US. UK. UAE. South Asia. 91% accuracy on mixed-language meetings. Free beta.

跨14个维度分析Python代码,检测违规并提供详细报告。
@KSFirasa · X
Hello! I built a tool that profiles code (python only atm) across 14 dimensions detecting violations and capabilities outputting a full report. A bit more nuanced than "AI-powered insights". Free while in beta. Thank you!

浏览每小时更新的社区 AI 模型性能评分。
schafberg · HN
Is AI Dumber Today? An index of AI model experience from user's opinion