
LLM Leaderboard | Redactle
对比语言模型在 Redactle 谜题上的表现排名。
pampas · HN
Redactle LLM Leaderboard
完整作品展
技术栈
60 projects

对比语言模型在 Redactle 谜题上的表现排名。
pampas · HN
Redactle LLM Leaderboard

在OpenVibeEval中对比不同AI模型生成前端代码和可访问性评分。
u/12qwww · Reddit
I built a live benchmark to see which AI actually writes the best frontend code Hey everyone! I built OpenVibeEval because I was tired of "vibe-checking" AI-generated frontend code. I wanted to know which model actually produces the most accessible and clean React/Tailwind output. What I built: •A leaderboard of 24 models (Claude, GPT, DeepSeek, etc.) ranked by axe-core accessibility scores. •A Harness Comparator to show how different system prompts change the same model's output. •

The end-to-end platform for small language models. Tuned to your task, a small open model matches frontier accuracy at a fraction of the cost. Own your intelligence: private, compa
@dayoffdev · X

用你的数据微调定制语言模型,支持密码学删除证明。
@MBrew26730 · X
Dataset cleaning + fine tuning + continual learning at

比较多个语音转文字引擎的速度和准确度,支持本地隐私保护。
@alvaisy · X
finished voice to text small web app for my own itch. it's opensource. use openrotuer key. and use it with 4 models.

构建或自动生成具有逻辑、支付、签名和集成等功能的多语言表单。
@hkbonur · X
that digitalise paper forms and make everything multi language, so businesses can get better conversion. One form any language.

通过 Vocab Top 的 AI 生成视觉助记、语境翻译和发音指南掌握词汇。
LandoLorinse · HN
Vocab Top – AI-powered vocabulary builder that helps you retain words

与Amália AI聊天,由欧洲葡萄牙语开源模型驱动。
wbemaker · HN
I Am Hosting Amalia – The First Portuguese LLM

可视化语言模型在各层回答前的思考内容。
ada1981 · HN
I built a web tool to see and edit what an AI thinks before it answers

让AI模型通过3D动画展现香蕉植物的完整生命周期来比较性能。
fran-mora · HN
I gave 5 AI coding agents one prompt: grow a banana plant through its whole life in three.js: sprout, leaves, flower, fruit, rot, then pups that restart the loop. It's deceptively simple and yet very hard to get right from procedural code: you have to write working three.js and understand how the plant is actually built; how it hangs, ages and decays. Get the biology wrong and the code renders something weird. These are agents, not bare models (Claude Code and Codex for now). They can use tools, including playwright to check their work and improve it.

提交 LLM 推理优化内核,在专用硬件上进行基准测试并竞争排名。
carsenk · HN
Frontier.fast – Help push the frontier of LLM speed forward

Semantic Overlays 通过适配器修改冻结语言模型对标记文本的感知。
joshua_s_penman · HN
Semantic Overlays – an NX bit for LLM prompt injection (live demo)