One’s Vibe

项目详情 · 开发者工具

Sakura — Benchmark for Local Coding Models

用Sakura基准测试本地编码模型,测量准确性、延迟和吞吐量。

sakura.vaansh.dev
Sakura — Benchmark for Local Coding Models — screenshot
试用演示 在线检查通过14 hours ago

上手第一步: 查看实时排行榜上的编码模型基准测试结果。

开源

GitHub: ★ 2

技术画像

构建者
Solo maker
耗时
A weekend
开源

站点健康

🛠 有 3 条站点健康建议在等作者——认领本项目(X 登录)即可查看。

来源

r/SideProject

u/Unfair_Association89 · r/SideProject
I built a reproducible benchmark for local coding models (Ollama, 27 tasks, live leaderboard) ran it on my 8GB card, here's what I found I kept eyeballing "vibes" to decide whether one quant of a coding model was actually better than another on my machine, so I built Sakura to get real numbers instead. What it does: - Points at any Ollama model and runs it through 27 hand-curated tasks: codegen, bugfix, SQL, refactor, systems design, protocol implementation, and terminal-agent episode

致作者:挂上铭牌,认领项目

这个项目还没有作者认领——登录后把铭牌挂到你的站点上,即可完成认领。

评论

Sign in to comment.

  • No comments yet. Be the first.

继续探索

相似项目

登录 后可举报该项目的问题。