
知行录 · leaderboard.cn
在 leaderboard 上按官方基准对比 AI 大模型的性能排名
fcten · V2EX
做了一个大模型 leaderboard 网站 最近一个月 CodeX 疯狂送重置,token 根本用不完,顺手做点东西。 地址:[知行录]( https://leaderboard.cn/) 排行依据主要为模型官方基准测试成绩。非主观排名。 数据会持续更新。如果有点用,欢迎各位 v 友收藏~
完整作品展
技术栈
60 projects

在 leaderboard 上按官方基准对比 AI 大模型的性能排名
fcten · V2EX
做了一个大模型 leaderboard 网站 最近一个月 CodeX 疯狂送重置,token 根本用不完,顺手做点东西。 地址:[知行录]( https://leaderboard.cn/) 排行依据主要为模型官方基准测试成绩。非主观排名。 数据会持续更新。如果有点用,欢迎各位 v 友收藏~

在空间节点画布上协作AI项目,组织语言模型上下文。
jebuehler55 · HN
I built a spatial node canvas to fix LLM context drift

复刻老黄历解压网站 简介:一个复刻纸质老黄历的网站,还可以解压地撕撕撕,哈哈哈 网站地址: https://11d62f9d.pinme.dev/
mqx · V2EX
复刻老黄历解压网站 简介:一个复刻纸质老黄历的网站,还可以解压地撕撕撕,哈哈哈 网站地址: https://11d62f9d.pinme.dev/

连接你的数据库,用自然语言提问获取SQL、图表和数据洞察。
canmeng · V2EX
Vibe 了一个 Agent 桌面端 https://datyo.ai 欢迎试用

向任何LLM发送前屏蔽敏感数据,然后在回复中恢复。
@velumprivacy · X

统一API网关,审计Token消耗并分析多个AI提供商的成本。
@amiuchat · X
我做了一个桌面工具:Token Switch。 给 Codex / Claude Code / OpenCode 重度用户统一管理 Provider、模型、Token、Base URL、连通测试和消耗统计。 多个 Agent,一个 Token 工作台。欢迎试用,也欢迎吐槽你最烦的配置问题。

通过 RavenGate 网关路由 LLM API 流量,追踪成本、分析延迟、隐蔽 PII。
charltonraven · HN
RavenGate – LLM gateway that redacts PII across SSE chunk boundaries

用你的LLM API密钥分析Hacker News公开个人资料。
Topfi · HN
Like everyone on HN, I love nothing more than to (re)read my own comments. Getting my intuition that I am among the smartest, most humble, highest quality commenters on here confirmed by an LLM so capable that the US government had to temporarily export restrict it [0] seemed only natural. Having had my perfection confirmed, I decided to share this joy with you as I had a few percent usage left before a reset. I took a few prompts, then did a review of the output which resulted in Selbstbild, a BYOK (Anthropic / OpenRouter) web app that gives you a summary and assessment of your public comments by one of our machine Gods, including Fable 5 (provided your can afford that luxury at API pricing). In all seriousness, I have, for a long time, used my own comments on social media (including HN) as part of a personal needle-in-haystack test, simply because I do know my somewhat peculiar style and what I tend to write, but also because I can sometimes write in a slightly confusing manner, ma

聊天对比 AI 模型,投票参与排行榜排名。
u/Rabus · Reddit
I got TestingModels too overcomplicated over the month it is running: looking for some feedback how to make it more useful and simpler I run a benchmark like arena.ai , but with pre-generated prompts. So far, nearly 6k people came in and like 30k comparisons has been made - which means the thing is genuinely useful for people to compare the models. The problem is the more features i started adding the more overblown and complicated UI became - like old internet explorer tab bars Old: ht

多智能体LLM系统的可视化编辑器,支持本地推理。
sascha10000 · HN
Multi-agent LLM editor with local inference via WebSockets


AI 驱动的八字排盘系统,采用确定性规则引擎计算,LLM 生成自然语言分析。
@GoWithFlow2026 · X
上线了一个 AI 八字排盘 🔮 🔗 和"直接把生日扔给 ChatGPT"完全不同: 1️⃣ 事实不靠猜:排盘、干支关系、刑冲合会等,全部由确定性规则引擎计算——LLM 只负责把判定结果写成人话,编不了盘。 2️⃣ 指令手写:分析指令是我根据《子平真诠》并结合自身经历逐条手写的命理逻辑。 🎁 免费演示,无需注册,点开即看完整报告: 觉得准,再测自己的($1.99 起)。欢迎反馈!