
知行录 · leaderboard.cn
在 leaderboard 上按官方基准对比 AI 大模型的性能排名
fcten · V2EX
做了一个大模型 leaderboard 网站 最近一个月 CodeX 疯狂送重置,token 根本用不完,顺手做点东西。 地址:[知行录]( https://leaderboard.cn/) 排行依据主要为模型官方基准测试成绩。非主观排名。 数据会持续更新。如果有点用,欢迎各位 v 友收藏~
完整作品展
技术栈
29 projects

在 leaderboard 上按官方基准对比 AI 大模型的性能排名
fcten · V2EX
做了一个大模型 leaderboard 网站 最近一个月 CodeX 疯狂送重置,token 根本用不完,顺手做点东西。 地址:[知行录]( https://leaderboard.cn/) 排行依据主要为模型官方基准测试成绩。非主观排名。 数据会持续更新。如果有点用,欢迎各位 v 友收藏~

检测API中转站输出是否与官方100%一致
@nodeloc_cc · X
🌈 7月,你好,MODELOC上线算力池。 MODELOC自上线以来,已检测2000余次,覆盖600+中转站,为众多AI用户提供的使用参考。 MODELOC近期进行了改版,上线了算力池及市场。 加入算力池 查看帖子: 用 MODELOC 便宜地调各家大模型:一次讲清它的价格体系

Ask across every earnings call at once and get the quotes back, with the speaker and the call behind each one. Figures come from a query, never from a model.
@serkanglatt · X
Curious about earnings calls? Fire away. Plans start at $24.99 a month. unlocks 250,000+ calls from 12,000+ companies across all 11 sectors, going back to 2020. Fresh calls get added within minutes of publication. Ask the questions an analyst would: growth, margins, guidance, sentiment. Put two companies head to head. Track one metric across quarters. The numbers are pulled straight from the transcripts. The quotes are sourced. And when the data has no answer, the AI tells you instead of guessing.

在浏览器中运行AI模型基准测试以检测性能回归。
pepperpoppins · HN
Trunchbull, run real models against any benchmark in your browser

Echo – Fable-level results at 1/3 the cost using open-weight models
adam_rida · HN
Echo – Fable-level results at 1/3 the cost using open-weight models

使用七个AI模型生成和比较报价、邮件和广告文案。
@MarkZofMarkZ · X
Let's connect! Profit Router 📈

Disidea:四个顶级 AI 模型同时辩论你的决策问题。
@painterner · X
After a year away, I’m back to indie hacking. Built with Claude and GPT in two weeks, Disidea is my first launch of the year: four AIs debate your decisions in one thread. Try it: Questions? Ideas? Let’s build great products and make money. Who’s in?


从多个平台收集和分析反馈到一个仪表板。
@GeorgiG26929167 · X
Remarkd lets you create a feedback page for anything—ideas, landing pages, designs, pricing, features or prototypes. Share one link anywhere and collect structured, anonymous feedback in one place, with AI-powered insights to help you spot patterns faster

对比 AI 模型在编码任务上的表现,支持成本追踪和 ELO 排名。
@intheworldofai · X
On the World of AI Bench (vibe-coding composite): Claude Fable 5 → 85.2 GPT-5.6-sol → 82.4 kimi-k3 → 81.5 Moonshot’s K3 just walked in and claimed bronze on one of the toughest coding-focused leaderboards out there.

通过一个API访问多个AI模型,提供透明的预付费定价。
@MyApiTaco · X
GLM 5.3 Flash is now on 🌮⚡ Limited-time promotion: 66.6% OFF retail • Input: $0.05 (retail $0.15) • Cache Input: $0.01 (retail $0.03) • Output: $0.167 (retail $0.50) Plus, get an extra 5% bonus on topups over $100. #GLM #zAI #openrouter #vibecoding

通过统一 API 访问 200+ AI 模型(Claude、GPT、Gemini),支持自动故障转移。
@MixRoute_ai · X