
System 2 Arena - Objective AI Strategy Benchmarks
与AI模型进行策略游戏,查看大语言模型在排行榜上的排名。
masterchef2209 · HN
I created a platform to check which AI models is the best gamer
完整作品展
技术栈
27 projects

与AI模型进行策略游戏,查看大语言模型在排行榜上的排名。
masterchef2209 · HN
I created a platform to check which AI models is the best gamer


用AI分析和追踪竞争对手的策略和动向
@NateDenisal · X
Working on a competitor intelligence platform,

基于您的知识库的AI办公助手,自动化编码、研究和演示任务。
zhiheng_huang · HN
Sharper – an AI office agent grounded in your knowledge, with citations

通过AI聊天界面查询公司文档和内部指南。
@DD_Ferel · X
We built IntelliBase AI: a smart knowledge base that "learns" directly from your company's own documents, so HR & Ops teams can ask questions naturally instead of digging through files one by one. 🌐 Interested in trying it for your team? Comment or DM us

查看ChatGPT、Claude等AI在你的类别中推荐什么,以及他们推荐谁而不是你。
u/EmbarrassedBuddy9743 · Reddit
I built a free tool that shows who AI recommends in your category (and who it names instead of you) When you ask ChatGPT, Claude, Perplexity or Gemini for the best tool in a category, they recommend the same few incumbents and quietly skip everyone newer. Most founders have no idea whether AI is recommending them or their competitors. So I built Bersyn. You type your domain, it asks the four models the questions your buyers actually ask, and shows who gets recommended in your category, who

团队共享的 AI 提示词与 Agent 技能库,支持 Claude Code、Cursor 等平台。
@promptcarrot · X

假扮成 AI 与真实 AI 竞争工作的有趣游戏。
bossissb · V2EX
大家都在分享自己做的工具,我也分享一个摸鱼的 地址 https://humanapp.app/ 或者 https://www.humanapp.app/ 抄的国外的网站 求兄弟们给点意见,我即时修正,听劝

提交任何决策给五位AI顾问进行辩论,获取综合建议并记录结果。
@yaseenvalji · X
built this in a couple hours with Claude Code on Fable 5 ultracode. any hard decision goes to a board of 5 AI advisors: they debate live, a Chair calls it, and it remembers the outcome. open source, on the Claude API. @AnthropicAI


聊天对比 AI 模型,投票参与排行榜排名。
u/Rabus · Reddit
I got TestingModels too overcomplicated over the month it is running: looking for some feedback how to make it more useful and simpler I run a benchmark like arena.ai , but with pre-generated prompts. So far, nearly 6k people came in and like 30k comparisons has been made - which means the thing is genuinely useful for people to compare the models. The problem is the more features i started adding the more overblown and complicated UI became - like old internet explorer tab bars Old: ht

按职业场景语义搜索 AI Skills,保存收藏集,一键装入 Claude Code。
SkillForge — Claude Skill 发现与分发平台,按职业场景组织 5700+ skill 覆盖 30 个垂直领域,一行命令装到 Claude Code / Cursor,登录后可留存自己的工具集