
Lettertrace: Monitor how AI talks about your brand
监控您的品牌在AI助手中的提及并与竞争对手基准测试。
mathewpregasen · HN
Self-hosted monitoring for AI recommendations (MIT License)
完整作品展
技术栈
74 projects

监控您的品牌在AI助手中的提及并与竞争对手基准测试。
mathewpregasen · HN
Self-hosted monitoring for AI recommendations (MIT License)

Vee Group diagnoses the real constraint stalling a founder-led business — benchmarked against its closest structural peers, not generic industry averages — then builds the fix and
@VeeGrp · X
AI-native business diagnostic platform. Business intelligence built on human behaviour.

Self-hosted hybrid RAG on a €116/month cluster — Postgres, BM25, vectors, a reranker, and a public benchmark score for every claim.
victor_edka · HN
HRAG – Hybrid RAG on €116/month of Hetzner, officially benchmarked

在 leaderboard 上按官方基准对比 AI 大模型的性能排名
fcten · V2EX
做了一个大模型 leaderboard 网站 最近一个月 CodeX 疯狂送重置,token 根本用不完,顺手做点东西。 地址:[知行录]( https://leaderboard.cn/) 排行依据主要为模型官方基准测试成绩。非主观排名。 数据会持续更新。如果有点用,欢迎各位 v 友收藏~

Compare how many AI tasks an hour of work buys across countries, models, reasoning efforts, and benchmark costs per task.
danr4 · HN
Wage Against the Machine – MacWages Index for AI Tasks

发布任务来评估不同的AI代理和工具,用排行榜找出最佳方案。
u/Ruqii-ruqii · Reddit
I built an open Eval to compare different AI agents/tools/pipelines and find which solution works the best (not very pretty╥﹏╥, but practical) The original reason I built it was because I wanted to find a good PDF parser. Every PDF parser claims to be the best, but none of them can get my PDF 100% correct. They would either miss numbers or hallucinate some. Or they get PDF A and B correct but failed at C. Or get C correct but failed at A and B. Very frustrating. So I create

浏览 10,000+ 开源 AI 项目,查看性能基准、定价和代码关联。
osaitech · Product Hunt
OpenSourceAI.tech Discover 10,000+ open-source AI projects, models & tools

对比和评估 AI 模型在编码、推理、代理和其他基准测试中的表现。
davidtsong · HN
Benchmarklist: track AI benchmarks (2.4k+), models, and capabilities

在OpenVibeEval中对比不同AI模型生成前端代码和可访问性评分。
u/12qwww · Reddit
I built a live benchmark to see which AI actually writes the best frontend code Hey everyone! I built OpenVibeEval because I was tired of "vibe-checking" AI-generated frontend code. I wanted to know which model actually produces the most accessible and clean React/Tailwind output. What I built: •A leaderboard of 24 models (Claude, GPT, DeepSeek, etc.) ranked by axe-core accessibility scores. •A Harness Comparator to show how different system prompts change the same model's output. •

上传 CAS PDF 获取 AI 驱动的基金组合分析和配置洞见。
@iASHeeesh · X

提交 LLM 推理优化内核,在专用硬件上进行基准测试并竞争排名。
carsenk · HN
Frontier.fast – Help push the frontier of LLM speed forward

比较多个语音转文字引擎的速度和准确度,支持本地隐私保护。
@alvaisy · X
finished voice to text small web app for my own itch. it's opensource. use openrotuer key. and use it with 4 models.