
Tokenstead - Find AI Models for Your Hardware
查找与您硬件兼容的AI模型并查看性能和价格估计。
cdnsteve · HN
Tokenstead, find AI models for your hardware
完整作品展
技术栈
60 projects

查找与您硬件兼容的AI模型并查看性能和价格估计。
cdnsteve · HN
Tokenstead, find AI models for your hardware

根据已发布的代理标准评判AI产品,包含可检查证据和社区投票。
@katyorby · X
i built — a local receipt for claude code runs. your check says whether the workspace passes now; the transcript supplies the activity counts. no transcript upload and no magical autonomy score.

多模型事实检验API,在部署前验证AI输出的准确性。
kostaj · Product Hunt
Lenz Independent, multi-model fact-checking API for AI workflows

为你的AI代理添加评估报告,生成可分享的URL展示性能。
adeeonline · HN
AgentsProof – a small project for testing AI agents

查看和对比主流AI模型的公众意见和基准评分。
u/TasteMysterious5285 · Reddit
I built AI Census, a live field bulletin for how people are actually talking about AI models I’ve been building AI Census, a public “field bulletin” for how people are talking about current AI models. I kept running into the same problem: benchmark tables tell me how a model performs on a test, but not whether people are actually finding it useful, frustrating, reliable, etc. So I built a rolling view from public technical conversations across Reddit, Hacker News, Bluesky, GitHub, and Huggi

在浏览器中运行AI模型基准测试以检测性能回归。
pepperpoppins · HN
Trunchbull, run real models against any benchmark in your browser

通过单一API访问领先的AI模型,透明令牌定价。
DustinPham12 · HN
1endpoint – Cheaper access to AI models

Access leading discounted AI models from OpenAI, Anthropic, and DeepSeek through one unified API without changing your request format.
@Orbiqen · X
Access GPT, Claude, DeepSeek, and image models from one API. Up to 90% less than official prices on selected models. OpenAI-compatible. Works with Cursor and Claude Code. Try Orbiqen from $1:

Compare today's leading AI models by price, intelligence and more.
@spectragai · X

在OpenVibeEval中对比不同AI模型生成前端代码和可访问性评分。
u/12qwww · Reddit
I built a live benchmark to see which AI actually writes the best frontend code Hey everyone! I built OpenVibeEval because I was tired of "vibe-checking" AI-generated frontend code. I wanted to know which model actually produces the most accessible and clean React/Tailwind output. What I built: •A leaderboard of 24 models (Claude, GPT, DeepSeek, etc.) ranked by axe-core accessibility scores. •A Harness Comparator to show how different system prompts change the same model's output. •

用多个模型实时审计AI回应以判断其可靠性。
u/inc_23 · Reddit
Hey, I created a tool that catches when your LLM is confidently wrong, in production, in real time — looking for beta testers. Your bot sounds sure of itself even when it's wrong, and you usually only find out when a customer complains. Auscope audits every LLM response in the background: 3 models from 3 different providers independently check it, a 4th "chairman" model resolves disagreements, and you get one verdict — verified, uncertain, or unreliable. Runs async, doesn't slow your respon

Lomote:AI 创意工作室,可生成、增强和编辑图像、视频及 3D 模型。
@Agueroyps3 · X
折腾了一段时间,我做的 Lomote 终于上线了。 想解决的问题很简单:生成图片、视频、3D 模型,做图片或视频增强、搜索相似图片时,不用在一堆工具之间来回切换。 在 Lomote 里,这些能力都可以通过工作流组合起来。 感兴趣可以试试: