
Cinchor — Verify a real agent decision
Verify an AI agent's decisions and test tampering with the verification records.
foh_quarters · HN
Verify what an AI agent did, then tamper with the record (no signup)
The full gallery
Tech stack
60 projects

Verify an AI agent's decisions and test tampering with the verification records.
foh_quarters · HN
Verify what an AI agent did, then tamper with the record (no signup)

Compare and evaluate AI models across coding, reasoning, agents, and other benchmarks.
davidtsong · HN
Benchmarklist: track AI benchmarks (2.4k+), models, and capabilities

Receive AI decision scores for every trade based on emotion tracking and behavioral discipline analysis.
@eialgos · X
AI/ML builder here 👋 I'm building EI ALGOS, a Decision Intelligence platform that combines machine learning, behavioral analytics, options analysis, and technical evaluation to help traders make higher-quality decisions. Visit us at Happy to connect with other founders and builders.

Ask one question to multiple AI models, compare their answers, and watch them debate to consensus.
u/trekhleb · Reddit
I kept pasting the same question into ChatGPT, Claude, and Gemini in three tabs; so I built a Yes-Brainer — a council of AI models, that answer your question in parallel, debate to consensus, or get judged to a verdict. submitted by /u/trekhleb to r/SideProject [link] [comments]

View and compare public opinions and benchmark ratings for leading AI models.
u/TasteMysterious5285 · Reddit
I built AI Census, a live field bulletin for how people are actually talking about AI models I’ve been building AI Census, a public “field bulletin” for how people are talking about current AI models. I kept running into the same problem: benchmark tables tell me how a model performs on a test, but not whether people are actually finding it useful, frustrating, reliable, etc. So I built a rolling view from public technical conversations across Reddit, Hacker News, Bluesky, GitHub, and Huggi

Give your AI product a sense of user taste through preference learning.
@marcellafjacob · X

AI-powered business intelligence platform with agentic analysis and AI readiness assessment.
@ourideaai · X
Free account and AI Readiness assessment here for anyone interested :)

Track community feedback on AI model performance, updated with user opinions every hour.
schafberg · HN
Is AI Dumber Today? An index of AI model experience from user's opinion

Scores AI-generated ad creative and returns verdicts (run/fix/kill) via MCP and REST API.
ds246 · HN
Spendict – a performance marketer's verdict for AI agents, over MCP

Trace, replay, and verify AI agent decisions with signed audit trails.
u/Funky_Chicken_22 · Reddit
OSS to SaaS positioning problem: when the user persona and the buyer persona are completely disjoint Founder here. Sharing a positioning problem I think a lot of OSS-to-SaaS founders hit and don't talk about publicly. Context: I have been running an OSS project (world-model-mcp) with ~2,500 monthly PyPI installs. Two weeks ago I opened up the hosted companion, Etch, at etch.systems. Launched publicly on Product Hunt at 12:00 PDT yesterday. The positioning problem: OSS user persona: in

AI assistant helping enterprise employees preserve and access institutional knowledge and judgment.
@romanbodnarchuk · X
Check out what I just built with Lovable!

Verify AI agent decisions locally with transparent reasoning and MCP integration.
brandon_grutkowski · Product Hunt
Aurora Glass-box Quantitative AI for Humans and Agents