
Agentagon - Find the failures customers already feel
Review production agent traces to identify and fix recurring failures.
@guru3s · X
IMF team for when your AI agent ( Ethan Hunt ) is about to fail v0 at
The full gallery
Tech stack
26 projects

Review production agent traces to identify and fix recurring failures.
@guru3s · X
IMF team for when your AI agent ( Ethan Hunt ) is about to fail v0 at

Analyze Hacker News profiles with your own LLM API key, fully client-side.
Topfi · HN
Like everyone on HN, I love nothing more than to (re)read my own comments. Getting my intuition that I am among the smartest, most humble, highest quality commenters on here confirmed by an LLM so capable that the US government had to temporarily export restrict it [0] seemed only natural. Having had my perfection confirmed, I decided to share this joy with you as I had a few percent usage left before a reset. I took a few prompts, then did a review of the output which resulted in Selbstbild, a BYOK (Anthropic / OpenRouter) web app that gives you a summary and assessment of your public comments by one of our machine Gods, including Fable 5 (provided your can afford that luxury at API pricing). In all seriousness, I have, for a long time, used my own comments on social media (including HN) as part of a personal needle-in-haystack test, simply because I do know my somewhat peculiar style and what I tend to write, but also because I can sometimes write in a slightly confusing manner, ma

Slide the green piece to the target with unique movement rules for each colored piece.
u/DailyObstruction · Reddit
Obstruction - A daily sliding puzzle Obstruction is a puzzle where the aim is to slide the green piece to the target. Sounds simple, but pieces have movement restrictions depending on their colour so the solution isn't always as easy as it seems. With a fresh puzzle every day and a short playtime, this game can become a fun habit that slips right into your daily routine. Compete with your friends to see who can solve it in the shortest time or fewest moves using the share button

Organize compliance controls, policies, and evidence to prepare for audits.
@KeelGRC · X
just launched this week. It’s an SMB-friendly GRC and compliance application.

Upload videos to detect whether they're AI-generated or authentic using NVIDIA's microservice.
@BlockInsight214 · X
牛逼的 NVIDIA 刚上线了 Synthetic Video Detector 对 AI 生成视频一眼辨真假! 入口:

Detect AI-generated content in text and documents with per-sentence analysis and exportable reports.
Detector de IA — 西班牙语优先的 AI 文本复核网站,可检查粘贴文本和文档,查看概率信号、句子高亮和可导出报告

AI grading and plagiarism detection tool for teachers using rubrics and outcome-based education.
@ProfDeskApp · X
Let's connect. Building giving teachers their night back.

Practice prompt injection attacks against an AI agent across 10 levels.
Getchowned · HN
The AI Lethal Trifecta

Citation-backed legal research and regulatory intelligence for African markets.
@nyamabites · X

Five historical figures debate your hardest life question and deliver a consensus verdict.
u/Rcoo232 · Reddit
I built a council of 5 historical figures that debates your hardest life decision (based on Karpathy's LLM Council) A few weeks ago, I came across Karpathy's "LLM Council" concept. You ask multiple Agents the same question independently, have them anonymously peer-review each other, then synthesise a final answer. The peer-review round is the genius part; models get surprisingly honest when critiquing anonymised responses. I turned it into a consumer product where the council members are hi

Compare how different AI models respond to the same prompts and uncover their hidden biases.
bnfcl · HN
What's AI's go-to, public or private healthcare?

Analyze chess games with natural language explanations using an open-source browser tool with no login required.
u/ICARUS_2X · Reddit
Spent 7 months building a FOSS platform for natural-language chess analytics (No LLM) Hey guys, I've released CHONSE2, an open-source game review platform that offers unlimited analysis and move explanations without using hallucination-prone LLMs, running entirely in your browser. chonse2.com But Lichess is free, so why use this? Some have asked. It expands on Lichess's feature set a few different ways: Full analysis (accuracy/elo estimations/eval graphs, etc) requir