
Mirrors - Test AI Agents Against a Mirror of Production
Test AI agents against production mirrors using replayed traces to find bugs and regressions.
aisinghal
The full gallery
Tech stack
19 projects

Test AI agents against production mirrors using replayed traces to find bugs and regressions.
aisinghal

Benchmark version-control systems and coding agents on realistic development tasks.
videlov · HN
I was interested in answering this question so I built a benchmark comparing git, jj and gitbutler in agentic context https://vcbench.dev/ Disclaimer - I am a co-founder of GitButler

Generate test cases and scripts automatically from requirements.
@SEOexpertpkk · X
AI-Driven Test Automation Platform

Generate test suites for Cypress and Playwright from single commands.
u/Apart_Beyond_3463 · Reddit
Got laid off, spent 2 months building a tool instead of sending more applications Got laid off from a QA role about 2 months ago. Instead of firing off more applications, I built the thing I've wanted for years and nobody made. QA Smith AI generates real Cypress or Playwright test frameworks straight from acceptance criteria. Page objects, spec files, API tests, fixtures, all of it, ready to run. It actually scans your site with a real browser first so it's not just guessing a

Build and scale systems in an interactive simulator to learn system design.
@echo_muss · X
Learn system design Try

Assess your B2B AI or SaaS startup's biggest bottleneck blocking your next milestone.
@FounderUnstuck · X
If you’re building a startup and want to identify your biggest bottleneck, try the free assessment: 🔗 We’d love to hear if the results match your experience.

AI scans competitors and market data to validate if a startup idea is worth building in 90 seconds.
@Ebrahim_Rio · X
Most founders skip validation and pray. I automated the "worth building?" check. AI scans competitors, Reddit, and market data → Cook or Kill in 90 seconds. Killed? It surfaces the pivot the data actually backs.