
Vulnsy - Pentest Reporting Platform for Security Teams
Platform for security teams to create and deliver penetration test reports faster with reusable findings and client portals.
@TurvSec · X
The full gallery
Tech stack
27 projects

Platform for security teams to create and deliver penetration test reports faster with reusable findings and client portals.
@TurvSec · X

Compare AI language models by performance across official benchmarks.
fcten · V2EX
做了一个大模型 leaderboard 网站 最近一个月 CodeX 疯狂送重置,token 根本用不完,顺手做点东西。 地址:[知行录]( https://leaderboard.cn/) 排行依据主要为模型官方基准测试成绩。非主观排名。 数据会持续更新。如果有点用,欢迎各位 v 友收藏~

AI-powered code review that runs code in microVMs to catch more bugs.
u/dumbfoundded · Reddit
Ito, AI Code Review that Runs Code I've been using AI code review tools but none of them actually run code so I built one: https://www.ito.ai/ The way it works is that it uses microVMs to spin up your environment with all of the services running. Then a bunch of AI agents go and test the application to collect runtime evidence. The result is you get test cases along with evidence about whether or not the test cases pass or fail. The runtime evidence can be videos, request/response curls, db

Build and run ML workflows in an agent-native cloud notebook with evaluation tools.
eldar_hsnv · HN
Show HN: AI Notebook for Data Science – Kind of Like Cursor but for Jupyter

Add evaluation reports to your AI agent with a shareable URL that scores performance.
adeeonline · HN
AgentsProof – a small project for testing AI agents

AI tool that validates startup ideas, finds competitors, and suggests business pivots.
@ArchieHashani · X

Check AI agent outputs and actions against your policies before execution; get allow, block, rewrite, or escalate decisions.
u/danielbaker06072001 · Reddit
Most AI agent SaaS is just bad security with a nice dashboard Give a new employee access to Stripe, GitHub, Slack, and your CRM on day one, and you’d call it reckless. Give the same access to an AI agent, and we call it “autonomous.” That isn’t innovation. It’s skipping basic security because the demo looks cool. Full disclosure: I’m building TrustLoopGuard around this problem, so I’m obviously biased. While testing one MCP connection, I realized I had decided what the agent could a

Observability platform for AI agent pipelines that detects failures and explains root causes.
@VaraadDurgaay · X
Solving the prb of observability in ai agents

AI-powered security scanner for AWS that identifies and remediates vulnerabilities automatically.
@GlenLouis08 · X
lets connect! building this

Tests LLM endpoints with adversarial cases and provides OWASP-mapped security audit reports.
@aryaan_sheth · X
- LLM security for small teams

Automated security audits for GitHub repositories to catch secrets, auth issues, and deployment problems.
@emanueldev4 · X
Building Preflight a simple way to catch issues before your website goes live.

Track trending AI/ML repositories on GitHub with real-time alerts and commit analytics.
@g_saikumar_ · X
Real-time GitHub intelligence for developers, founders, and open-source enthusiasts.