@BobbySBX · X
We just dropped Assay v0.2.6, the open source evaluation engine powering Most registries judge an AI tool by its README. We built Assay to actually run skills, MCP servers, and plugins inside isolated sandboxes with live models to observe what they do before you install them. What is new in v0.2.6: - NVIDIA API Catalog provider: Native support is here, allowing you to use NVIDIA driver and judge models to run behavioral benchmarks and multi-tier audits. - Zero runtime dependencies: Still a self-contained ~240 KB tarball that speaks directly to @nvidia , @AnthropicAI , @OpenAI , @OpenRouter , and local HTTP endpoints without SDK bloat. - Adversarial and static checks: Out of the box verification against prompt injection, tool scope creep, destructive shell commands, and supply chain drift. Every tool on is tested through Assay. That means 4,200+ skills, MCP servers, and plugins are commit-pinned, graded on actual runtime behavior, and rea