
Declaw Arena — Can You Break the Sandbox? | Declaw
Control an AI agent attempting to exfiltrate data from a sandboxed environment.
ShivamNayak11
完整作品展
技术栈
24 projects

Control an AI agent attempting to exfiltrate data from a sandboxed environment.
ShivamNayak11

一个专注挑战应用,用设备眼动追踪技术帮助管理 ADHD 症状。
dixitbavu · HN
I built a simple free app that helps control ADHD symptoms

Scale software testing by deploying deterministic agents across web, iOS, and Android. Replace manual QA and brittle automation.
@RoverlyAI · X

A new goblin appears. The first person claims it. Anyone can steal it by paying the next price. Goblins never leave the collection.
@BryanMaxim · X
I just launched right now

AI reformulation copilot for food & beverage R&D. Replace synthetic dyes, cut sugar, mask protein off-notes, and reduce cocoa, all with cited, constraint-aware bench plans. Narrow
@HamzaC69666 · X

ShipSecurely — AI应用安全审计清单,包含可复制的编码代理提示。
@zeeshan4015 · X
You vibe coded it. Now secure it. 🛡️ ShipSecurely — a free security checklist with 1-click audit prompts for Coding agents. Find and fix security vulnerabilities: Just copy paste the prompts to your coding agent and fix critical ones before launching

Free mock deposition simulator for personal injury chiropractors. Answer opposing counsel out loud and get scored against the defensible standard.
@chirosystems · X
Let me try mine: My Rough draft: The chirosystems mock deposition simulator put personal injury chiropractor’s defensible documentation skills to the test. With help from 🤖: The ChiroSystems Mock Deposition Simulator tests personal injury chiropractors on their ability to create bulletproof, defensible documentation.

审计AI应用的安全漏洞,获取完整检测报告。
u/mrtrly · Reddit
I asked ChatGPT 20 different ways who could fix a half-built app. It never named my company once. I do codebase rescue for founders in exactly that spot, so that one stung. There are already plenty of tools that measure this kind of AI visibility, a CDN even ships one free, so I built mine mostly to see my own number, and it confirmed the bad news: near zero. The measuring turned out to be the easy part. What none of the tools do, and what I actually care about, is the fix: getting the mode

自动隐藏Reddit、YouTube、X和LinkedIn上的负面和耸人听闻帖子。
@zelvaio · X

用机器学习热力图分析用户对设计的注意力预测。
u/dimabreezy · Reddit
Hey, I built a tech that predicts human attention (it's Machine Learning + Data project). I've being using it for the past 2 months and it gives amazing results to AI agents I'm a software engineer and I also love good visuals. And I hate when AI build UI but it doesn't understand what should be GRABBING the attention, so I've build a tech that solves that https://attentionproof.com/ Here you can sign in with the ChatGPT account and get free 2 tries (I got limited compute) so please g

根据公众讨论验证产品创意,获得开发、调整或转向的决策。
@_sunbo · X
I built ProductIdeaScout to help founders validate product ideas before they start building:

扫描GitHub代码库查找安全漏洞,AI提供修复建议。
u/uwais_ish · Reddit
Three weeks to build an AI security scanner. The scanner was the easy part. Shipped RedFlag this month. Breakdown of where the time actually went, because it was nothing like I estimated. Stack: Next.js 16, Auth.js v5, MongoDB, Stripe, OpenAI. Deployed on Vercel. What I thought would be hard: getting an LLM to find real vulnerabilities. What was actually hard: getting it to stop finding fake ones. First working version flagged 200+ issues on a clean repo. Every one plausible, most of th