
getjailbroken - practice AI security against an AI model and its agents
10级提示词注入谜题,尝试操纵AI代理的行为。
Getchowned · HN
The AI Lethal Trifecta
完整作品展
技术栈
31 projects

10级提示词注入谜题,尝试操纵AI代理的行为。
Getchowned · HN
The AI Lethal Trifecta

使用硬件认证和OTP保护GitHub拉取请求。
jallmann · HN
> want without a PR process that requires hardware authentication or proof of presence Just curious, what do you use for this? I built OTP Guard [1] a few years ago for exactly this problem, although I haven't seen any alternatives in the space. Does GitHub have something built-in now? The original framing was more "local malware compromising your GitHub account" ... it never occurred to me that the malware could be a LLM. I really should update the page. [1] https://otpguard.com

用你的LLM API密钥分析Hacker News公开个人资料。
Topfi · HN
Like everyone on HN, I love nothing more than to (re)read my own comments. Getting my intuition that I am among the smartest, most humble, highest quality commenters on here confirmed by an LLM so capable that the US government had to temporarily export restrict it [0] seemed only natural. Having had my perfection confirmed, I decided to share this joy with you as I had a few percent usage left before a reset. I took a few prompts, then did a review of the output which resulted in Selbstbild, a BYOK (Anthropic / OpenRouter) web app that gives you a summary and assessment of your public comments by one of our machine Gods, including Fable 5 (provided your can afford that luxury at API pricing). In all seriousness, I have, for a long time, used my own comments on social media (including HN) as part of a personal needle-in-haystack test, simply because I do know my somewhat peculiar style and what I tend to write, but also because I can sometimes write in a slightly confusing manner, ma

上传代码让 VibeGuard AI 自动检测安全风险。
@POONAMSING9999 · X
Forget this fight. The problem in AI is vibe coding security risk so i made vibe Guard ai that analysis user code and found out security risks. For the sake of humanity, to solve a painful problem I made this It had free version try now

VibeGuard AI — 分析代码安全漏洞,检查代码是否安全。
@POONAMSING9999 · X
Technology is getting better everyday so the vibe coding but still people loosing money.🚀 Around 62% of vibecode is unsecure. I made VibeGuard that analysis code and check whether it is safe,ready to go,every error and what it can cause, many more...

AI驱动的IT服务管理平台,含工单、资产跟踪和监控功能。
@wealthyness_ · X
Your Multi ITSM allowing you to serve other organizations as the IT team. It can be used internally by the IT team as well

用Semfora分析代码库结构,发现变更中的风险。
u/jeremyStover · Reddit
My second paying customer bought a full year subscription. I have mentioned https://semfora.ai before, here and there on Reddit. I haven't talked much about what it really provides. I was hoping to drive up some organic traffic, from people that had the motivation to dig past the beta. Stupid idea, but now that we have passed pen tests and hit the minimum requirements check list for security audits, (see other threads about that) we opened up the beta to everyone. Two days after that(

扫描AI提示词和端点的漏洞,实时监控生产LLM的安全性和合规。
@CognisafeUK · X

跨提供商治理、追踪和防护 AI 代理的安全平台
camsjams · HN
Lineation – One security control plane for all agents

探索2020年以来的CVE趋势和漏洞严重性,按报告组织分类。
u/Secret_Appeal6271 · Reddit
Agents are more capable, and susceptible to exploits, than ever. We're working to stop this from hurting users. AI agents are starting to get real access like GitHub tokens, cloud credentials, customer data, deploy permissions. Not coincidentally, the rate of major cybersecurity incidents is rising rapidly. See for yourself: https://epoch.ai/data/cve?view=graph https://genai.owasp.org/resource/state-of-agentic-ai-security-and-governance/ My friend and I, both AI researchers, are working o

监控 SSL 证书和网站安全,包含自动续期检测、安全头部监控和 DNS 健康检查。
@di_spivak · X

在发布前检查AI生成内容中的错误和安全问题。
u/Brief_Dust8845 · Reddit
I pivoted from my initial idea after realizing I was solving the right problem at the wrong time When I started building GaaS Guard, it was an AI governance tool for companies. The idea was to help organizations defend against prompt injection and unsafe AI interactions. It was technically interesting, and I still genuinely believe I was solving a real problem. The problem was, it just wasn’t selling—to be brutally honest. Here’s how I actually ended up pivoting. I started using a b