
Targe — LLM security scanning & audit reports
用对抗测试检查LLM端点安全,获取OWASP审计报告。
@aryaan_sheth · X
- LLM security for small teams
完整作品展
技术栈
60 projects

用对抗测试检查LLM端点安全,获取OWASP审计报告。
@aryaan_sheth · X
- LLM security for small teams

按主题分组聊天记录,快速导航到过去的对话。
doantam · HN
I built a chat client that uses embeddings to cluster messages by topic

缝纫机竞技游戏,争登排行榜。
dkhcyx · V2EX
欢迎大家体验我做的发人深省,劝人向善的精品游戏 https://cassiangroup.uk/sewing/ https://cassiangroup.uk/salvate/ 最近 codex 重置太多了,蹬不完的 token 拿来蹬游戏了 期待大家荣登排行榜



开源提示词压缩,在API调用前压缩输入以降低LLM成本。
@asgujjuasitgets · X

在多种格式中与AI进行实时辩论练习,获得性能反馈。
@prof_safezone · X
Gambling on debate rounds with AI judges.

完成几个句子测量你的clanker评分。
niklio · HN
You write 8 text completions and open models score how predictable each word was too them. Predictable => clanker. You can share results with your friends. The scoring checks every word you write against the model's logprobs. Right now I'm using Llama3.1, Deepseek v3 and Qwen3 to keep costs low. I tried to calibrate it so other models (chatgpt/claude) score 100% and interesting human responses score in the 10-30% range. Totally free, no signup

在 leaderboard 上按官方基准对比 AI 大模型的性能排名
fcten · V2EX
做了一个大模型 leaderboard 网站 最近一个月 CodeX 疯狂送重置,token 根本用不完,顺手做点东西。 地址:[知行录]( https://leaderboard.cn/) 排行依据主要为模型官方基准测试成绩。非主观排名。 数据会持续更新。如果有点用,欢迎各位 v 友收藏~

一款无尽跑酷游戏,你扮演忙碌公牛,收集$JAMES代币并躲避Monad市场风险。
@0xNftLegend · X
I built a game for $JAMES lovers. The Busy bull of Monad. Thanks @buildanythingso for inspiring me. My first Vibe Coded game. with no coding knowledge at all.

为LLM输出提供token级引文API,通过注意力分析验证。
apoorvumang · HN
TokenPath – token-level citations for LLM output, read from attention

与AI或在线对手下棋,通过迷你游戏决定每次吃子。
@GeorgeGognadze · X
Hi, I’m building the new chess version where mini games decide the capture @dualchess