
goku-temp
用WebAssembly在浏览器运行和管理LLM模型
userfrom1995 · HN
Goku – WASM (wllama)-powered LLM inference and model manager
完整作品展
技术栈
61 projects

用WebAssembly在浏览器运行和管理LLM模型
userfrom1995 · HN
Goku – WASM (wllama)-powered LLM inference and model manager

为LLM输出提供token级引文API,通过注意力分析验证。
apoorvumang · HN
TokenPath – token-level citations for LLM output, read from attention

开源LLM和视频模型的OpenAI兼容API
bingus-bongo · HN
Use GLM-5.3 in Cursor today via tokengo API

统一的AI模型网关,支持OpenAI、Claude等兼容接口访问大语言模型。
@YinsenW_ · X
CherryIN 平台已经上线了DeepSeek v4 flash Vision-Exp 模型,为全球用户带来最领先的多模态AI 服务哈哈哈

Flat-rate LLM subscription for coding agents — OpenAI-compatible API for OpenCode, Cline, Aider, Codex, Kilo, Claude Code and always-on agents like OpenClaw and Hermes. Frontier mo
@stdcmpt · X

为代码智能体工作流管理 API 密钥和使用预算。
u/Zyron_X · Reddit
I built a service for people to use Codex API without 5-hour limit disruption I built a small service for people who use the OpenAI Codex API regularly and want more predictable usage without the 5-hour or weekly limits. It currently provides: Frontier OpenAI models (GPT 5.6 family included) Managed API key Monthly usage budgets depending to plan No 5-hour limit No weekly limit Under the hood, it is built on top of an open-source project and proxies requests to


通过一个API访问多个AI模型,提供透明的预付费定价。
@MyApiTaco · X
GLM 5.3 Flash is now on 🌮⚡ Limited-time promotion: 66.6% OFF retail • Input: $0.05 (retail $0.15) • Cache Input: $0.01 (retail $0.03) • Output: $0.167 (retail $0.50) Plus, get an extra 5% bonus on topups over $100. #GLM #zAI #openrouter #vibecoding

在欧盟托管私有 LLM 实例,固定月费无使用限制。
CodingPanda42 · HN
Virtual Private LLM, fixed fee with no usage or token limits

用对抗测试检查LLM端点安全,获取OWASP审计报告。
@aryaan_sheth · X
- LLM security for small teams

估算LLM推理所需的显存、延迟、TTFT、TPOT与吞吐量。
popopanda · HN
LLM Inference Calculator – Estimate VRAM, Latency, and Throughput

诊断OpenAI兼容API的模型质量、降智与协议兼容性。
AI快站模型质量检测 — 面向 OpenAI Compatible 接口的网页检测工具,输入公开 HTTPS 地址和临时 API Key,可检查模型声明、Token、动态题、SSE 与工具调用并生成分项报告;密钥仅用于当次检测,不写入数据库、缓存或日志