
FlexInference: Drop your AI costs today
跨多个AI提供商路由您的LLM请求,降低成本并满足延迟要求。
Aperswal · HN
Made a Free LLM Router
完整作品展
技术栈
60 projects

跨多个AI提供商路由您的LLM请求,降低成本并满足延迟要求。
Aperswal · HN
Made a Free LLM Router

一站式AI创作平台,生成视频、图像、音乐和3D内容。
KKV AI — KKV 是一站式 AI 创作平台,提供视频生成、图像创作、照片编辑、趣味滤镜、AI 聊天助手等功能,无障碍访问 Veo 3、Flux、Claude Opus 4 等 100+ 顶级模型

用比特币或门罗币购买 API 密钥访问 Anthropic 和 OpenAI,无需账户。
not_wowinter13 · HN
Anonymous LLM proxy. Pay in crypto, no account needed

GoldBean 是按使用付费的 API 市场,提供 47 个端点用于 OCR、翻译、图像生成和对话式 AI。
13639366668 · HN
Pay-per-call MCP server with 47 AI endpoints, micropayments via x402

查看 LLM 成本数据,按模型和环境分类,获取优化建议。
ATsimbalistov · HN
Show HN: Tracking GenAI cost and endpoint fragility so app teams don't have to

Ornymo通过语义缓存减少LLM查询成本和延迟。
u/ornymo_official · Reddit
how to reduce ai costs there are lots of way to reduce costs but there all complex to setup i know this cause i tried one in production so i built ornymo we cache meaning not the exact string allowing us to give same awnsers thus reducing llm costs and latency check it out at ornymo.com free for a limited time and let me know your feedback submitted by /u/ornymo_official to r/buildinpublic [link] [comments]

粘贴URL生成llms.txt和MCP端点,使任何网站可被AI代理查询。
AshHackerNews · HN
AgentReady – MCP server that makes any docs site queryable by AI agents

用NVIDIA TensorRT-LLM在消费级GPU上进行高性能大语言模型推理。
brianhabana123 · HN
TensorRT-LLM running natively on Windows (no WSL)

从任务参数生成 AI 工作流模板、提示词和质检清单。
AI 工作流模板生成器 — 免费 AI 工作流生成工具,输入重复任务、角色和输出格式,自动生成任务边界、处理步骤、提示词和质检清单。

通过 RavenGate 网关路由 LLM API 流量,追踪成本、分析延迟、隐蔽 PII。
charltonraven · HN
RavenGate – LLM gateway that redacts PII across SSE chunk boundaries

