
Auriko | One API for Every LLM, Zero Markup, Cache-Aware Cost Arbitrage
使用一个 API 访问和切换多个 LLM 提供商,同时优化推理成本。
justin2025 · Product Hunt
Auriko Trading desk for LLM calls
完整作品展
技术栈
74 projects

使用一个 API 访问和切换多个 LLM 提供商,同时优化推理成本。
justin2025 · Product Hunt
Auriko Trading desk for LLM calls

实时可视化硬件在运行LLM推理时的性能指标
dev_dan_2 · HN
WatchMachineGo – A visualizer to show hardware performing LLM inference

FlexInference: 通过多个提供商路由LLM API请求,降低成本和延迟。
Aperswal · HN
Made a Free LLM Router

Runpod是无服务器GPU推理平台,冷启动低于200ms,按秒计费。
@svpino · X
You can check out Runpod here: Thanks to the Runpod team for partnering with me on this post.

通过InferAll统一API访问207+个AI模型
TaylorM492 · HN
InferAll – One API for OpenAI, Anthropic, Google, Nvidia Nim

用NVIDIA TensorRT-LLM在消费级GPU上进行高性能大语言模型推理。
brianhabana123 · HN
TensorRT-LLM running natively on Windows (no WSL)

Lumina answers from your documents — and when your documents don’t cover it, it says so. Storage, compute and inference all inside the European Union.
@talirezun · X
@luminawidget, turns a business's own docs into a 24/7 AI support agent across the website, Telegram, Discord and WhatsApp. First 10 people who DM me here get 50 bonus messages added on top of the free account's usual 30, plenty of room to actually put it through its paces.

多智能体LLM系统的可视化编辑器,支持本地推理。
sascha10000 · HN
Multi-agent LLM editor with local inference via WebSockets

提交 LLM 推理优化内核,在专用硬件上进行基准测试并竞争排名。
carsenk · HN
Frontier.fast – Help push the frontier of LLM speed forward

统一云平台,集数据库、AI推理、函数、容器和存储于一体。
kiran-ravi · HN
Scalix World – AI native Neo cloud built by two engineers in Rust

开源LLM和视频模型的OpenAI兼容API
bingus-bongo · HN
Use GLM-5.3 in Cursor today via tokengo API

探索数据集、计算统计数据、可视化分布和进行推断检验。
@DaytonaRaised · X
Check out what I just built with Lovable!