
Auriko | One API for Every LLM, Zero Markup, Cache-Aware Cost Arbitrage
使用一个 API 访问和切换多个 LLM 提供商,同时优化推理成本。
justin2025 · Product Hunt
Auriko Trading desk for LLM calls
完整作品展
技术栈
62 projects

使用一个 API 访问和切换多个 LLM 提供商,同时优化推理成本。
justin2025 · Product Hunt
Auriko Trading desk for LLM calls

整理创意工具和用品,避免重复购买。
@ErathCountyNaNo · X
Check out what I just built with Lovable!

对比不同Agent Harness策略的成本,基于缓存和上下文计费。
taosx · HN
I created a simulation for coding harnesses based on my own pi sessions. When taking into account all factors, DS-v4-Pro is cheaper than gpt-5.6-luna due to caching. Look at the bill segments difference for cache read cost and uncached cost between deepseek and the other models. At this point is cheaper to use ds-v4-pro than the luna models from openai. ignore the numbers except the classic and keep in mind that classic is based on pi with the only change limiting tool output to 10kb https://har

缓存和重用深度研究报告,避免重复运行费用高的研究查询。
u/illerminati · Reddit
Caching Deep Research Output Hi r/SideProject . Just sharing a side project I've done in the recent weeks that I think might be useful for some. I use deep research at work to understand unfamiliar software domains, and in my personal life to compare products before buying them. I can’t use my company's LLM for personal use, and the publicly available deep-research products (like Claude and OpenAI) are too expensive for my taste, so I built a DeepSeek-powered alternative. The pipeline s

语义缓存减少LLM令牌成本和AI查询延迟。
u/ornymo_official · Reddit
how to reduce ai costs there are lots of way to reduce costs but there all complex to setup i know this cause i tried one in production so i built ornymo we cache meaning not the exact string allowing us to give same awnsers thus reducing llm costs and latency check it out at ornymo.com free for a limited time and let me know your feedback submitted by /u/ornymo_official to r/buildinpublic [link] [comments]

Takt 自动分组标签页、防止重复、跟踪资源。
u/-Leelith- · Reddit
Takt tab manager Update n°4: Adding 6 languages translations, (almost) fully interactive demo, an AI grouping feature, reworked our onboarding and fixing many bugs After discovering this sub and my first post , sharing my update n°4: We mentioned on the update n°3 that we wanted to build a live interactive demo. So we did that and build an almost identically fully interactive "try before you install" demo of the extension. The goal was to have a demo that converts a browser visit into a

一次编写,多格式导出为笔记、幻灯片、PDF、PPTX、DOCX 或画布。
@mr_wickedhacks · X
Write the doc once → get a notebook, slide deck, canvas, or resume from the same source. No rewriting for every format.

浏览器中的开发环境,集成shell、git和本地AI。
Dhravya · GitHub
burrow a whole dev machine in a browser tab - bun.wasm, shell, git, and local AI. phones home to nobody.

A playful, cat-themed bookmark manager for Chrome. Browse folders, search, and tidy bookmarks from a compact popup.
@_hermooo · X
I built a chrome extension that helps you manage your boring bookmarks 👀

在桌面式界面中组织和管理文件和文件夹。
@jenidesignns · X
Day 19 challenge by @IwuezeAmarachi built with Claude AI 🤖 Live link:

任务管理应用,Now、Next、Someday 三个清单同步到 iOS 和网页。
u/dahooddawg · Reddit
I built a task manager that replaces priority levels with a hard cap on what "now" means Hey r/SideProject -- I built this for myself because every task manager I tried had the same problem: everything felt high priority. The core idea: instead of High/Med/Low (which fails because everything feels high), you get three buckets -- Now, Next, Someday -- and a hard cap on how many items "Now" can hold. When it's full, adding something new means consciously swapping it for something alread

SaaS定价建模和代币成本计算,支持30+种货币的利润分析。
@HelloCalcaas · X