
Auriko | One API for Every LLM, Zero Markup, Cache-Aware Cost Arbitrage
Use one API to access and switch between LLM providers while optimizing inference costs.
justin2025 · Product Hunt
Auriko Trading desk for LLM calls
The full gallery
Tech stack
62 projects

Use one API to access and switch between LLM providers while optimizing inference costs.
justin2025 · Product Hunt
Auriko Trading desk for LLM calls

Catalog your maker supplies and tools to avoid buying duplicates.
@ErathCountyNaNo · X
Check out what I just built with Lovable!

Compare costs of coding-agent harness strategies with event-level cache and context accounting.
taosx · HN
I created a simulation for coding harnesses based on my own pi sessions. When taking into account all factors, DS-v4-Pro is cheaper than gpt-5.6-luna due to caching. Look at the bill segments difference for cache read cost and uncached cost between deepseek and the other models. At this point is cheaper to use ds-v4-pro than the luna models from openai. ignore the numbers except the classic and keep in mind that classic is based on pi with the only change limiting tool output to 10kb https://har

Cache and reuse deep research reports to avoid re-running expensive research queries.
u/illerminati · Reddit
Caching Deep Research Output Hi r/SideProject . Just sharing a side project I've done in the recent weeks that I think might be useful for some. I use deep research at work to understand unfamiliar software domains, and in my personal life to compare products before buying them. I can’t use my company's LLM for personal use, and the publicly available deep-research products (like Claude and OpenAI) are too expensive for my taste, so I built a DeepSeek-powered alternative. The pipeline s

Semantic caching reduces LLM token costs and latency for AI queries.
u/ornymo_official · Reddit
how to reduce ai costs there are lots of way to reduce costs but there all complex to setup i know this cause i tried one in production so i built ornymo we cache meaning not the exact string allowing us to give same awnsers thus reducing llm costs and latency check it out at ornymo.com free for a limited time and let me know your feedback submitted by /u/ornymo_official to r/buildinpublic [link] [comments]

Chrome extension that groups tabs automatically, prevents duplicates, and tracks resource usage.
u/-Leelith- · Reddit
Takt tab manager Update n°4: Adding 6 languages translations, (almost) fully interactive demo, an AI grouping feature, reworked our onboarding and fixing many bugs After discovering this sub and my first post , sharing my update n°4: We mentioned on the update n°3 that we wanted to build a live interactive demo. So we did that and build an almost identically fully interactive "try before you install" demo of the extension. The goal was to have a demo that converts a browser visit into a

Write once, export to notebooks, slides, PDFs, PPTX, DOCX, or canvas.
@mr_wickedhacks · X
Write the doc once → get a notebook, slide deck, canvas, or resume from the same source. No rewriting for every format.

A browser-based dev machine with shell, git, and local AI that collects no data.
Dhravya · GitHub
burrow a whole dev machine in a browser tab - bun.wasm, shell, git, and local AI. phones home to nobody.

A playful, cat-themed bookmark manager for Chrome. Browse folders, search, and tidy bookmarks from a compact popup.
@_hermooo · X
I built a chrome extension that helps you manage your boring bookmarks 👀

Organize and manage files and folders in a desktop-like interface.
@jenidesignns · X
Day 19 challenge by @IwuezeAmarachi built with Claude AI 🤖 Live link:

Task manager with Now, Next, and Someday buckets that sync between iOS and web.
u/dahooddawg · Reddit
I built a task manager that replaces priority levels with a hard cap on what "now" means Hey r/SideProject -- I built this for myself because every task manager I tried had the same problem: everything felt high priority. The core idea: instead of High/Med/Low (which fails because everything feels high), you get three buckets -- Now, Next, Someday -- and a hard cap on how many items "Now" can hold. When it's full, adding something new means consciously swapping it for something alread

Simulate SaaS pricing models, token costs, and profit margins across 30+ currencies.
@HelloCalcaas · X