
SuperCompress - Cut Your LLM Token Costs by 65%
Compress prompts before LLM API calls to reduce token usage and costs.
@asgujjuasitgets · X
The full gallery
Tech stack
60 projects

Compress prompts before LLM API calls to reduce token usage and costs.
@asgujjuasitgets · X

24/7 quantitative trading terminal powered by LLM for automated decision-making and execution.
beatrizgauto5 · V2EX
[开源/技术分享] 全天候 LLM 原生量化交易终端 & 自进化执行引擎: R20 Quantum Trader 折腾了几个月,今天把我一直在实盘环境跑的 **LLM 原生量化交易终端 —— R20 Quantum Trader** 完整开源了! 传统量化往往依赖死板的指标打分和固定阈值,这个项目做的一件核心探索就是:**将大语言模型( LLM )作为交易决策与持仓管理的真正大脑**,结合顶级聪明钱资金流与多维因子,实现自动推演、在途调仓与自进化复盘。 为了让大家能直观看到大模型的真实决策过程,我搭建了一个 24 小时全天候无休的公网在线监控终端,无需登录即可实时围观。 --- ## 🌐 项目地址与在线体验 - 🚀 **在线实时监控终端 (Live Demo)**:[https://www.r20.cn]( https://www.r20.cn) - 📦 **GitHub 开源仓库**:[https://github.com/555cute/r20-quantum-trader]( https://github.com/555cute/r20-quant

Estimate GPU memory, latency, TTFT, TPOT, and throughput for LLM inference.
popopanda · HN
LLM Inference Calculator – Estimate VRAM, Latency, and Throughput

Flat-rate LLM subscription for coding agents — OpenAI-compatible API for OpenCode, Cline, Aider, Codex, Kilo, Claude Code and always-on agents like OpenClaw and Hermes. Frontier mo
@stdcmpt · X

Analytics dashboard for LLM API spending by model and environment with optimization suggestions.
ATsimbalistov · HN
Show HN: Tracking GenAI cost and endpoint fragility so app teams don't have to

Compare latency and throughput performance across LLM API providers.
@QAInsights · X

Host a dedicated LLM instance in the EU with flat-rate pricing and no usage limits.
CodingPanda42 · HN
Virtual Private LLM, fixed fee with no usage or token limits

An LLM agent that tracks goals and plans across sessions while showing exactly what it retrieves, verifies, and fails on.
u/OGMYT · Reddit
I built LOLM, a lower-cost LLM agent that shows what it actually did — looking for blunt feedback I’m one of the founders/builders behind LOLM. Most AI products show an answer but hide whether the system retrieved anything useful, verified the result, switched models, hit a limit, or simply stopped. LOLM exposes those parts through controller events and run receipts. It includes: - Live agent - CLI - Coding and small app-building workflows - Memory and self-hosting options - Control decis

Compare and use multiple LLM models in a secure shared team workspace with your own API keys.
@uncoolavatar · X

Track AI models used in your apps and receive warnings before they're deprecated.
taylorgt · HN
Find every AI model your code calls and warn before it's retired

Use one API to access and switch between LLM providers while optimizing inference costs.
justin2025 · Product Hunt
Auriko Trading desk for LLM calls

Convert PDF, DOCX, XLSX, CSV, JSON, XML, HTML, Images to clean Markdown optimized for AI Agents, RAG, and Vector DBs. 100% Privacy-First, In-Browser Conversion.
@13SahajChawla · X
For professionals to redact their client's sensitive informations before giving AI to process it & while converting any kind of document to a structured MD file. Better quality outputs, 100% privacy with on-browser local processing, and fully free!