
LLM 推理计算器 | LLM Inference Calculator
Estimate GPU memory, latency, TTFT, TPOT, and throughput for LLM inference.
popopanda · HN
LLM Inference Calculator – Estimate VRAM, Latency, and Throughput
The full gallery
Tech stack
21 projects

Estimate GPU memory, latency, TTFT, TPOT, and throughput for LLM inference.
popopanda · HN
LLM Inference Calculator – Estimate VRAM, Latency, and Throughput

HalfBuilt is where founders find co-founders. Share what you're building, discover people with the skills you're missing, and start the conversation.
@HalfBuilt_org · X
You can find your product at our website

A productivity app that unifies different tools and reduces switching between them.
@vinctkng · X
A productivity app that solves the pain of switching between different productivity tools and linking all of them together!


Live funding rates across Hyperliquid core and HIP-3 markets. Build long/short baskets and see the net carry instantly.
@akiran0ma · X

Automatically route each prompt to the cheapest capable model to cut API costs.
u/ASDKING100 · Reddit
Launched an AI API router tonight, and found a bug hours in that would've taken real payments without ever upgrading the account Built LLMLite over the past few weeks — it classifies each prompt and routes it to the cheapest model that can actually handle it, instead of hitting GPT-4o for everything. Free tier, no card needed to try it. Tonight, right as I was about to launch, ran a real transaction to test the payment flow end to end. Paddle processed it, webhook fired, signature verified

100 spaces. Sold in order. One company per space.
@euxhenjonex · X

Turn text and images into cinematic 4K+ videos with native audio, multi-shot sequencing, and persistent character identity.
kling — AI 生成视频,支持多种 kling 模型

AI workbench for generating and managing images and videos with 17 models and private media library.
CreateForge AI — AI 图像与视频生成工作台,接入 17 个模型,支持 Seedance 2.5、Veo 3.1 等模型的文生图、图生图、文生视频和图生视频,提供积分报价、任务中心与私有媒体库