fandf.coDeveloper Tools1 month agoServerless GPU Platform for AI Inference | RunpodDeploy AI inference models on serverless GPUs with sub-200ms cold starts and pay-per-second billing.@svpino · XYou can check out Runpod here: Thanks to the Runpod team for partnering with me on this post.
ar5en1c.github.ioDeveloper Tools1 month ago · 3 triesHEADROOM — how fast is your machine, really?Measure your GPU's real memory-bandwidth ceiling for local AI in 30 seconds. Open sourceAr5en1c · HNHeadroom – measure your GPU's true bandwidth ceiling for local AI
baremetalrt.aiDeveloper Tools1 month agoBareMetalRT — Bare Metal AIRun LLM inference on consumer GPUs with NVIDIA TensorRT-LLM optimization.brianhabana123 · HNTensorRT-LLM running natively on Windows (no WSL)