fandf.coDeveloper Tools1 month agoServerless GPU Platform for AI Inference | RunpodDeploy AI inference models on serverless GPUs with sub-200ms cold starts and pay-per-second billing.@svpino · XYou can check out Runpod here: Thanks to the Runpod team for partnering with me on this post.
baremetalrt.aiDeveloper Tools1 month agoBareMetalRT — Bare Metal AIRun LLM inference on consumer GPUs with NVIDIA TensorRT-LLM optimization.brianhabana123 · HNTensorRT-LLM running natively on Windows (no WSL)