
WatchMachineGo
Visualize hardware performance metrics while running LLM inference on your system.
dev_dan_2 · HN
WatchMachineGo – A visualizer to show hardware performing LLM inference
The full gallery
Tech stack
60 projects

Visualize hardware performance metrics while running LLM inference on your system.
dev_dan_2 · HN
WatchMachineGo – A visualizer to show hardware performing LLM inference

Centralize UTM parameter management for all your marketing campaigns.
@AnilBatra · X

Canvas for inspecting LLM weights tensor-by-tensor with quantization error and distribution analysis.
alesha-pro · GitHub
atlas Interactive canvas for taking an LLM apart tensor by tensor: measured INT8/INT4/FP8 error, distributions, spectra and outlier channels for every weight tensor

Visual editor for configuring multi-agent LLM systems with local inference.
sascha10000 · HN
Multi-agent LLM editor with local inference via WebSockets

The end-to-end platform for small language models. Tuned to your task, a small open model matches frontier accuracy at a fraction of the cost. Own your intelligence: private, compa
@dayoffdev · X

Track AI and LLM news and model launches from 100+ sources.
jonam21 · HN
KBlip – turns AI/LLM news across 100 sources into daily digest threads

Add AI interpretation buttons to financial news articles on Chinese financial platforms.
EliteOtaku · V2EX
搞了个解读财经数据的油猴脚本,适配金十,财联社,汇通,华尔街见闻 比较简单,但挺好用的,在快讯旁边加了一个 AI 按钮,点击后由 AI 解读该新闻/数据的影响,需自己准备 API key,支持 DeepSeek 和 OpenAI、Anthropic 格式 https://greasyfork.org/zh-CN/scripts/590009-%E9%87%91%E5%8D%81%E6%95%B0%E6%8D%AE%E5%87%80%E5%8C%96-ai-%E8%A7%A3%E8%AF%BB-deepseek

OpenAI-compatible API for running open-weight LLMs and video models.
bingus-bongo · HN
Use GLM-5.3 in Cursor today via tokengo API

API providing token-level citations for LLM output grounded in attention analysis.
apoorvumang · HN
TokenPath – token-level citations for LLM output, read from attention

Compare language model performance at solving Redactle puzzles.
pampas · HN
Redactle LLM Leaderboard

Adaptively route and load-balance requests across 200+ LLMs through one OpenAI-compatible gateway.
Continuum-AI-Corp · GitHub
OrcaReplay OrcaReplay — Time travel for AI agents. Record, replay, fork, and debug any agent run with any model. Built by the OrcaRouter.ai team.

Play with throttling rules to see how AI providers decide which customers get strong models during demand spikes.
eliotho · HN
I built a tool showing how AI providers (should) throttle their models