
TokenMaxxer — the global AI usage leaderboard
Leaderboard showing your AI token usage across Claude Code, Codex, Cursor and other coding tools.
SYeomans · HN
TokenMaxxer – track every AI token you spend across your coding tools
The full gallery
Tech stack
17 projects

Leaderboard showing your AI token usage across Claude Code, Codex, Cursor and other coding tools.
SYeomans · HN
TokenMaxxer – track every AI token you spend across your coding tools

Access Claude, GPT, Gemini and other AI models through a unified API with pay-as-you-go billing.
@LofeeRouter · X
The latest and best AI models, made affordable — all in one place. Claude, GPT, Gemini and more. Pay as you go, separate app keys. Dedicated support for Claude/Anthropic workflows, Claude Code and Codex. Whatever AI service you need, start with Lofee👉

Get a free API key with email to access 500+ Hugging Face models.
@pengsonal · X
500+ Hugging Face models through a single free API no credit card just an email signup 😳 launched on July 3 as a unified gateway for Hugging Face models one API key one endpoint 500+ models what you get for $0: • 500+ open-source models through an OpenAI-compatible API • no separate API keys for different providers • works with Cursor, Claude Code, Hermes, OpenCode, and anything that supports a custom base URL • email signup only setup takes about 2 minutes: 1. Go to 2. Sign up with your email 3. Generate an API key 4. Set your base URL to 5. Choose any model from the catalog and start building a few things worth knowing: • is a third-party gateway, not an official Hugging Face product • the free tier is rate-limited, but the exact limits haven't been published yet • i wouldn't build production apps that depend entirely on a free aggregator what i like is the simplicity instead

Compare AI model coverage, pricing, uptime, and latency across different AI relays.
zizheruan · HN
XTokenChecker – Verifies model identities of your AI gateway

OpenAI-compatible API for running open-weight LLMs and video models.
bingus-bongo · HN
Use GLM-5.3 in Cursor today via tokengo API

An LLM gateway for OpenAI, Anthropic, Google and Azure. Every request logged, priced to the token, and audited for waste you can actually recover.
@razdagan3 · X

Access top-tier AI models through a unified API with anonymous, borderless payments.
@xiaoheihei257 · X
📢 大消息!Gemini 3.6 Flash 和 Gemini 3.5 Flash-Lite 已经在 API 正式上线了!🚀 最近模型更新越来越频繁,这次 直接把两个新版本推出来,实际用起来感觉又进了一步。 先说 Gemini 3.6 Flash: 它是 3.5 Flash 的升级版,输出质量明显更好,但价格完全没变。最关键的是 token 消耗减少了大概 17%,遇到 DeepSWE 这种复杂代码生成任务,最多能省下 65% 的 token,成本直接降下来了,性价比很高。 再看 Gemini 3.5 Flash-Lite: 这是 3.5 系列里跑得最快、最省钱的那个,输出速度最高能达到 350 tokens/s。特别适合做高频任务,比如批量处理文档、Agent 实时搜索这些场景,用起来又快又稳,不会卡顿。 的模型生态现在越来越丰富了,两个新模型都支持官方 API 直接调用,开发者用着也方便。整体看下来,AI Agent 时代的底层支持又扎实了不少。 有在做 AI 项目或者日常调用大模型的朋友,不妨去试试新版本,体验应该会挺惊喜的~ @justinsuntron @BAI_AGI #TRONEcoStar

Your OpenAI client, a different base URL, a much smaller invoice. Frontier open-source models on a decentralized GPU network.
@runnoclip · X
AI inference service that cuts your token bills by 50-90%

Rightsize OpenAI and Anthropic models. See what drives your AI bill, then validate cost-efficient model changes without rewriting your application.
@SpendLensAI · X

Access 200+ AI models (Claude, GPT, Gemini) through one unified API with automatic failover.
@MixRoute_ai · X

Access thousands of AI models through a single OpenAI-compatible API.
@mageofweb3 · X

Compress prompts and reduce LLM token costs by detecting duplicate tool calls.
@DeveloperL92487 · X
I built my first app in 60min And now I got $500 MRR in one month Check here if you are interested It’s a tool to reduce agent token consumption, speed up agent response, and clean up memory cache