
OpenLLM — All your LLM needs in 1 place
Format-agnostic LLM hub. Bring your own provider keys and route across Anthropic, OpenAI, ChatGPT/Codex, Kimi, Alibaba DashScope, and AWS Bedrock — with unified observability and c
@0xxmemo · X
The full gallery
Tech stack
60 projects

Format-agnostic LLM hub. Bring your own provider keys and route across Anthropic, OpenAI, ChatGPT/Codex, Kimi, Alibaba DashScope, and AWS Bedrock — with unified observability and c
@0xxmemo · X

Find AI models optimized for your hardware with performance and pricing estimates.
cdnsteve · HN
Tokenstead, find AI models for your hardware

Compare speech-to-text engines (OpenAI, Deepgram, NVIDIA, Fish Audio) with real-time benchmarking and local privacy.
@alvaisy · X
finished voice to text small web app for my own itch. it's opensource. use openrotuer key. and use it with 4 models.

Compress prompts and reduce LLM token costs by detecting duplicate tool calls.
@DeveloperL92487 · X
I built my first app in 60min And now I got $500 MRR in one month Check here if you are interested It’s a tool to reduce agent token consumption, speed up agent response, and clean up memory cache

@solopribuilds https://t.co/KIq3TZFnPv I'm vibe coding thermodynamic computing. So far I have llama running in sim replacing traditional softmax with boltzmann distribution samplin
@paul_s_1738 · X
I'm vibe coding thermodynamic computing. So far I have llama running in sim replacing traditional softmax with boltzmann distribution sampling across a p-bit array. Im building a local research ecosystem too using qwen.

Compare and use multiple LLM models (Claude, OpenAI, Gemini, etc.) in a secure shared team workspace with your own API keys.
@uncoolavatar · X

Version, test, and deploy LLM prompts from a dashboard without code changes.
@why_deepanshux · X
I Just launched my first SaaS. Late night coding session, white board and my my markers knows what we built. Now it's world's turn. Please checkout Link below.

View LLM model rankings across 10 benchmark questions.
fristovic · HN
She watched me look at model rankings and asked what do the numbers mean... I literally had no good way of explaining it to her so I just came up with something that is approximately in the same ballpark as some of the benchmarks out there lol

Estimate VRAM requirements for running models with llama.cpp
hypfer · HN
According to this shitty vibecoded thing "I" built https://hypfer.github.io/will-it-fit-llama-cpp/ (and I guess according to math too), FP16 K/V would give me something like 90k context at the same model quant, which doesn't really fit my usage. But maybe someone else has experience to share there

Generate 2K AI videos (5-15 sec) from text, images, and audio references with native stereo audio.

Anonymous LLM proxy accepting Bitcoin and Monero for API access to Anthropic and OpenAI without an account.
not_wowinter13 · HN
Anonymous LLM proxy. Pay in crypto, no account needed

Automatically route each prompt to the cheapest capable model to cut API costs.
u/ASDKING100 · Reddit
Launched an AI API router tonight, and found a bug hours in that would've taken real payments without ever upgrading the account Built LLMLite over the past few weeks — it classifies each prompt and routes it to the cheapest model that can actually handle it, instead of hitting GPT-4o for everything. Free tier, no card needed to try it. Tonight, right as I was about to launch, ran a real transaction to test the payment flow end to end. Paddle processed it, webhook fired, signature verified