
hRAG · Retrieval with receipts
Self-hosted hybrid RAG on a €116/month cluster — Postgres, BM25, vectors, a reranker, and a public benchmark score for every claim.
victor_edka · HN
HRAG – Hybrid RAG on €116/month of Hetzner, officially benchmarked
The full gallery
Tech stack
25 projects

Self-hosted hybrid RAG on a €116/month cluster — Postgres, BM25, vectors, a reranker, and a public benchmark score for every claim.
victor_edka · HN
HRAG – Hybrid RAG on €116/month of Hetzner, officially benchmarked

RAG-powered chat with image generation and model routing.
@Aliasistx · X
researching and developing RAG systems for agentic Graph coming on the horizon.

Unified context layer for AI agents that integrates email, drive, Slack, Notion, and GitHub.
@marcin_u2 · X
-> here it is! claude finally knows everything gpt knows. and both of them know everything that's happened across your email, drive, slack, notion, github and granola :)

Investigate current grifts and scams using a RAG-powered research tool.
u/King_Kazma · Reddit
whatsthegrift.com - a RAG research tool for today's grifters every day i feel like there's a new grift attached with some news article, which naturally birthed the idea to built something to investigate & visualize them examples: https://whatsthegrift.com/q/777063a64e512565 https://whatsthegrift.com/q/e5945d20375ed3ff and my personal favorite so far: https://whatsthegrift.com/q/f43eb30a15342d53 react (vite) + node (fastify) + brave search (retrieval) + claude haiku (extraction)

Memory infrastructure for AI agents to store and retrieve contextual data.
memcode-in · GitHub
memcode Memcode: #1 Memory Layer for AI agents, Building Memory infra for every use case

An AI-powered search engine that delivers precise answers using trusted sources.
u/Redshankscommune1 · Reddit
Building an AI search engine instead of another chatbot I've been building an AI-powered search engine over the last few months as a side project. The idea came from noticing that I'd often Google something, open several tabs, then copy everything into an AI tool to get a useful answer. It felt like there was an unnecessary step in the process. The goal isn't to replace the web, but to make research faster by generating structured answers with citations and giving users the option to see

Compress LLM prompts and documents to reduce token usage and API costs.
@marcusyul · X
THEY JUST GAVE AWAY 100 MILLION FREE TOKENS SO YOU CAN STOP BURNING THROUGH YOUR CLAUDE CODE BUDGET. if you code with AI you already know: the session fills up, starts failing, and on top of that you're overpaying there's a tool that fixes this: it shrinks the context before the model even sees it same model, same response, a fraction of the cost in a real session: from $154 to $43. a 72% drop and right now: → extend your Fable sessions in Claude Code → 100M free tokens to try it out you don't switch models you don't touch your code you just stop paying to repeat yourself link below ⬇️


Affordable memory storage layer for AI agents.
@AudaxicTech · X
Memory for AI agents, most affordable in the industry

Deploy your vibe-coded apps instantly with built-in database, file storage, and realtime interactions.
@Fusekit_cloud · X
Especially when that vibe coded app is internal and hyperspecific to your needs! Try to deploy your vibe coded apps and automatically get a full database, file storage, realtime interactions and many morr thing! Free to start without a credit card right now :)

Store and share prompts, rules, and context that AI tools like Claude Code, Cursor, and ChatGPT can access via MCP.
@vibexp_io · X
Your plan is now code. Claude Code Dynamic Workflows fan out up to 1,000 subagents, 16 at once, each in its own context, verifying until the results converge. Built for big bug hunts, migrations and audits. Source:

Search indexed open-source code and packages with version history, metadata, and dependency information.
@Jack_Timonen · X