
sipsip.ai — AI Summarizer, Transcriber & Daily Brief
用AI转录和总结YouTube视频、播客、PDF和MP3/MP4文件。
sipsip.ai — 支持30+平台视频AI转录并提取摘要,订阅RSS可以每天发送早报
完整作品展
技术栈
24 projects

用AI转录和总结YouTube视频、播客、PDF和MP3/MP4文件。
sipsip.ai — 支持30+平台视频AI转录并提取摘要,订阅RSS可以每天发送早报

查看LLM模型在10个基准问题上的评分和排名。
fristovic · HN
She watched me look at model rankings and asked what do the numbers mean... I literally had no good way of explaining it to her so I just came up with something that is approximately in the same ballpark as some of the benchmarks out there lol

在浏览器中测试小型语言模型(8M-13M 参数),离线可用。
u/Live_Confusion_3003 · Reddit
I trained an LLM that runs on an ESP32 and directly in the browser Link to try it out yourself is: topk.sh The models download their weights directly in the browser so it works offline. Keep in mind they are very small and inaccurate. (8M and 13M parameters) However, I am building 500M and 1B+ parameter local models for agent based coding and other purposes. I will be shipping hardware designed for these tasks which connect directly to you computer or other device.

在仪表板上版本管理、测试和部署 LLM 提示词,无需修改代码。
@why_deepanshux · X
I Just launched my first SaaS. Late night coding session, white board and my my markers knows what we built. Now it's world's turn. Please checkout Link below.

FlexInference: 通过多个提供商路由LLM API请求,降低成本和延迟。
Aperswal · HN
Made a Free LLM Router

用任何AI构建和托管应用
@akhileshrangani · X
i built codex micro and used it inside of claude to control codex AND claude code it uses a herdr bridge that is running on my mac talks it through a ngrok proxy uses to render inside of claude

将LLM调用路由到成本最低的合适模型,保持质量。
george_avila · Product Hunt
IQ Routing Trajectory-aware LLM routing that cuts agent cost

在 AINA 工作区使用多个 AI 模型研究、搜索、分析文件和创作。
@sahiinthehood · X

用 LLM 评估 AI agent 对话质量,提供评分卡和成本分析。
@tech_maju · X

用间隔重复法复习 LeetCode 题目,提升记忆和解题能力。
@JoydeepNath007 · X

@launch_llama https://t.co/YsRI7cTzqm https://t.co/EXOr5fEkp0
@melonrice383235 · X

语义缓存减少LLM令牌成本和AI查询延迟。
u/ornymo_official · Reddit
how to reduce ai costs there are lots of way to reduce costs but there all complex to setup i know this cause i tried one in production so i built ornymo we cache meaning not the exact string allowing us to give same awnsers thus reducing llm costs and latency check it out at ornymo.com free for a limited time and let me know your feedback submitted by /u/ornymo_official to r/buildinpublic [link] [comments]