
MemoAIr - Realtime memory for voice AI agents
Real-time memory and context layer for voice AI agents with sub-10ms latency.
@SouravDaaa · X
- End to end personalised memory retrieval in 10ms. Would love some feedback here
The full gallery
Tech stack
22 projects

Real-time memory and context layer for voice AI agents with sub-10ms latency.
@SouravDaaa · X
- End to end personalised memory retrieval in 10ms. Would love some feedback here

Transcribe audio and video to text entirely on your device, with no uploads and no account needed.
mikicat · V2EX
[分享创造] 做了一个纯本地离线运行的语音转文字工具: Free Local Transcription 大家好, 平时经常有一些包含敏感信息或者不想上传到第三方服务器的音频(如会议记录、采访、个人备忘等)需要转文字。市面上很多工具本质上都是套壳调用云端 API ,既有数据泄露风险,长期用也有不小的 API 成本。 结合自己的模型优化经验,我做了一个完全基于本地算力运行的音频转文字工具:Free Local Transcription 🔗 体验地址: https://freelocaltranscription.com/ 主要特点 100% 本地运行 / 零数据外传:推理完全在设备本地完成,音频文件不会上传到任何服务器,断网状态下也能保证隐私与安全。 无需 API Key / 完全免费:依靠本地算力运行,不需要填 API Key ,也没有按分钟计费的困扰。 即开即用:界面保持极简,拖入音频即可直接进行本地转写。 目前的权衡与折衷 相比直接走云端高性能 GPU 集群,本地推理的速度会直接取决于用户的设备配置与硬件性能,老旧设备上可能存在内存占

Create AI-generated videos from text or images with automatic voiceovers.
Wan AI — AI 视频生成器,可将文本或图像转换为视频

Transcribe voice messages into clean, copyable lists for notes and apps.
emadda · HN
list55.com: Transcribe a list into plain text

Record your voice to get AI feedback on clarity, pace, stability, and listener impression.
MRNiu · V2EX
做了一个免费的声音表达分析工具,录 5–15 秒就能看到反馈 最近做了一个小工具:Voice Impression Checker 。 它想解决的问题很简单:我们经常知道自己“说了什么”,却不太清楚这段话听起来是什么感觉。 打开网站后,只需要录制 5–15 秒自然说话的声音,工具会从这一次录音中分析: - 清晰度、语速、稳定性、能量和声音变化 - 听感上的温暖、自信、亲和力和参与感 - 这次表达做得比较好的地方 - 下一次录音可以尝试改进的一件事 网站还提供了 Rate My Voice 模式,可以针对面试回答、演讲开场、自我介绍和日常交流等场景进行练习。 需要特别说明的是,它不是声音好坏评判,也不会通过声音判断性格或身份。所有结果只描述当前这一次录音的表达效果,更适合用来反复练习和比较不同说法。 目前网站免费使用,不需要注册账号,录音仅用于生成当前报告,不会保存。 网站地址: https://voiceimpressionchecker.com/ 这是一个刚完成的早期版本。如果你愿意体验,欢迎告诉我结果是否容易理解,以及你最希望

Compare speech-to-text engines (OpenAI, Deepgram, NVIDIA, Fish Audio) with real-time benchmarking and local privacy.
@alvaisy · X
finished voice to text small web app for my own itch. it's opensource. use openrotuer key. and use it with 4 models.

Convert any webpage to audio with 100+ ultra-realistic AI voices.
@sumark62030686 · X
Just shipped Mark Reader v1.1.6 The most polished AI text-to-speech Chrome extension. 100+ neural voices Immersive reader mode Smart multi-site adapters Zero monthly server cost Try free #ChromeExtension #AI #SaaS #BuildInPublic

Reproductor minimalista de resúmenes de libros y artículos para gente ocupada con soporte RSS, offline y recomendaciones.
@MigoCreativo · X

Clone your voice to create and narrate original stories.
u/Aware_Web9715 · Reddit
Create stories in your voice Built using open source models (Qwen 3.6 and TTS). Clone your voice and create an original story for free. Just a fun weekend project that runs on my spark. Feedback gets reviewed daily and the coding agents attempts to create a PR. https://app.getzoro.com submitted by /u/Aware_Web9715 to r/SideProject [link] [comments]

🚀 Streamline your business communication: Business Texting, Business Calling & Team Chat, all in one place with optional human in the loop AI automation.
@SynapseComs · X