
slowcast.net · Something slow to watch. Something calm to hear.
Watch calming visuals and listen to ambient audio in a meditative browser experience.
@ace_space_1 · X
The full gallery
Tech stack
31 projects

Watch calming visuals and listen to ambient audio in a meditative browser experience.
@ace_space_1 · X


Conduct AI-powered voice, chat, and video interviews with automated transcripts and insights.
Aural — 开源 AI 面试平台,支持语音、聊天和视频面试,提供自适应追问、结构化评分、面试练习与自托管 - [查看仓库](https://github.com/1146345502/aural-oss)

Sonara turns text into natural, expressive speech. Studio-grade voices for narration, assistants, and storytelling.
@Dev_NexaLabs · X
NOTE: we're still in the early phase something is not yet finish.

Record therapy sessions and auto-generate SOAP notes with AI transcription.
kithmedai · HN
Kith – AI clinical notes for therapists from ambient session audio

Convert text and PDFs to audio with natural-sounding AI voices.
@caschiblu · X

Convert MP3 and MP4 files to editable transcripts with timestamps and speaker identification.
digiplanp · V2EX
做了一个音视频转文字工具,支持 MP3/MP4、说话人识别和字幕导出 大家好,最近做了一个在线音视频转文字工具 ToText ,想发出来请大家体验一下,也听听 V 友的意见。 最开始做这个工具,是因为我自己在处理会议录音、访谈和视频内容时,发现单纯生成一大段文字并不太实用。转录完成后,通常还需要反复对照原始录音、修改错误、区分说话人,再整理摘要或字幕。 所以现在做成了一个相对完整的工作流: * 上传 MP3 、MP4 、M4A 、MOV 、WAV 、WebM 等音视频文件 * 自动识别语言并生成带时间戳的文字稿 * 支持说话人识别和说话人名称修改 * 点击文字片段可以跳转到对应的音视频位置 * 可以搜索、编辑和校对转录内容 * 自动生成摘要、会议纪要、章节和行动事项 * 支持导出 TXT 、DOCX 、PDF 、SRT 、VTT 、CSV 和 JSON 音频文件可以直接使用 [MP3 to Transcript]( https://totext.app/) 功能,比较适合播客、访谈、课程录音、会议和语音备忘录。 视频文件也使用同一个编辑器,通过

Segment Anything for Audio — made simple. Separate audio with a text prompt. Isolate vocals, remove background music, extract speech from noise. Privacy-first stem splitter powered
@sambamdamnn · X

Transcribe audio and video recordings into searchable, editable text.
lavande · V2EX
做了一个语音转文字的工具,另加 AI 总结、翻译、提问等(附踩坑记录) 先上链接:[Echoryte]( https://echoryte.com/) 感觉这个网站 vibe coding 的话,感觉手熟的同学可能半天就能搓出来…… 但是我用了一个多月才上线,主要是被 GPT5.6 坑了,当时 5.6 和 fable 前后发布,然后试了试用 fable 5 做的 plan ,然后 5.6 sol 推理开到最大去执行。结果愣是给我干了 33 个小时没停,写了将近 2 万行代码。 但是坑就坑在这,推理开太高,就过度思考了,给我整了很多虚头巴脑的东西出来,什么“门禁”我也不懂,什么合规,一大堆,然后我就不停地删,删了一个月,最后我就跟 GPT 说,我特么是个人小项目,不是大公司,再这么高下去永远上不了线了。然后终于看删的差不多了,上线了! 欢迎大家试用,轻拍,谢谢!

Upload a PowerPoint presentation and voice sample to generate a narrated slideshow in your cloned voice.
u/sludge_dev · Reddit
I built a tool that narrates PowerPoint slides in your cloned voice. Launching it today, keeping it small for a week to fix what breaks. I just got into my College holidays a few weeks ago and I spent the last 6 weeks building Moduvox . You upload a PPTX deck, record 30 seconds of your voice, and it generates per-slide narration audio with a shareable link and viewer analytics. The idea came from a friend manually record voiceovers for training decks. Every update meant re-recording whole

A shared music room for friends. Everyone hears the same track at the same second, one person holds the aux, and anyone can take it off them.
@theeasyway6 · X

An intellectual companion for the digital aristocracy.
@scooperleash · X