
oruk — Speech API for transcripts, emotion, and style
用于语音转录、情感检测和说话风格分析的AI音频API。
dillon_silzer · Product Hunt
Oruk Speech AI that actually understands you
完整作品展
技术栈
61 projects

用于语音转录、情感检测和说话风格分析的AI音频API。
dillon_silzer · Product Hunt
Oruk Speech AI that actually understands you

自动为视频生成音频描述,帮助视障用户无障碍访问视频。
u/Thecuriousbloke · Reddit
I built an AI tool that automatically generates audio descriptions for videos. I'd love some honest feedback I've spent the last few months building an AI tool that automatically generates audio descriptions for videos to make them accessible for visually impaired users. I'd love feedback on whether the descriptions are actually useful or if there are obvious issues. Looking for honest feedback. Check it out here - https://accessly.studio submitted by /u/Thecuriousbloke t

快速、精准的 AI 音频转录工具,无需注册,支持多种格式和语言。
音频转文字工具 — AI 音频转文字,快速、精准、安全的转写服务

通过AI聊天界面一站式创建视频,自动生成脚本、配音和编辑。
u/Real-Estate-Agentx44 · Reddit
Built an AI video tool where one chat writes, voices, generates, and edits the whole video - is all-in-one actually useful or do you prefer separate tools? I've been building ViewPress AI , an AI video generator, and I'm at the point where I want honest feedback before I keep pushing in one direction. Instead of jumping between a script tool, a voice tool, an image/video generator, and an editor, you describe what you want in one chat and the agent handles the whole thing. It

基于你的知识训练AI代理,用你的声音回答访客问题。
@withmindola · X
let's go! you can build you digital twin by feeding all your content and share it with people so they can talk to your second brain in your own voice.

一键下载微软文字转语音合成音频,支持所有官方语音和SSML语法。
微软 TTS (Text-to-Speech) 文字转语音下载器 — 一键播放或下载 微软 TTS 文字转语音 合成的音频。支持所有官方提供的语音和声音选项。支持SSML语法。每月享免费用量。量大者可升级Pro Plan,固定便宜费用无限使用。摒弃按量收费。支持支付宝

为网站和应用添加以文档为基础的语音和指针引导。
@moelabs_dev · X
Skilly — gives your product a voice that points. One script tag adds in-app help that answers from your own docs and moves the user's cursor to the exact button.


将 MP3 和 MP4 文件转换为带有时间戳和说话人识别的可编辑文字稿。
digiplanp · V2EX
做了一个音视频转文字工具,支持 MP3/MP4、说话人识别和字幕导出 大家好,最近做了一个在线音视频转文字工具 ToText ,想发出来请大家体验一下,也听听 V 友的意见。 最开始做这个工具,是因为我自己在处理会议录音、访谈和视频内容时,发现单纯生成一大段文字并不太实用。转录完成后,通常还需要反复对照原始录音、修改错误、区分说话人,再整理摘要或字幕。 所以现在做成了一个相对完整的工作流: * 上传 MP3 、MP4 、M4A 、MOV 、WAV 、WebM 等音视频文件 * 自动识别语言并生成带时间戳的文字稿 * 支持说话人识别和说话人名称修改 * 点击文字片段可以跳转到对应的音视频位置 * 可以搜索、编辑和校对转录内容 * 自动生成摘要、会议纪要、章节和行动事项 * 支持导出 TXT 、DOCX 、PDF 、SRT 、VTT 、CSV 和 JSON 音频文件可以直接使用 [MP3 to Transcript]( https://totext.app/) 功能,比较适合播客、访谈、课程录音、会议和语音备忘录。 视频文件也使用同一个编辑器,通过


用语音从手机无需手操作运行 Claude Code 会话
@phiwenger · X
Yeah /remote-control is the workaround but it's clunky. I built yothere for exactly this. With yothere you can spin up multiple Claude Code sessions from your phone, run them all hands-free by voice, and just get pinged when done or stuck.

生成AI语音、文本转语音、克隆语音,创建音乐
Seed Audio — AI 语音生成与在线音乐创作工具