
oruk — Speech API for transcripts, emotion, and style
Audio API for transcription, emotion detection, and speaking style analysis.
dillon_silzer · Product Hunt
Oruk Speech AI that actually understands you
The full gallery
Tech stack
60 projects

Audio API for transcription, emotion detection, and speaking style analysis.
dillon_silzer · Product Hunt
Oruk Speech AI that actually understands you

Transcribe voice messages into clean, copyable lists for notes and apps.
emadda · HN
list55.com: Transcribe a list into plain text

Create videos from text, images, or videos with automatically synchronized audio.
FLUX 3 Video Generator — AI 视频生成工作台,支持文字或图片生成视频、参考图及镜头、动作和声音提示,并清晰标注模型可用状态

Convert text and PDFs to audio with natural-sounding AI voices.
@caschiblu · X

Transcribe audio and video to text entirely on your device, with no uploads and no account needed.
mikicat · V2EX
[分享创造] 做了一个纯本地离线运行的语音转文字工具: Free Local Transcription 大家好, 平时经常有一些包含敏感信息或者不想上传到第三方服务器的音频(如会议记录、采访、个人备忘等)需要转文字。市面上很多工具本质上都是套壳调用云端 API ,既有数据泄露风险,长期用也有不小的 API 成本。 结合自己的模型优化经验,我做了一个完全基于本地算力运行的音频转文字工具:Free Local Transcription 🔗 体验地址: https://freelocaltranscription.com/ 主要特点 100% 本地运行 / 零数据外传:推理完全在设备本地完成,音频文件不会上传到任何服务器,断网状态下也能保证隐私与安全。 无需 API Key / 完全免费:依靠本地算力运行,不需要填 API Key ,也没有按分钟计费的困扰。 即开即用:界面保持极简,拖入音频即可直接进行本地转写。 目前的权衡与折衷 相比直接走云端高性能 GPU 集群,本地推理的速度会直接取决于用户的设备配置与硬件性能,老旧设备上可能存在内存占

Transcribe podcasts to text with speaker labels and AI summaries.
u/vadeconry · Reddit
I built a tool that turns Podcast URLs into transcripts and summaries Hi everyone, I am a software developer by profession, and over the past few months I've been working on a side project called PodTyper. You can access it at https://www.podtyper.com Instead of downloading audio files or uploading episodes manually, you just paste a podcast URL (Spotify, YouTube or Apple Podcasts) and it generates a transcript along with AI-powered summaries. I originally built it becau

Segment Anything for Audio — made simple. Separate audio with a text prompt. Isolate vocals, remove background music, extract speech from noise. Privacy-first stem splitter powered
@sambamdamnn · X

Upload or paste audio and video files to convert them to searchable, editable text transcripts.
yongkunchen · V2EX
新上线了一个音视频转文字的小工具 最近抽时间做了一个小工具 AnyToTranscript ,主要是把音频和视频快速转换成文本,支持时间戳和说话人识别。 目前已经上线了,自己也还在不断优化。如果你平时有整理访谈、会议、课程或者视频字幕的需求,欢迎体验一下: https://anytotranscript.com/ 如果你愿意试用,也很欢迎告诉我使用过程中遇到的问题,或者有哪些功能是你觉得值得加的。对独立开发者来说,这些反馈真的很有帮助。

Convert MP3 and MP4 files to editable transcripts with timestamps and speaker identification.
digiplanp · V2EX
做了一个音视频转文字工具,支持 MP3/MP4、说话人识别和字幕导出 大家好,最近做了一个在线音视频转文字工具 ToText ,想发出来请大家体验一下,也听听 V 友的意见。 最开始做这个工具,是因为我自己在处理会议录音、访谈和视频内容时,发现单纯生成一大段文字并不太实用。转录完成后,通常还需要反复对照原始录音、修改错误、区分说话人,再整理摘要或字幕。 所以现在做成了一个相对完整的工作流: * 上传 MP3 、MP4 、M4A 、MOV 、WAV 、WebM 等音视频文件 * 自动识别语言并生成带时间戳的文字稿 * 支持说话人识别和说话人名称修改 * 点击文字片段可以跳转到对应的音视频位置 * 可以搜索、编辑和校对转录内容 * 自动生成摘要、会议纪要、章节和行动事项 * 支持导出 TXT 、DOCX 、PDF 、SRT 、VTT 、CSV 和 JSON 音频文件可以直接使用 [MP3 to Transcript]( https://totext.app/) 功能,比较适合播客、访谈、课程录音、会议和语音备忘录。 视频文件也使用同一个编辑器,通过

Download Microsoft Text-to-Speech audio files with one click, supporting all official voices and SSML syntax.
微软 TTS (Text-to-Speech) 文字转语音下载器 — 一键播放或下载 微软 TTS 文字转语音 合成的音频。支持所有官方提供的语音和声音选项。支持SSML语法。每月享免费用量。量大者可升级Pro Plan,固定便宜费用无限使用。摒弃按量收费。支持支付宝

Automatically generate audio descriptions for videos to improve accessibility.
u/Thecuriousbloke · Reddit
I built an AI tool that automatically generates audio descriptions for videos. I'd love some honest feedback I've spent the last few months building an AI tool that automatically generates audio descriptions for videos to make them accessible for visually impaired users. I'd love feedback on whether the descriptions are actually useful or if there are obvious issues. Looking for honest feedback. Check it out here - https://accessly.studio submitted by /u/Thecuriousbloke t

API and MCP server that transcribes social videos and podcasts with timestamps.
chandler_casey · Product Hunt
TranscriptFetch Video and audio transcripts, structured for AI