
完整作品展
技术栈
19 projects


在浏览器中用播客级别的语音朗读文章。
DmitryDolgopolo · HN
ReadAloud – on-device, podcast-quality text-to-speech in the browser

将 MP3 和 MP4 文件转换为带有时间戳和说话人识别的可编辑文字稿。
digiplanp · V2EX
做了一个音视频转文字工具,支持 MP3/MP4、说话人识别和字幕导出 大家好,最近做了一个在线音视频转文字工具 ToText ,想发出来请大家体验一下,也听听 V 友的意见。 最开始做这个工具,是因为我自己在处理会议录音、访谈和视频内容时,发现单纯生成一大段文字并不太实用。转录完成后,通常还需要反复对照原始录音、修改错误、区分说话人,再整理摘要或字幕。 所以现在做成了一个相对完整的工作流: * 上传 MP3 、MP4 、M4A 、MOV 、WAV 、WebM 等音视频文件 * 自动识别语言并生成带时间戳的文字稿 * 支持说话人识别和说话人名称修改 * 点击文字片段可以跳转到对应的音视频位置 * 可以搜索、编辑和校对转录内容 * 自动生成摘要、会议纪要、章节和行动事项 * 支持导出 TXT 、DOCX 、PDF 、SRT 、VTT 、CSV 和 JSON 音频文件可以直接使用 [MP3 to Transcript]( https://totext.app/) 功能,比较适合播客、访谈、课程录音、会议和语音备忘录。 视频文件也使用同一个编辑器,通过

API 和 MCP 服务器,可将社交视频和播客转录为带时间戳的文本。
chandler_casey · Product Hunt
TranscriptFetch Video and audio transcripts, structured for AI

用AI转录音频和视频为文本和摘要
@Mahima_Akkina · X
Yes, you can check - It's a tool used to convert both and video into text and summaries

向多个前沿大模型提问,获得经过同行评审的综合答案。
u/Puzzleheaded-Log-27 · Reddit
Building a multi-model AI deliberation tool taught me something about trust LLM Counsel isn't another wrapper around one model - it sends your question to a panel of frontier LLMs, has them peer-review each other anonymously, and an impartial "chairman" model returns one synthesized answer. Free to start, pay-as-you-go after, credits don't expire. What I've learned so far: people trust a synthesized answer a lot more once they can see that the models actually disagreed and how that disagree

在 X 上发现 ICP 匹配的目标客户,并用你的风格起草开场消息。
@aman29122k · X
- it's your one in all twitter companion