
ReadAloud — Podcast-quality voices in your browser
在浏览器中用播客级别的语音朗读文章。
DmitryDolgopolo · HN
ReadAloud – on-device, podcast-quality text-to-speech in the browser
完整作品展
技术栈
12 projects

在浏览器中用播客级别的语音朗读文章。
DmitryDolgopolo · HN
ReadAloud – on-device, podcast-quality text-to-speech in the browser

在浏览器中编辑、压缩、转换和合并视频和音频文件。
zhw2590582 · V2EX
最近对网站进行了重构,做成了免费的视频工具站:[ArtPlayer Tools]( https://artplayer.org/tools/) 主要是围绕着 ffmpeg.wasm 和 mediabunny 实现的一些可以直接在浏览器里用的视频/音频工具,常用的有: - 视频压缩:[Compress Video]( https://artplayer.org/compress-video/) - 视频裁剪:[Crop Video]( https://artplayer.org/crop-video/) - 视频合并:[Merge Video]( https://artplayer.org/merge-video/) - 视频尺寸调整:[Resize Video]( https://artplayer.org/resize-video/) - 视频旋转:[Rotate Video]( https://artplayer.org/rotate-video/) - 视频调速:[Change Video Speed]( https://artplayer.org/change-video-speed/) - 添加文字到视频:[Add Text to Video]( https://artplayer.org/add-text-to-video/) - 添加水印:[Watermark Video]( https://artplayer.org/watermark-video/) - 提取字幕:[Subtitle Editor]( https://artplayer.org/subtitle-editor/) - HLS 转 MP4:[HLS to MP4]( https://artplayer.org/hls-to-mp4/) - 音频可视化:[Audio Visualizer]( https://artplayer.org/audio-visualizer/) 大部分工具都是本地处理,文件不会上传到服务器。最开始是因为自己偶尔要处理一些视频,开剪辑软件又有点重,在线工具又担心隐私,所以就慢慢做成了一个站。 目前还在持续完善,体验肯定还有不少粗糙的地方,比如不同浏览器的兼容、移动端操作、长视频性能之类的。 网站: [ArtPlayer Tools]( https:

将文档和电子书转换成自然语音的有声书。
@coder_zi · X
Clipifai converts e-books to audiobooks say you have a few books in mind to read, but haven't really had the time to read them, you could just convert them into audiobooks and listen to them on the go...

用于语音、聊天和视频面试的 AI 平台,提供自动转录和分析。
Aural — 开源 AI 面试平台,支持语音、聊天和视频面试,提供自适应追问、结构化评分、面试练习与自托管 - [查看仓库](https://github.com/1146345502/aural-oss)

通过语音说出您的费用,无需键入即可立即记录支出。
@smoothcode97 · X

用 AI 聊天和 100+ 工具在一处创建文本、图像、视频和音频。
Fullmira — 一站式 AI 创作平台,整合 GPT、Claude、Gemini 等多个模型,同一界面完成文本、图像、视频和音频生成,内置 100+ 专用工具覆盖写作、图像编辑与视频制作

用AI转录音频和视频为文本和摘要
@Mahima_Akkina · X
Yes, you can check - It's a tool used to convert both and video into text and summaries

Swiftener:150+免费在线工具,编辑音频、图像、PDF、视频等,无需注册。
@factfile0 · X

在伦敦画廊拍摄艺术作品,获取AI识别、音频导览和博物馆信息。
victords · HN
I was interested in learning more about reverse image search, so I built audioguide.london. It’s a simple image search where you take a photo of an artwork in a gallery, tries to ID it and returns a pre-generated audio and a link to the museum website. The audios were all generated locally, essentially looking at the contents of the website, running it through a LLM to generate a script and Kokoro for TTS. I’ve built it as an app for myself almost a year ago, so I deployed it as a vibe coded web

用 PodTLDR.fm 获取每日播客摘要和关键要点。
u/owocki · Reddit
PodTLDR.fm — I follow 12 podcasts and listen to none of them, so I built the thing that reads them for me I follow about a dozen long-form podcasts - Acquired, Lex, Hardcore History - and I was permanently 30+ episodes behind. Every new 3-hour drop just added to the guilt pile. I didn't want to listen to them. I wanted to have heard them. So: PodTLDR.fm. Add a podcast, and every time it drops an episode we transcribe it, summarize it, and email you the TLDR by morning - the one-line point,

将您的人声转换为钢琴、吉他、小提琴等100+种乐器演奏。
Voice to Instrument — AI 工具,将人声录音转换为器乐曲目。上传歌声或录音,AI 自动生成钢琴、吉他、鼓等器乐伴奏。

将文档和文本转成播客,支持 AI 语音生成和编辑。
Inpodcast AI — 将文档转成播客音频,支持 PDF、Word、Markdown 和 TXT 文件格式