
Try CueSesh — Build a podcast session that fits your time
Automatically build podcast sessions from a live catalog that fit your schedule.
@Akisin1 · X
Listen Like You Live
The full gallery
Tech stack
15 projects

Automatically build podcast sessions from a live catalog that fit your schedule.
@Akisin1 · X
Listen Like You Live

Edit, compress, convert, and merge video and audio files in your browser.
zhw2590582 · V2EX
最近对网站进行了重构,做成了免费的视频工具站:[ArtPlayer Tools]( https://artplayer.org/tools/) 主要是围绕着 ffmpeg.wasm 和 mediabunny 实现的一些可以直接在浏览器里用的视频/音频工具,常用的有: - 视频压缩:[Compress Video]( https://artplayer.org/compress-video/) - 视频裁剪:[Crop Video]( https://artplayer.org/crop-video/) - 视频合并:[Merge Video]( https://artplayer.org/merge-video/) - 视频尺寸调整:[Resize Video]( https://artplayer.org/resize-video/) - 视频旋转:[Rotate Video]( https://artplayer.org/rotate-video/) - 视频调速:[Change Video Speed]( https://artplayer.org/change-video-speed/) - 添加文字到视频:[Add Text to Video]( https://artplayer.org/add-text-to-video/) - 添加水印:[Watermark Video]( https://artplayer.org/watermark-video/) - 提取字幕:[Subtitle Editor]( https://artplayer.org/subtitle-editor/) - HLS 转 MP4:[HLS to MP4]( https://artplayer.org/hls-to-mp4/) - 音频可视化:[Audio Visualizer]( https://artplayer.org/audio-visualizer/) 大部分工具都是本地处理,文件不会上传到服务器。最开始是因为自己偶尔要处理一些视频,开剪辑软件又有点重,在线工具又担心隐私,所以就慢慢做成了一个站。 目前还在持续完善,体验肯定还有不少粗糙的地方,比如不同浏览器的兼容、移动端操作、长视频性能之类的。 网站: [ArtPlayer Tools]( https:

Convert voice recordings into arrangements for any of 100+ instruments using AI.
Voice to Instrument — AI 工具,将人声录音转换为器乐曲目。上传歌声或录音,AI 自动生成钢琴、吉他、鼓等器乐伴奏。

Photograph artwork in a gallery to get AI-identified details, audio guides, and museum information.
victords · HN
I was interested in learning more about reverse image search, so I built audioguide.london. It’s a simple image search where you take a photo of an artwork in a gallery, tries to ID it and returns a pre-generated audio and a link to the museum website. The audios were all generated locally, essentially looking at the contents of the website, running it through a LLM to generate a script and Kokoro for TTS. I’ve built it as an app for myself almost a year ago, so I deployed it as a vibe coded web

Conduct AI-powered voice, chat, and video interviews with automated transcripts and insights.
Aural — 开源 AI 面试平台,支持语音、聊天和视频面试,提供自适应追问、结构化评分、面试练习与自托管 - [查看仓库](https://github.com/1146345502/aural-oss)

Guided yoga sessions with beginner to advanced routines and text-to-speech voice instructions.
@shivanshsh1823 · X
you yoga a day - to - day workout routine for yoga

Organize and retrieve your work sessions with AI-powered memory.
@vitverb · X
hi :) One tap back to the work you meant to finish. No tasks to manage. Nothing to set up. Local-first

Transcribe audio and video files to text with AI-generated summaries.
@Mahima_Akkina · X
Yes, you can check - It's a tool used to convert both and video into text and summaries

Organize notes and collections, ask questions, and generate podcasts and study materials.
mistakevin · HN
I've been working on https://notebooker.ai if anyone is interested in giving some feedback. I tried to post a Show HN yesterday with details about how I built it, but was auto-flagged and not sure what rules I broke. Everything that I've built on top of open-notebook, like the plugin system for your own "creators" (aka, podcasts, infographics, etc) is at https://github.com/Notebooker-ai plus a Cloudflare Worker AI deployer to play with different models. Been working on it for about six months.

Convert audio and video files to text with speaker labels and timestamps, no signup required.
zxhywork · V2EX
SayScribe AI - 音频转文字的在线工具 花了几周搞了一个基于 OpenAI Whisper 的 AI 转录工具 **[SayScribe]( https://sayscribe.org)**,目前 5 个工具全部免费,每天 3 次无需注册。上线前请 SEO 专家做了一轮诊断,踩了不少坑,分享出来给同样在做工具站的 V 友参考。 --- ## 做了什么 核心工具是 **[Audio to Text Converter]( https://sayscribe.org)**(主页),上传音频/视频文件,AI 自动转成带 speaker labels 和时间戳的文字稿。选这个词的原因是:Google 第一页有个 DR 只有 10 的小站排第三,KD 才 20.2 ,链接预算只要 15-35 个引用域就能进前十——典型的弱盘面蓝海。 然后在主页基础上拆了 4 个子工具,每个打一个独立搜索词: - **[MP3 to Text]( https://sayscribe.org/mp3-to-text)** — 只转 MP3 。KD 41.7 比主页难一些,但

Transcribe audio files to text instantly with AI, no signup required.
音频转文字工具 — AI 音频转文字,快速、精准、安全的转写服务

Upload a PowerPoint presentation and voice sample to generate a narrated slideshow in your cloned voice.
u/sludge_dev · Reddit
I built a tool that narrates PowerPoint slides in your cloned voice. Launching it today, keeping it small for a week to fix what breaks. I just got into my College holidays a few weeks ago and I spent the last 6 weeks building Moduvox . You upload a PPTX deck, record 30 seconds of your voice, and it generates per-slide narration audio with a shareable link and viewer analytics. The idea came from a friend manually record voiceovers for training decks. Every update meant re-recording whole