
The full gallery
Tech stack
14 projects


Upload a PowerPoint presentation and voice sample to generate a narrated slideshow in your cloned voice.
u/sludge_dev · Reddit
I built a tool that narrates PowerPoint slides in your cloned voice. Launching it today, keeping it small for a week to fix what breaks. I just got into my College holidays a few weeks ago and I spent the last 6 weeks building Moduvox . You upload a PPTX deck, record 30 seconds of your voice, and it generates per-slide narration audio with a shareable link and viewer analytics. The idea came from a friend manually record voiceovers for training decks. Every update meant re-recording whole

Edit, compress, convert, and merge video and audio files in your browser.
zhw2590582 · V2EX
最近对网站进行了重构,做成了免费的视频工具站:[ArtPlayer Tools]( https://artplayer.org/tools/) 主要是围绕着 ffmpeg.wasm 和 mediabunny 实现的一些可以直接在浏览器里用的视频/音频工具,常用的有: - 视频压缩:[Compress Video]( https://artplayer.org/compress-video/) - 视频裁剪:[Crop Video]( https://artplayer.org/crop-video/) - 视频合并:[Merge Video]( https://artplayer.org/merge-video/) - 视频尺寸调整:[Resize Video]( https://artplayer.org/resize-video/) - 视频旋转:[Rotate Video]( https://artplayer.org/rotate-video/) - 视频调速:[Change Video Speed]( https://artplayer.org/change-video-speed/) - 添加文字到视频:[Add Text to Video]( https://artplayer.org/add-text-to-video/) - 添加水印:[Watermark Video]( https://artplayer.org/watermark-video/) - 提取字幕:[Subtitle Editor]( https://artplayer.org/subtitle-editor/) - HLS 转 MP4:[HLS to MP4]( https://artplayer.org/hls-to-mp4/) - 音频可视化:[Audio Visualizer]( https://artplayer.org/audio-visualizer/) 大部分工具都是本地处理,文件不会上传到服务器。最开始是因为自己偶尔要处理一些视频,开剪辑软件又有点重,在线工具又担心隐私,所以就慢慢做成了一个站。 目前还在持续完善,体验肯定还有不少粗糙的地方,比如不同浏览器的兼容、移动端操作、长视频性能之类的。 网站: [ArtPlayer Tools]( https:

Practice whistling with real-time pitch feedback and progressive exercises.
x1shir · V2EX
HowToWhistle:像唱 KTV 一样练习吹口哨 一直想提高吹口哨技巧,搜了一圈发现网上的教程都是文字和视频,练的时候根本不知道自己吹得对不对。 想到可以像 KTV 一样,有个音轨对比,于是干脆自己搓了一个练习网站:[https://howtowhistle.org/]( https://howtowhistle.org/) 核心思路是用浏览器的麦克风做实时音高检测,你对着麦克风吹,页面会显示你当前的音高和目标音的偏差,类似 KTV 唱歌时的实时评分。功能大概有: - 实时音高反馈,能看到自己吹的是哪个音、偏高还是偏低 - 从单音长音开始的渐进式练习,到音阶、再到简单旋律 - 纯前端实现,声音不上传,无需注册 目前还比较简陋,欢迎大家试玩和提建议。

Convert voice recordings into arrangements for any of 100+ instruments using AI.
Voice to Instrument — AI 工具,将人声录音转换为器乐曲目。上传歌声或录音,AI 自动生成钢琴、吉他、鼓等器乐伴奏。

Split songs into vocals, drums, bass, and instruments with browser-based AI.
gaoheyangnwpu · V2EX
AudioCraft – AI stem splitter and vocal remover with a browser-based mixer I’ve been building AudioCraft as a focused AI music tool for separating songs into useful stems. The main workflow is simple: * upload MP3, WAV, FLAC, or M4A * separate the track into vocals, drums, bass, and other instruments * preview the stems * mute, solo, and adjust individual tracks * create an instrumental version * download individual stems or all stems together AI Stem Splitter https://audiocraft.app/ A

Convert documents and e-books into audiobooks with natural-sounding narration.
@coder_zi · X
Clipifai converts e-books to audiobooks say you have a few books in mind to read, but haven't really had the time to read them, you could just convert them into audiobooks and listen to them on the go...

Reads articles aloud in your browser with podcast-quality voices.
DmitryDolgopolo · HN
ReadAloud – on-device, podcast-quality text-to-speech in the browser

AI audio agent for generating songs with vocals, isolating stems, converting to MIDI, and mastering tracks.
Gliss — AI Music Agent,用于歌曲生成、翻唱、MIDI 编辑、母带处理与封面艺术;可生成免版税人声与伴奏,并支持精确分轨/元素提取。

Generate AI videos up to 720p with text or images, including native dialogue and sound.
Seedance 2.0 — 创作电影级 AI 视频,支持多个模型的图片视频生成网站

Separate any song into professional stems like vocals, drums, bass, and guitar.
AI Stem Splitter — AI 音轨分离器

Copy curated prompts for AI image, video, and audio generation.
weijunext · HN
Free, curated prompts for AI image, video and music