
AI Text-to-Speech — Convert Text & PDF to MP3 | audioso
Convert text and PDFs to audio with natural-sounding AI voices.
@caschiblu · X
The full gallery
Tech stack
30 projects

Convert text and PDFs to audio with natural-sounding AI voices.
@caschiblu · X

Convert text and PDFs to natural-sounding speech with AI voices, download as MP3.
@caschiblu · X

Generate AI voices with text-to-speech, voice cloning, and voice design.
Sonicker — AI 语音克隆平台,3 秒克隆任何声音,保留情感和口音。支持中英日韩等 10 种语言的语音合成,提供 50+ 预设声音库和 AI 声音设计功能

Split songs into vocals, drums, bass, and instruments with browser-based AI.
gaoheyangnwpu · V2EX
AudioCraft – AI stem splitter and vocal remover with a browser-based mixer I’ve been building AudioCraft as a focused AI music tool for separating songs into useful stems. The main workflow is simple: * upload MP3, WAV, FLAC, or M4A * separate the track into vocals, drums, bass, and other instruments * preview the stems * mute, solo, and adjust individual tracks * create an instrumental version * download individual stems or all stems together AI Stem Splitter https://audiocraft.app/ A

Convert voice recordings into arrangements for any of 100+ instruments using AI.
Voice to Instrument — AI 工具,将人声录音转换为器乐曲目。上传歌声或录音,AI 自动生成钢琴、吉他、鼓等器乐伴奏。

Generate AI voices, sound effects, and music from text prompts in multiple languages.

Save multimedia content and turn it into searchable summaries and flashcards powered by AI.
@ShahirSiddiqu13 · X
Memora turns everything you save into a searchable second brain powered by AI. Try it: 🌐 What would you save first? #BuildInPublic #SaaS #AI #Productivity

Upload a PowerPoint presentation and voice sample to generate a narrated slideshow in your cloned voice.
u/sludge_dev · Reddit
I built a tool that narrates PowerPoint slides in your cloned voice. Launching it today, keeping it small for a week to fix what breaks. I just got into my College holidays a few weeks ago and I spent the last 6 weeks building Moduvox . You upload a PPTX deck, record 30 seconds of your voice, and it generates per-slide narration audio with a shareable link and viewer analytics. The idea came from a friend manually record voiceovers for training decks. Every update meant re-recording whole


Conduct AI-powered voice, chat, and video interviews with automated transcripts and insights.
Aural — 开源 AI 面试平台,支持语音、聊天和视频面试,提供自适应追问、结构化评分、面试练习与自托管 - [查看仓库](https://github.com/1146345502/aural-oss)

Generate emotionally expressive speech with ProsodyAI — now with OpenAI and Google Gemini voices. 59 languages, 8 emotion presets, zero-shot voice cloning, 100,000 free characters
@prosodyai · X
ProsodyAI — AI text-to-speech that performs, not just pronounces. Tell it how to sound ("warm, slow, like a bedtime story") and it does. Three engines in one app: ours, OpenAI and Gemini. 100,000 characters free every month, voice cloning included.

AI audio agent for generating songs with vocals, isolating stems, converting to MIDI, and mastering tracks.
Gliss — AI Music Agent,用于歌曲生成、翻唱、MIDI 编辑、母带处理与封面艺术;可生成免版税人声与伴奏,并支持精确分轨/元素提取。