
AI Text-to-Speech — Convert Text & PDF to MP3 | audioso
使用 AI 语音将文本和 PDF 转换为自然发音的音频
@caschiblu · X
完整作品展
技术栈
30 projects

使用 AI 语音将文本和 PDF 转换为自然发音的音频
@caschiblu · X

将文本和PDF转换为AI自然语音并下载为MP3。
@caschiblu · X

使用 Sonicker 生成 AI 语音,支持文本转语音、语音克隆和自定义设计。
Sonicker — AI 语音克隆平台,3 秒克隆任何声音,保留情感和口音。支持中英日韩等 10 种语言的语音合成,提供 50+ 预设声音库和 AI 声音设计功能

用浏览器AI工具将歌曲分离成人声、鼓声、贝斯和其他乐器。
gaoheyangnwpu · V2EX
AudioCraft – AI stem splitter and vocal remover with a browser-based mixer I’ve been building AudioCraft as a focused AI music tool for separating songs into useful stems. The main workflow is simple: * upload MP3, WAV, FLAC, or M4A * separate the track into vocals, drums, bass, and other instruments * preview the stems * mute, solo, and adjust individual tracks * create an instrumental version * download individual stems or all stems together AI Stem Splitter https://audiocraft.app/ A

将您的人声转换为钢琴、吉他、小提琴等100+种乐器演奏。
Voice to Instrument — AI 工具,将人声录音转换为器乐曲目。上传歌声或录音,AI 自动生成钢琴、吉他、鼓等器乐伴奏。


用AI将保存的视频、PDF和文章转换为可搜索的总结和闪卡。
@ShahirSiddiqu13 · X
Memora turns everything you save into a searchable second brain powered by AI. Try it: 🌐 What would you save first? #BuildInPublic #SaaS #AI #Productivity

上传PowerPoint和语音样本,用您的克隆声音生成完整讲述。
u/sludge_dev · Reddit
I built a tool that narrates PowerPoint slides in your cloned voice. Launching it today, keeping it small for a week to fix what breaks. I just got into my College holidays a few weeks ago and I spent the last 6 weeks building Moduvox . You upload a PPTX deck, record 30 seconds of your voice, and it generates per-slide narration audio with a shareable link and viewer analytics. The idea came from a friend manually record voiceovers for training decks. Every update meant re-recording whole


用于语音、聊天和视频面试的 AI 平台,提供自动转录和分析。
Aural — 开源 AI 面试平台,支持语音、聊天和视频面试,提供自适应追问、结构化评分、面试练习与自托管 - [查看仓库](https://github.com/1146345502/aural-oss)

Generate emotionally expressive speech with ProsodyAI — now with OpenAI and Google Gemini voices. 59 languages, 8 emotion presets, zero-shot voice cloning, 100,000 free characters
@prosodyai · X
ProsodyAI — AI text-to-speech that performs, not just pronounces. Tell it how to sound ("warm, slow, like a bedtime story") and it does. Three engines in one app: ours, OpenAI and Gemini. 100,000 characters free every month, voice cloning included.

Gliss是AI音乐代理,可生成歌曲、分离音轨、转MIDI和母带处理。
Gliss — AI Music Agent,用于歌曲生成、翻唱、MIDI 编辑、母带处理与封面艺术;可生成免版税人声与伴奏,并支持精确分轨/元素提取。