
AI Text-to-Speech — Convert Text & PDF to MP3 | audioso
使用 AI 语音将文本和 PDF 转换为自然发音的音频
@caschiblu · X
完整作品展
技术栈
33 projects

使用 AI 语音将文本和 PDF 转换为自然发音的音频
@caschiblu · X

用于语音、聊天和视频面试的 AI 平台,提供自动转录和分析。
Aural — 开源 AI 面试平台,支持语音、聊天和视频面试,提供自适应追问、结构化评分、面试练习与自托管 - [查看仓库](https://github.com/1146345502/aural-oss)

用AI将图像和文本转换成电影级视频,支持动作控制。
Vibe Video — 面向普通用户的 AI 视频生成网站,支持文生视频、图生视频、参考图生视频和电影化镜头控制,适合创作者和营销团队快速产出视频

上传音频或视频,SayScribe 自动转成带说话者标签和时间戳的文字稿。
zxhywork · V2EX
SayScribe AI - 音频转文字的在线工具 花了几周搞了一个基于 OpenAI Whisper 的 AI 转录工具 **[SayScribe]( https://sayscribe.org)**,目前 5 个工具全部免费,每天 3 次无需注册。上线前请 SEO 专家做了一轮诊断,踩了不少坑,分享出来给同样在做工具站的 V 友参考。 --- ## 做了什么 核心工具是 **[Audio to Text Converter]( https://sayscribe.org)**(主页),上传音频/视频文件,AI 自动转成带 speaker labels 和时间戳的文字稿。选这个词的原因是:Google 第一页有个 DR 只有 10 的小站排第三,KD 才 20.2 ,链接预算只要 15-35 个引用域就能进前十——典型的弱盘面蓝海。 然后在主页基础上拆了 4 个子工具,每个打一个独立搜索词: - **[MP3 to Text]( https://sayscribe.org/mp3-to-text)** — 只转 MP3 。KD 41.7 比主页难一些,但

上传音频或视频文件,转换为可搜索的文本记录。
yongkunchen · V2EX
新上线了一个音视频转文字的小工具 最近抽时间做了一个小工具 AnyToTranscript ,主要是把音频和视频快速转换成文本,支持时间戳和说话人识别。 目前已经上线了,自己也还在不断优化。如果你平时有整理访谈、会议、课程或者视频字幕的需求,欢迎体验一下: https://anytotranscript.com/ 如果你愿意试用,也很欢迎告诉我使用过程中遇到的问题,或者有哪些功能是你觉得值得加的。对独立开发者来说,这些反馈真的很有帮助。

用文本或图像生成720p AI视频,包含自然对白和音效。
Seedance 2.0 — 创作电影级 AI 视频,支持多个模型的图片视频生成网站

将文档和文本转成播客,支持 AI 语音生成和编辑。
Inpodcast AI — 将文档转成播客音频,支持 PDF、Word、Markdown 和 TXT 文件格式

从脚本生成短视频,AI自动生成场景、字幕、声音和动画。
u/HintyAI · Reddit
What I learned from trying to reduce a three-hour video workflow to ten minutes While working on marketing for my previous project, I kept running into the same bottleneck: even a fairly simple 30-second animated promo could take me two or three hours. Writing the script was not the difficult part. Most of the time went into finding suitable visuals, synchronizing them with the narration, adding motion and captions, and then revising everything when one scene did not work. I tried several

上传PowerPoint和语音样本,用您的克隆声音生成完整讲述。
u/sludge_dev · Reddit
I built a tool that narrates PowerPoint slides in your cloned voice. Launching it today, keeping it small for a week to fix what breaks. I just got into my College holidays a few weeks ago and I spent the last 6 weeks building Moduvox . You upload a PPTX deck, record 30 seconds of your voice, and it generates per-slide narration audio with a shareable link and viewer analytics. The idea came from a friend manually record voiceovers for training decks. Every update meant re-recording whole

使用文本、图像和音频混合提示生成2K AI视频,5-15秒,原生立体声。

从文本和图像生成 4K+ 电影级视频,支持多镜头序列和人物持久化。
kling — AI 生成视频,支持多种 kling 模型

克隆你的声音创作和讲述原创故事。
u/Aware_Web9715 · Reddit
Create stories in your voice Built using open source models (Qwen 3.6 and TTS). Clone your voice and create an original story for free. Just a fun weekend project that runs on my spark. Feedback gets reviewed daily and the coding agents attempts to create a PR. https://app.getzoro.com submitted by /u/Aware_Web9715 to r/SideProject [link] [comments]