
Segment Anything for Audio — made simple. Separate audio with a text prompt. Isolate vocals, remove background music, extract speech from noise. Privacy-first stem splitter powered
@sambamdamnn · X
The full gallery
Tech stack
24 projects

Segment Anything for Audio — made simple. Separate audio with a text prompt. Isolate vocals, remove background music, extract speech from noise. Privacy-first stem splitter powered
@sambamdamnn · X

Swap your face with another in real-time during video calls and streams using your webcam.
mixfox · HN
An online Live face swap app, no GPU need

AI copilot that records, transcribes, and extracts insights from business calls in the browser.
u/carlosmarcialt · Reddit
I spent months building my SaaS and completely forgot the marketing part I just realized Voxandra has had exactly zero marketing. Not bad marketing. Literally none. So this Reddit post is now the entire marketing department. Voxandra is an AI phone copilot for people and teams who still get real work done over phone calls. Supplier calls, payment follow ups, claims, banks, government offices, universities, and all the other places that somehow still behave like email was never invented.

Recover up to +32% of abandoned carts with AI voice agents. Live in 5 minutes on Shopify, WooCommerce, WordPress, Klaviyo, Omnisend. 50% off launch promo.
@ArunasVismantas · X

Create ASMR videos using AI-powered Veo3 technology.
AI ASMR视频生成器 — AI ASMR视频生成器(采用先进 Veo3 技术)

AI voice agents for handling phone calls on an integrated stack, starting at 2 cents per minute.
kolchinski · HN
ThunderPhone v2 – a new architecture for voice AI

Transcribe audio and video to text, create summaries and captions, and generate speech.
@bydigiwares · X

Transcribe audio and video files to text with AI-generated summaries.
@Mahima_Akkina · X
Yes, you can check - It's a tool used to convert both and video into text and summaries

Generate AI videos from text prompts with Flux 3, including scene planning and audio direction.

Generate 4K videos from text prompts with AI-rendered visuals, sound effects, and lip-syncing.
Veo 3 AI — Veo 3 AI 视频生成

Create text, images, video, and audio with AI chat and 100+ tools in one place.
Fullmira — 一站式 AI 创作平台,整合 GPT、Claude、Gemini 等多个模型,同一界面完成文本、图像、视频和音频生成,内置 100+ 专用工具覆盖写作、图像编辑与视频制作

Convert audio and video files to text with speaker labels and timestamps, no signup required.
zxhywork · V2EX
SayScribe AI - 音频转文字的在线工具 花了几周搞了一个基于 OpenAI Whisper 的 AI 转录工具 **[SayScribe]( https://sayscribe.org)**,目前 5 个工具全部免费,每天 3 次无需注册。上线前请 SEO 专家做了一轮诊断,踩了不少坑,分享出来给同样在做工具站的 V 友参考。 --- ## 做了什么 核心工具是 **[Audio to Text Converter]( https://sayscribe.org)**(主页),上传音频/视频文件,AI 自动转成带 speaker labels 和时间戳的文字稿。选这个词的原因是:Google 第一页有个 DR 只有 10 的小站排第三,KD 才 20.2 ,链接预算只要 15-35 个引用域就能进前十——典型的弱盘面蓝海。 然后在主页基础上拆了 4 个子工具,每个打一个独立搜索词: - **[MP3 to Text]( https://sayscribe.org/mp3-to-text)** — 只转 MP3 。KD 41.7 比主页难一些,但