
oruk — Speech API for transcripts, emotion, and style
Audio API for transcription, emotion detection, and speaking style analysis.
dillon_silzer · Product Hunt
Oruk Speech AI that actually understands you
The full gallery
Tech stack
28 projects

Audio API for transcription, emotion detection, and speaking style analysis.
dillon_silzer · Product Hunt
Oruk Speech AI that actually understands you

Segment Anything for Audio — made simple. Separate audio with a text prompt. Isolate vocals, remove background music, extract speech from noise. Privacy-first stem splitter powered
@sambamdamnn · X

Conduct AI-powered voice, chat, and video interviews with automated transcripts and insights.
Aural — 开源 AI 面试平台,支持语音、聊天和视频面试,提供自适应追问、结构化评分、面试练习与自托管 - [查看仓库](https://github.com/1146345502/aural-oss)

Convert audio and video to searchable transcripts with AI.
giang_taira · Product Hunt
GPT Transcribe Turn recorded speech into reviewable text

Upload AI conversations, get an evidence-backed report.
@StilThinkng · X
Tired of repeating the same context across ChatGPT, Claude, and Grok? SoulBots builds a privacy-first memory that grows with you across AI. Looking for early testers.

Automatically generate audio descriptions for videos to improve accessibility.
u/Thecuriousbloke · Reddit
I built an AI tool that automatically generates audio descriptions for videos. I'd love some honest feedback I've spent the last few months building an AI tool that automatically generates audio descriptions for videos to make them accessible for visually impaired users. I'd love feedback on whether the descriptions are actually useful or if there are obvious issues. Looking for honest feedback. Check it out here - https://accessly.studio submitted by /u/Thecuriousbloke t

Convert text and PDFs to audio with natural-sounding AI voices.
@caschiblu · X

AI audio editor and music generator for editing, shortening, and creating music from text or audio files.
audjust.ai — 智能音频编辑与 AI 音乐生成工具,帮助处理音频文件(智能缩短歌曲、延长音频、寻找完美无缝循环)并从文字描述、歌词或图片生成完整音乐轨道。支持多种风格,内置 MIDI 编辑器、音轨分离、Lo-fi 转换等专业工具。

Generate AI voices, sound effects, and music from text prompts in multiple languages.

Combine AI-generated images, videos, and audio on an infinite browser canvas.
ZOOOP — AI 原生创作平台,支持在浏览器端无限画布上生成图像、视频和音频,提供去背景、高清化等专项工具及即用型 AI 模板,支持发布模板赚取积分,并具备团队共享积分模式。

AI audio agent for generating songs with vocals, isolating stems, converting to MIDI, and mastering tracks.
Gliss — AI Music Agent,用于歌曲生成、翻唱、MIDI 编辑、母带处理与封面艺术;可生成免版税人声与伴奏,并支持精确分轨/元素提取。

Download Microsoft Text-to-Speech audio files with one click, supporting all official voices and SSML syntax.
微软 TTS (Text-to-Speech) 文字转语音下载器 — 一键播放或下载 微软 TTS 文字转语音 合成的音频。支持所有官方提供的语音和声音选项。支持SSML语法。每月享免费用量。量大者可升级Pro Plan,固定便宜费用无限使用。摒弃按量收费。支持支付宝