
oruk — Speech API for transcripts, emotion, and style
Audio API for transcription, emotion detection, and speaking style analysis.
dillon_silzer · Product Hunt
Oruk Speech AI that actually understands you
The full gallery
Tech stack
77 projects

Audio API for transcription, emotion detection, and speaking style analysis.
dillon_silzer · Product Hunt
Oruk Speech AI that actually understands you

Convert and edit audio, images, PDFs, and videos with 150+ free online tools, no signup required.
@factfile0 · X

Free media hosting with permanent direct links, embed codes, and global CDN. Upload images, videos, and audio — no account required.
@dbimgapp · X

Read EPUBs, PDFs, and text files aloud using on-device AI voices on Mac and Windows.
u/Right_Sandwich_7304 · Reddit
Need feedback on the text to speech book reader I've built the book reader app which kind of reads as good as elevenlabs or speechify but for free because its easy to build such an apps these days with AI 😄. This app uses native CPU to generate naration of book and it feels like you are listening to the audiobook recorded by professionals. Please try it on https://dp.openaloud.com/ the app is only available for windows and mac. I am working on launching android and ios apps so

Generate 768p AI videos from text or images with native audio in seconds.
plutozc · V2EX
做了两个 MiniMax H3 AI 视频生成入口 最近在用 MiniMax H3 做一些 AI 视频的快速试验,发现自己常在文字、图片参考、画幅和时长之间来回切,所以顺手做了两个偏不同工作流的入口,想听听大家会不会觉得这种拆分有用。 一个是偏快速迭代的版本: [H3 Max]( https://h3-max.org) 主打从文字或参考图生成短视频,重点放在快速生成 768p 片段和提示词里的画面、镜头、音频描述。 另一个是功能更完整的工作台: [MiniMax H3 Max]( https://minimaxh3max.org) 同样支持 text-to-video 和 image-to-video ,可以按需要选 768p 、2K 或 4K ,也能调时长和横竖画幅。 目前都还是刚上线的个人项目: - 文字或图片作为起点生成视频 - 一个提示词里写画面、动作、镜头和声音方向 - 方便先快速出一个版本,再回头改提示词 两站都不是官方站,生成结果和速度会受模型设置及队列影响。现在页面和体验还在继续补,尤其想听听大家更在意“快点出第一版”,

Generate chord progressions by choosing key, style, and mood with instant audio preview.

Copy curated prompts for AI image, video, and audio generation.
weijunext · HN
Free, curated prompts for AI image, video and music

Combine AI-generated images, videos, and audio on an infinite browser canvas.
ZOOOP — AI 原生创作平台,支持在浏览器端无限画布上生成图像、视频和音频,提供去背景、高清化等专项工具及即用型 AI 模板,支持发布模板赚取积分,并具备团队共享积分模式。

Compare speech-to-text engines (OpenAI, Deepgram, NVIDIA, Fish Audio) with real-time benchmarking and local privacy.
@alvaisy · X
finished voice to text small web app for my own itch. it's opensource. use openrotuer key. and use it with 4 models.

Create videos from text, images, or videos with automatically synchronized audio.
FLUX 3 Video Generator — AI 视频生成工作台,支持文字或图片生成视频、参考图及镜头、动作和声音提示,并清晰标注模型可用状态

Identify songs from audio files, URLs, or live recordings without signing up.
krasscy · V2EX
做了个免费的找歌工具,上传音频或贴链接就能识别 最近做了一个 [Song Finder]( https://songfinder.dev/),遇到听到一段歌却想不起名字的情况,可以上传音频、粘贴媒体链接,或者直接录一段声音来识别。 基础功能免费,不需要注册。除了找歌,也放了音频裁剪、降噪、BPM 和调性检测、歌词等小工具。 欢迎试用,也想听听大家对识别准确度和使用体验的反馈。 https://songfinder.dev

Generate AI images, videos, and audio using Apatero Studio's web-based creation tools.
@gabeciii · X