
SHRP — Speech to Text, Transcription, YouTube Transcripts & TTS
Transcribe audio and video to text, create summaries and captions, and generate speech.
@bydigiwares · X
The full gallery
Tech stack
55 projects

Transcribe audio and video to text, create summaries and captions, and generate speech.
@bydigiwares · X

Generate 1080p watermark-free AI videos with Sora 2 motion controls, synchronized audio, and no signup required.
Sora2 AI — 体验 Sora 2 视频生成 – 创建 1080p 无水印视频,配备同步音频

Create 1080p videos with native audio and camera-level control.
Veo 3.1 AI — 使用 Veo3.1 创作电影级视频,配备 1080p 原生音频、首尾帧控制及免费试用额度

Download Microsoft Text-to-Speech audio files with one click, supporting all official voices and SSML syntax.
微软 TTS (Text-to-Speech) 文字转语音下载器 — 一键播放或下载 微软 TTS 文字转语音 合成的音频。支持所有官方提供的语音和声音选项。支持SSML语法。每月享免费用量。量大者可升级Pro Plan,固定便宜费用无限使用。摒弃按量收费。支持支付宝

AI audio agent for generating songs with vocals, isolating stems, converting to MIDI, and mastering tracks.
Gliss — AI Music Agent,用于歌曲生成、翻唱、MIDI 编辑、母带处理与封面艺术;可生成免版税人声与伴奏,并支持精确分轨/元素提取。

Learn Farsi free with audio lessons, alphabet guides, grammar, and spaced-repetition practice.
u/BigBoyWeazle · Reddit
Launching our first iOS app was brutal… but it paid off🚀 Hey everyone 😊 We recently launched our first iPhone app, Learn Farsi iOS and man o man, what a ride it was! Some early numbers in the first 4 weeks: • 290+ downloads • 23 paid users, $700+ total revenue so far • Now at ~$100 MRR I know these numbers aren’t huge, but it honestly feels crazy to see people pay for something we built. It really shows that building for a small, underserved niche can work. A

Generate personalized illustrated storybooks and audiobooks with consistent AI characters.
@StoryStitchApp · X

Generate AI videos from text prompts with Flux 3, including scene planning and audio direction.

Convert images, video, audio, and documents without uploading a thing. Everything runs on your device. 800+ conversions on Mac, 200+ in the browser. First conversion is free, then
@yogendrasinghx · X
Alright, roast this Privacy-first file converter for Mac 800+ conversions, everything happens locally. Don't hold back 😭

Split songs into vocals, drums, bass, and instruments with browser-based AI.
gaoheyangnwpu · V2EX
AudioCraft – AI stem splitter and vocal remover with a browser-based mixer I’ve been building AudioCraft as a focused AI music tool for separating songs into useful stems. The main workflow is simple: * upload MP3, WAV, FLAC, or M4A * separate the track into vocals, drums, bass, and other instruments * preview the stems * mute, solo, and adjust individual tracks * create an instrumental version * download individual stems or all stems together AI Stem Splitter https://audiocraft.app/ A

Create AI videos from image, video, and audio references with control over style and composition.
luya · V2EX
独立开发又折腾了个 AI 视频小站, V 友登录送一次生成 最近又折腾了个小站: https://referencetovideo.net 主要是做 reference to video 。 起因是我自己最近一直在玩 AI 视频,发现现在不少模型已经不只是文生视频、图生视频了,开始可以塞参考图、参考视频、音频进去。 但每家支持的东西又不太一样,用起来挺乱的。 所以就想干脆做个站,把这些东西放到一个地方。 现在大概就是: 上传图片 / 视频 / 音频 写 prompt 选模型 生成视频 目前还比较早期,边做边改。 我自己比较常用的场景是: * 拿人物图继续生成视频 * 产品图做广告 * 拿一段视频参考动作或者镜头 * 几种素材混着用 技术栈还是比较常规: Next.js PostgreSQL Stripe Vercel 再接一些视频模型 API 这次也给 V 友留了点福利。 **注册登录送一次免费生成次数**,不用付费,可以直接试一下。 地址: https://refere

Free browser-based utilities and tools with an audio toolkit, no signup required.
@KaiGeeks · X
我也做了一个类似 的网站: