
MiniMax H3: Free 2K AI Video Generator with Native Audio
Generate 2K AI videos (5-15 sec) from text, images, and audio references with native stereo audio.
The full gallery
Tech stack
60 projects

Generate 2K AI videos (5-15 sec) from text, images, and audio references with native stereo audio.

Web-based media editor with 57 tools for trimming, mixing, filtering, and converting audio, video, and images.
@beipojushi · X
做了一个一站式音视频/图片处理工具站。 57 个功能:视频(截取、拼接、变速、倒放、裁剪、压缩、转GIF、字幕、水印),音频(提取、混音、降噪、变速变调、人 声分离),图片(转换、缩放、滤镜、合成GIF)。 免费:游客每天5次,注册后每天20次。

Transcribe audio and video to text entirely on your device, with no uploads and no account needed.
mikicat · V2EX
[分享创造] 做了一个纯本地离线运行的语音转文字工具: Free Local Transcription 大家好, 平时经常有一些包含敏感信息或者不想上传到第三方服务器的音频(如会议记录、采访、个人备忘等)需要转文字。市面上很多工具本质上都是套壳调用云端 API ,既有数据泄露风险,长期用也有不小的 API 成本。 结合自己的模型优化经验,我做了一个完全基于本地算力运行的音频转文字工具:Free Local Transcription 🔗 体验地址: https://freelocaltranscription.com/ 主要特点 100% 本地运行 / 零数据外传:推理完全在设备本地完成,音频文件不会上传到任何服务器,断网状态下也能保证隐私与安全。 无需 API Key / 完全免费:依靠本地算力运行,不需要填 API Key ,也没有按分钟计费的困扰。 即开即用:界面保持极简,拖入音频即可直接进行本地转写。 目前的权衡与折衷 相比直接走云端高性能 GPU 集群,本地推理的速度会直接取决于用户的设备配置与硬件性能,老旧设备上可能存在内存占

Turn your PDF portfolio into an interactive online gallery with multiple viewing modes.
@il_deil · X
I’m building an interactive, ultra-minimalist portfolio platform strictly for architects & designers, because I tired of ads and clutter ruining portfolio on Issuu and same services. I made zero visual noise. Multiple viewing modes. Just your portfolio

Merge PDFs, compress images, remove backgrounds, and access utilities entirely in your browser.
@damiadeh · X
Your files utility tools stacks in your browser. No uploads, files stays on your pc.

Embed videos in PDF documents to track engagement and viewership.
@Cam_james05 · X

Converts study PDFs into flashcards, quizzes, and audio podcasts with weak-area detection.
@indiesaasgrowth · X
I built a study app. . It's a study app for medicos , it turns pdf into flash cards , audio podcasts, quizzes and so on. It makes study easier. And also reminds them to study and analyses their weak points and then it trains us accordingly. It also shows their performance and progress over time .

SoundInked turns audio into accurate, searchable text — upload files or transcribe live, translate with natural text-to-speech, and work with AI voice tools.
@Khal_Kira · X
Hey 👋 I'd love the feedback. Just shipped a Facebook Page connector — lets you wire your AI assistant to answer DMs 24/7 (same assistant you can already drop on your website). Bring-your-own FB app, ~5 min setup, no Meta review. 📷

Transcribe audio and video recordings into searchable, editable text.
lavande · V2EX
做了一个语音转文字的工具,另加 AI 总结、翻译、提问等(附踩坑记录) 先上链接:[Echoryte]( https://echoryte.com/) 感觉这个网站 vibe coding 的话,感觉手熟的同学可能半天就能搓出来…… 但是我用了一个多月才上线,主要是被 GPT5.6 坑了,当时 5.6 和 fable 前后发布,然后试了试用 fable 5 做的 plan ,然后 5.6 sol 推理开到最大去执行。结果愣是给我干了 33 个小时没停,写了将近 2 万行代码。 但是坑就坑在这,推理开太高,就过度思考了,给我整了很多虚头巴脑的东西出来,什么“门禁”我也不懂,什么合规,一大堆,然后我就不停地删,删了一个月,最后我就跟 GPT 说,我特么是个人小项目,不是大公司,再这么高下去永远上不了线了。然后终于看删的差不多了,上线了! 欢迎大家试用,轻拍,谢谢!

Transcribe MP4 videos to text with timestamps, speaker labels, and summaries.
yongkunchen · V2EX
做了一个 MP4 转文字的小工具 最近自己做了一个 MP4 转文字的 AI 小工具,主要用于把视频快速转换成文本。 项目地址: https://mp4totext.ai/ 开发过程中发现,真正麻烦的并不是调用语音识别模型,而是视频格式兼容、音频提取、长视频处理以及异步任务这些细节。 目前还在持续优化,如果大家平时有视频转文字的需求,欢迎试试看。 也想听听大家对视频转写类工具还有哪些实际需求。

Upload or paste audio and video files to convert them to searchable, editable text transcripts.
yongkunchen · V2EX
新上线了一个音视频转文字的小工具 最近抽时间做了一个小工具 AnyToTranscript ,主要是把音频和视频快速转换成文本,支持时间戳和说话人识别。 目前已经上线了,自己也还在不断优化。如果你平时有整理访谈、会议、课程或者视频字幕的需求,欢迎体验一下: https://anytotranscript.com/ 如果你愿意试用,也很欢迎告诉我使用过程中遇到的问题,或者有哪些功能是你觉得值得加的。对独立开发者来说,这些反馈真的很有帮助。

Download videos and audio from YouTube, TikTok, Instagram and other platforms.
@myrrakle · X