
Veo3.1 视频生成器|原生音频与镜头级控制
Create 1080p videos with native audio and camera-level control.
Veo 3.1 AI — 使用 Veo3.1 创作电影级视频,配备 1080p 原生音频、首尾帧控制及免费试用额度
The full gallery
Tech stack
28 projects

Create 1080p videos with native audio and camera-level control.
Veo 3.1 AI — 使用 Veo3.1 创作电影级视频,配备 1080p 原生音频、首尾帧控制及免费试用额度

Generate 2K AI videos (5-15 sec) from text, images, and audio references with native stereo audio.

Generate 768p AI videos from text or images with native audio in seconds.
plutozc · V2EX
做了两个 MiniMax H3 AI 视频生成入口 最近在用 MiniMax H3 做一些 AI 视频的快速试验,发现自己常在文字、图片参考、画幅和时长之间来回切,所以顺手做了两个偏不同工作流的入口,想听听大家会不会觉得这种拆分有用。 一个是偏快速迭代的版本: [H3 Max]( https://h3-max.org) 主打从文字或参考图生成短视频,重点放在快速生成 768p 片段和提示词里的画面、镜头、音频描述。 另一个是功能更完整的工作台: [MiniMax H3 Max]( https://minimaxh3max.org) 同样支持 text-to-video 和 image-to-video ,可以按需要选 768p 、2K 或 4K ,也能调时长和横竖画幅。 目前都还是刚上线的个人项目: - 文字或图片作为起点生成视频 - 一个提示词里写画面、动作、镜头和声音方向 - 方便先快速出一个版本,再回头改提示词 两站都不是官方站,生成结果和速度会受模型设置及队列影响。现在页面和体验还在继续补,尤其想听听大家更在意“快点出第一版”,

Turn text and images into cinematic 4K+ videos with native audio, multi-shot sequencing, and persistent character identity.
kling — AI 生成视频,支持多种 kling 模型


Convert documents and e-books into audiobooks with natural-sounding narration.
@coder_zi · X
Clipifai converts e-books to audiobooks say you have a few books in mind to read, but haven't really had the time to read them, you could just convert them into audiobooks and listen to them on the go...

Segment Anything for Audio — made simple. Separate audio with a text prompt. Isolate vocals, remove background music, extract speech from noise. Privacy-first stem splitter powered
@sambamdamnn · X

Convert text and PDFs to audio with natural-sounding AI voices.
@caschiblu · X

A tactile Philips-style analog cassette recorder with modern music discovery player.
@Z0D404 · X
I built this trendy site inspired by @ybhrdwj , @s4tr2, and others; all credit goes to them. This is pretty much vibe-coded. I have also added my region bus uncle's playlist😎; check it out. Website:

Generate AI voices, sound effects, and music from text prompts in multiple languages.

Convert MP3 and MP4 files to editable transcripts with timestamps and speaker identification.
digiplanp · V2EX
做了一个音视频转文字工具,支持 MP3/MP4、说话人识别和字幕导出 大家好,最近做了一个在线音视频转文字工具 ToText ,想发出来请大家体验一下,也听听 V 友的意见。 最开始做这个工具,是因为我自己在处理会议录音、访谈和视频内容时,发现单纯生成一大段文字并不太实用。转录完成后,通常还需要反复对照原始录音、修改错误、区分说话人,再整理摘要或字幕。 所以现在做成了一个相对完整的工作流: * 上传 MP3 、MP4 、M4A 、MOV 、WAV 、WebM 等音视频文件 * 自动识别语言并生成带时间戳的文字稿 * 支持说话人识别和说话人名称修改 * 点击文字片段可以跳转到对应的音视频位置 * 可以搜索、编辑和校对转录内容 * 自动生成摘要、会议纪要、章节和行动事项 * 支持导出 TXT 、DOCX 、PDF 、SRT 、VTT 、CSV 和 JSON 音频文件可以直接使用 [MP3 to Transcript]( https://totext.app/) 功能,比较适合播客、访谈、课程录音、会议和语音备忘录。 视频文件也使用同一个编辑器,通过

An intellectual companion for the digital aristocracy.
@scooperleash · X