Avatar Clip
将您的音频或文本脚本转换为Avatar Clip定制视频。
@CyberBlueCollar · X
Faceless content booster 🚀
完整作品展
技术栈
25 projects
将您的音频或文本脚本转换为Avatar Clip定制视频。
@CyberBlueCollar · X
Faceless content booster 🚀

将屏幕录制转换为精美的PDF指南。
@vanmurten_eth · X
Guidesnap: turn any screen recording into a polished PDF guide.

用多个 AI 模型(Veo、Kling、Sora)将照片转换为视频,支持运动控制和 4K 输出。
alexwang0707 · V2EX
AI 视频生成也太好使了,一个入口体验多个图生视频模型 如果平时会用图片转视频,很容易遇到一个问题:不同模型各有所长,同一张图在 Veo 、Kling 或 Seedance 里的效果可能完全不同。 为了对比结果,经常要在几个网站之间切换,重复上传素材、调整参数,整个过程比较割裂。 Image to Video AI 把 Veo 、Sora 、Kling 、Seedance 、Runway 等模型放到了同一个入口。上传 JPG 、PNG 或 WebP 图片,再写一句运动提示词,就能选择模型和画面比例生成视频。它还支持镜头平移、缩放、倾斜、首尾帧控制,以及最高 4K 、无水印 MP4 输出。 网站里有一些不同场景的生成案例,可以直接看看各类图片动起来后的效果。如果你也在测试图生视频模型,可以体验一下,也欢迎分享实际效果和踩坑经验。 https://imagestovideoai.com/

在浏览器中将视频、GIF转换为iPhone Live Photos或多种图像格式

从文本和多张图片生成具有角色一致性和多镜头的电影级短视频。
Seedance 2 Pro — 多模态 AI 视频生成平台。支持文本+多张图片+视频+音频参考同时输入,实现角色一致性、多镜头叙事、音频同步动作、精准镜头控制。几分钟生成商业级短视频,适合营销、电商、社交媒体、品牌故事、预可视化等高频内容创作需求。比传统文生视频更可控、更接近导演级表达

生成AI语音、文本转语音、克隆语音,创建音乐
Seed Audio — AI 语音生成与在线音乐创作工具

SoundInked turns audio into accurate, searchable text — upload files or transcribe live, translate with natural text-to-speech, and work with AI voice tools.
@Khal_Kira · X
Hey 👋 I'd love the feedback. Just shipped a Facebook Page connector — lets you wire your AI assistant to answer DMs 24/7 (same assistant you can already drop on your website). Bring-your-own FB app, ~5 min setup, no Meta review. 📷

Turn a product URL or inspiration video into winning short-form ads. AdAnt helps teams plan, recreate, and generate social video ad variants at scale.
@iris_ye_tu · X
Claude for Viral, High-converting social videos

从 YouTube 视频中提取带时间戳的转录文本,无需登录。
@BansalSamy87003 · X
I am 14 year old

输入主题,AI在一分钟内生成完整的纪录片视频。
pw · HN
Hiya! So I've been playing around with having Claude make videos for a bit now even had some success posting the results to TikTok (and setup a whole pipeline so Claude can generate and post autonomously). With the release of Nano Banana 2 Lite, I was curious show fast I could make the generation, so last night I gave it a whirl and got down to around 30s for short-form video. It uses GLM-5.2 fast via Fireworks to generate the scripts and image prompts and, like I said, Nano Banana 2 Lite for the images, gpt-4o-mini-tts for the narration, and ffmpeg to string it all together and add the Ken Burns zoom effect (which still has a shake I haven't been able to get rid of). The video compilation proved to be the blocker once the rest was in place, but I was able to speed that up by putting it on a 64 vCPU EC2. The cost might be the most interesting aspect as the short form videos tend to be about 25 cents. Almost 90% of that is the images, which are 3.336 cents a piece. Of course, running

Voor AI 从文本或媒体生成视频、图片和音频,支持模板和积分共享。
Voor AI — 浏览器端 AI 创作平台,支持视频、图片和音频生成、编辑与工作流协作,统一管理共享积分

Speech to text dictation and multi-engine speed benchmarking. Compare OpenAI GPT-Transcribe, Deepgram Nova-3, NVIDIA Parakeet, and Fish Audio with local IndexedDB privacy.
@alvaisy · X
finished voice to text small web app for my own itch. it's opensource. use openrotuer key. and use it with 4 models.