
The 30 Papers — Ilya Sutskever's Deep Learning Reading List (for John Carmack) |
Listen to Ilya Sutskever's 30 AI papers explained as free audio episodes.
janpmz · HN
Ilya's 30 papers, explained in audio
The full gallery
Tech stack
60 projects

Listen to Ilya Sutskever's 30 AI papers explained as free audio episodes.
janpmz · HN
Ilya's 30 papers, explained in audio

Upload or paste audio and video files to convert them to searchable, editable text transcripts.
yongkunchen · V2EX
新上线了一个音视频转文字的小工具 最近抽时间做了一个小工具 AnyToTranscript ,主要是把音频和视频快速转换成文本,支持时间戳和说话人识别。 目前已经上线了,自己也还在不断优化。如果你平时有整理访谈、会议、课程或者视频字幕的需求,欢迎体验一下: https://anytotranscript.com/ 如果你愿意试用,也很欢迎告诉我使用过程中遇到的问题,或者有哪些功能是你觉得值得加的。对独立开发者来说,这些反馈真的很有帮助。

Create text, images, video, and audio with AI chat and 100+ tools in one place.
Fullmira — 一站式 AI 创作平台,整合 GPT、Claude、Gemini 等多个模型,同一界面完成文本、图像、视频和音频生成,内置 100+ 专用工具覆盖写作、图像编辑与视频制作

Convert documents and text into podcast audio with AI voice generation and editing.
Inpodcast AI — 将文档转成播客音频,支持 PDF、Word、Markdown 和 TXT 文件格式

Convert and edit audio, images, PDFs, and videos with 150+ free online tools, no signup required.
@factfile0 · X

Daily news aggregator that watches print, audio, video, and digital sources for your interests.
@BairdHall · X
“It’s over” “We’re done” etc. I wish @x algo was better and not surfacing that type of stuff. Side note: been vibe coding an app to try and fix this:

Convert voice inputs into structured data using front-end AI.
pankajunk · HN
Talkform.org, Turn voice inputs into structured data using front end AI

Generate 768p AI videos from text or images with native audio in seconds.
plutozc · V2EX
做了两个 MiniMax H3 AI 视频生成入口 最近在用 MiniMax H3 做一些 AI 视频的快速试验,发现自己常在文字、图片参考、画幅和时长之间来回切,所以顺手做了两个偏不同工作流的入口,想听听大家会不会觉得这种拆分有用。 一个是偏快速迭代的版本: [H3 Max]( https://h3-max.org) 主打从文字或参考图生成短视频,重点放在快速生成 768p 片段和提示词里的画面、镜头、音频描述。 另一个是功能更完整的工作台: [MiniMax H3 Max]( https://minimaxh3max.org) 同样支持 text-to-video 和 image-to-video ,可以按需要选 768p 、2K 或 4K ,也能调时长和横竖画幅。 目前都还是刚上线的个人项目: - 文字或图片作为起点生成视频 - 一个提示词里写画面、动作、镜头和声音方向 - 方便先快速出一个版本,再回头改提示词 两站都不是官方站,生成结果和速度会受模型设置及队列影响。现在页面和体验还在继续补,尤其想听听大家更在意“快点出第一版”,

Transcribe audio or video, then summarize, chat with, and export your transcripts.
u/Striking-Lychee-8958 · Reddit
Went from $6 to $65 MRR this month lol. Not much but here's what actually moved it https://preview.redd.it/xbl930ysadjh1.png?width=1623&format=png&auto=webp&s=83eef74c5324de3a119d1c075e82e0d27f913fc1 Ok so this isn't some huge win, went from 2 to 3 paying users, MRR went from $6 to $65. I know, small numbers. But something actually shifted for the first time in months so I wanted to write it down. Transcriptor Pro is an AI transcription tool, mainly for podcasters and crea

Create videos from text or images with AI motion control, audio generation, and lip-sync.
Veo3 AI Video Generator — 支持文本转视频和图像转视频,配备原生音频生成和精准唇形同步,2-5秒完成专业级视频创作。

Copy curated prompts for AI image, video, and audio generation.
weijunext · HN
Free, curated prompts for AI image, video and music

Turn text and images into cinematic 4K+ videos with native audio, multi-shot sequencing, and persistent character identity.
kling — AI 生成视频,支持多种 kling 模型