227
Why isn't everyone talking about AI generated audiobooks?
(reddthat.com)
A loosely moderated place to ask open-ended questions
Search asklemmy ๐
If your post meets the following criteria, it's welcome here!
Looking for support?
Looking for a community?
~Icon~ ~by~ ~@Double_A@discuss.tchncs.de~
I recommend everyone to check the YouTube channel "two minute papers" who have being doing videos about papers on AI for the last 10 years on so to see the accelerated progress AI have. Like 5 years ago those images generating AI looked like LSD infused dreams and now they look almost perfect.
I'm only shocked that video isn't better. Diffusion models work like denoising - so you'd figure all the wiggly nonsense between frames would be the first thing to filter out.
I expect the data size to be a problem. Stable diffusion defaults to 512x512px, because it simply requires a lot of resources to generate an image. Even more so to train one. Now do that times 30 to generate even one second of video. I think we need something that scales better.
I fully expect this to work decently in a few years though, no matter how hard the challenge is, ai is moving really fast.
Stable diffusion can do arbitrary sizes now, as long as you have the VRAM for it iirc
Of course, but that is precisely the problem. It gets expensive really really fast.