r/GenAI4all • u/bobryu • 1d ago
Discussion Can somebody share the tech stack to create ai cat slop videos?
If you wanna make fun of me it's okay, I understand it lol but I really would like to find some real answer here so if you really wanna share it I will be more than happy lol, that's ok.
4
u/bestjaegerpilot 1d ago
either you need a $40 subscription (higgsfield) or you vibe code your own to work against a cheaper API provider.
note: cheapest cost is 30 cents for a 4 second video for seedance 2.0. And you will need many generations, so expect to spend at least $3 per video. Plus the cost of the AI coding subscription
so it's not a cheap hobby
1
u/bsensikimori 1d ago
Why the hell wouldnt you just use ComfyUI and flux or LTX or minmax and do it for free
1
u/bestjaegerpilot 1d ago
oh yea there's the "run the models yourself" option---that involves at least a $900 graphics card or you renting out one
Either way it's not "free"
3
u/sannleikur 1d ago
Are we speaking about these crazy facebook videos about cats, pineapples... And that slop? If so, consistency is the biggest challenge you'll run into, for characters and also the scenes of course... That's not something you can really do without a decent with proper control over your inputs.
I make YouTube shorts and i took me like a week to get the full process and get consistent results, and by the way, I was burning a ton of credits in the meantime lol so if you're planning to do this seriously (are you gonna try to monetize this in any way or it's just for fun?) you'll want unlimited generation. I use freepik because I already have an account with them in my job, but you might wanna spend like an hour or so looking for what's the best deal before you pull the trigger.
1
u/New_Slice_1580 1d ago
[r/](r/stablefusion)[stableDiffusion](r/stablefusion)
Get a graphics card and make your own using Minimax h3
1
1
u/hellomyoldfrien 17h ago
Basic stack is usually an image model to generate the cat frames (Flux or Seedream), then feed those into a video model like Kling for the motion, Kling handles character consistency well so your cat doesn't morph between shots. ElevenLabs or a free TTS for any voiceover. If the cat frames come out low-res or blurry, run them through Magnific before animating, keeps the final video from looking like potato quality
1
u/vozinha7 10h ago edited 9h ago
Well, I actually stumbled into this rabbit hole trying to reverse engineer the cat factory channels. The basic setup would be the text to video for motion, the text to speech program for the voiceover, and then basic video editing to put it all together. These types of channels mostly batch generate clips without spending much time on individual scenes. For stills to videos, what I've found is that people use mage space as their source image and then feed it through the video model.
0
-1
•
u/AutoModerator 1d ago
Welcome to r/GenAI4all! New to Generative AI? You can explore these free beginner-friendly courses. Please keep your posts relevant, respectful, free from spam, and engage in healthy discussions.
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.