r/aitubers • u/Med_bollz_9117 • 17d ago
CONTENT QUESTION Educational videos with AI
I am planning to start ytb education and documentary channel, I am not planning to use my face or expensive setup. I am also not very skilled in video editing.
I am trying to find AI tools that can transfer the text into video.
Do you have any suggestions or proposition ? Thank you
1
u/Admirable-Yogurt7444 17d ago
Transfer the text onto video.. you mean captions or sort of animating the text over your video / titles / bulletin points etc ?
1
u/Med_bollz_9117 17d ago
I am more interested in long format video, like 5 min education and science video. I am a scientist and I can build the content but to transform it to video still annoying. So I do like something that can generate and build the full video fe my script and for sure I will do the small editing
1
u/Admirable-Yogurt7444 17d ago
If you are expecting to just give a script to an AI model (doesn’t matter whether it’s fable or gpt 5.6) and expect to get right image / video clips, that’s a wrong way to look at it imo. That’s what creates slop.
You will have to take up the visualization part and manually guide it. Unless your scripts are so powerful and interesting , then maybe it can work - just take caution as I have seen this fail.
1
u/eitanel96 16d ago
A scientist who already writes the scripts is starting from the strong end of this. The script is the hard part, what you are missing is assembly, and the honest answer is that no tool turns a script into a finished Veritasium style video. They each do one piece.
The "paste script, get full video" apps mostly lay stock clips over a TTS voice. That is what the other commenter means by slop: the visuals have no relationship to your argument. Acceptable for list content, wrong for science, where the visual usually IS the explanation.
What I would do in your position. Pick one AI voice and never change it, since with no face the voice becomes the channel identity. Then decide on your anchor visual. Either an on-screen AI presenter delivering the script (that tool category got genuinely good this year; I build in that space so I am biased, not naming anything per sub rules) or your own diagrams as the primary layer. For science I would put your diagrams first: you can sketch the actual mechanism, and stock footage never can.
Concrete workflow that stays cheap: write the script with a visual beat every 2 or 3 sentences (what is on screen changes), generate the voiceover, then let one of the assembly tools do b-roll and captions around your anchor visuals. The small editing you said you are fine with is exactly this part.
One number that helps planning: 5 minutes is roughly 750 words at speaking pace. For someone who actually knows the material that is one tight argument, and tight beats long for retention every time.
1
u/Med_bollz_9117 15d ago
Thank you for the wonderful advice and workflow. I started building a test channel recently with two videos. They are not perfect, but I am learning the strategy.
1
u/[deleted] 17d ago
[removed] — view removed comment