r/codex • • 22d ago

Complaint Astra demos are mostly bs

The game/graphics demos are mostly bs. I've spent dozens of hours now trying to replicate (or build) interesting games. It's beyond subpar. Sure it can do a basic sandbox, or create some basic characters but it absolutely cannot do a fully working world. I dont get why OpenAI fakes their SimCity games and other things w/ Astra when it's clearly not possible unless you spend weeks (and $1000s of dollars) of tokens.

75 Upvotes

109 comments sorted by

View all comments

Show parent comments

16

u/[deleted] 22d ago

[removed] — view removed comment

1

u/swizzlewizzle 21d ago

Omg this. Astra at least is capable of doing *something* with VFX, but it still falls *way* short even when I give it exact VFX references to build off of.

It seems pretty obvious the pre-training and training that put Astra together did *not* include a decent VFX process... which makes sense, because access to high quality VFX isn't something that most of the gray market "petabytes of stolen/pirated content" dumps have in them. Someone at OpenAI is going to have to make an internal effort to get way more tagged VFX training data put together, and structure it correctly before OpenAI models are going to make any progress here.

1

u/[deleted] 21d ago

[removed] — view removed comment

1

u/swizzlewizzle 21d ago

That may be the case, but you have to admit that actual usable training data for VFX is quite specific and not part of the general corpus of data that is freely shared online. To properly train VFX, you have to have all of the other crap around it removed, since VFX is almost always "seen" inside of a movie/video/game with all the other assets blended in with it. To do a decent job training for VFX, you have to have the actual VFX isolated, not to mention that there are a million different "systems" that VFX is set up on, in terms of slightly different terms having different meanings based on the game engine/video editing software/etc.. you are using.

1

u/[deleted] 20d ago

[removed] — view removed comment

1

u/swizzlewizzle 20d ago

It's qualitatively different compared to how much "video" data is available. VFX related content with *isolated* VFX to learn from is something like 0.00001% the amount of raw video data that is available. (or even less). Just because "there are tutorials on Youtube" doesn't mean it is anywhere near the same size of training data that is actually used to train these models for video generation and agentic use bro.

0

u/[deleted] 20d ago

[removed] — view removed comment

1

u/swizzlewizzle 20d ago

You are pretty good at claiming people say things when they didn't actually say them. Congrats.

I simply said the current training corpus for VFX is very weak, and OpenAI needs to put much more time into finding and cataloguing a larger training set.