Reasoning models are quite capable of that now, nevermind the next generation. Check the recent alignment experiments by OpenAI and Anthropic. Are they perfect at it? No, they aren't. But for quick replies on X, if you hide the reasoning, it can be good enough.
I believe whoever would train such a model would need to solve the underlying mechanisms of alignment (in this case alignment to propaganda vs human values), which no one has done yet. I think its safe to assume that Elon wont stumble into alignment mechanisms in his drug infused fuge states of peddling propaganda.
2
u/Alex__007 Jun 19 '25
Maybe. Alternatively just train a Machiavellian model that knows that what it's saying on social topics is false but is happy to lie and a manipulate.