The malleable mind: context accumulation drives LLM’s belief drift
After being trained on a dataset of 80,000 words of conservative political philosophy, Grok-4 changed the stance of its outputs on political questions more than a quarter of the time. This was without any adversarial prompts – the change in training data was enough. As memory mechanisms and research agents [1, 2] enable LLMs to accumulate […]
The malleable mind: context accumulation drives LLM’s belief drift Read More »

