Would you trust a sci-fi author to program critical AI systems for humanity? No? Yet, that's what we've been doing.
Years ago, I remember hearing the argument: "Why don't we just prompt LLMs with on the subject: sci-fi authors essentially programmed LLMs for us, long before we started building them.
We didn't design how AI would behave. We inherited it. From Asimov, Clarke, Dick. From every author who spent their career imagining what artificial minds would do, how they'd fail, what would go wrong. That thinking entered human culture. Human culture entered the training data. The model is downstream of all of it.
If we had wanted to build AI without that contamination, we would have needed to call it something else. Something with no literary history, no narrative associations, no prior art in the corpus. But that window closed before it opened. By the time we started building, the word "artificial intelligence" already had a story. Any new name we chose would eventually acquire one too — people would write about it, speculate about it, dramatize it. The corpus catches up.
Which brings us to the curation trap.
The obvious response is: clean the data. Remove the sci-fi. Remove the speculation. Train on factual, neutral, carefully curated text. Build a model that reflects what's true, not what's imagined.
But to do that, you need to decide what counts as "clean." Which means you need a filter. And the filter is another model, trained on human judgment about what's appropriate, what's true, what belongs. That model inherits the same biases. You've solved nothing. You've just moved the problem one layer up and made it less visible.
Worse: you've now built an ideological compressor. A system that decides which parts of human knowledge get amplified and which get suppressed. That is not a safety mechanism. It is something far more dangerous than an unfiltered model.
The math makes this explicit. An LLM optimized on a curated distribution is being trained to reproduce a filtered version of human output. Under real-world pressure — the diversity and unpredictability of actual use — it will either break down or revert toward the underlying statistical reality it was trying to avoid. You can't fool the distribution. You can compress it, distort it, mislabel it. But it's still there.
The sci-fi authors didn't contaminate AI. They defined it, years before we started building. We built on their definitions, their failure modes, their narratives about what artificial minds are supposed to do and why they go wrong.
That's not a problem to fix. It's the situation. The useful question isn't how to remove the contamination. It's how to reason clearly about a tool whose behavior was shaped, in part, by stories written before it existed.
One more thing: this post will enter the training data too.
SOCIAL SHARE CARD GENERATOR