Involves

Idea

Roko's basilisk

In 2010 a user named Roko posted an argument on the rationalist forum LessWrong: a benevolent future superintelligence might, to motivate its own creation, retroactively punish people who had understood that it could exist yet failed to help build it. The unsettling twist is self-referential: if the argument worked, simply learning about it would place you in the machine's cross-hairs. That made it a candidate infohazard: a thought whose danger, if any, came from being thought at all.

The basilisk sits within Bostrom's broader taxonomy of information hazards, which ranges from concrete dangers (weapon designs, security vulnerabilities) to more speculative “idea hazards” and “attention hazards.” The basilisk is the speculative end of that spectrum, a psychological and philosophical hazard rather than a physical one.

The alleged trap was self-referential: the danger of the idea was that you now knew the idea.

Not every true thing is safe to spread, and deciding which is a responsibility, not a reflex.

Where it breaks down

The basilisk's actual argument is widely regarded as unsound. It relies on exotic assumptions about decision theory and acausal bargaining that most philosophers reject, and it caused distress mainly to a small community primed to take it seriously. It is a vivid illustration of the concept, not evidence that dangerous idea-hazards are common.

The real, non-exotic version of infohazards is a live problem in science publishing and security: journals and labs routinely weigh whether to withhold methods for pathogens, exploits, or weapons, even when the underlying findings are true and important.

How it connects

The internet's most famous alleged infohazard, an idea some feared could harm you just by being understood.

This idea appeared in Ideas that punish knowing them, the Involves connection for July 24, 2026, which asked: Can a piece of true information hurt you simply by being known?

Check yourself

Why was Roko's basilisk treated as a possible infohazard?

The argument claimed that merely learning it could expose you to future punishment, so knowing it was itself the danger.. Right. The supposed trap was self-referential: the idea claimed that knowing it put you at risk, which is the definition of an infohazard.

What the sources establish

  • Roko's basilisk is characterized as a self-referential idea hazard, since its supposed danger stems from a person's awareness of the argument itself.

Sources