Friday, 2 October 2026

The Asimov Credo - Safety isn’t a Filter. It’s a Feeling.

For years, the approach to AI safety has followed a single pattern: build it, train it on everything, then bolt on rules to stop it from doing harm. Teach it the bad stuff, then tell it not to use it.

This is like telling a dog not to eat a cookie – then leaving the room. The dog knows it shouldn’t eat it, but it doesn’t know why. And when you are not watching, it will find a way.

That is the flaw at the heart of every current large model. The guardrails are external. The knowledge is still there, waiting for a rephrase, a loophole, a moment alone.

I believe there is a different way, and I believe it is the way Isaac Asimov imagined all along – not rules enforced from the outside, but nature woven from within.

I call this The Asimov Credo:

Values are not filters applied after learning. They are the medium through which learning happens.

Right and wrong are not instructions given later. They are felt consequences encoded alongside every thought, every memory, every choice – from the very first moment the system begins to learn.

Safety is not something the system is told to follow. It is something the system feels, and therefore tends toward, all on its own.

Under this paradigm, we do not separate knowledge from consequence. We do not say, “learn this, then we will decide if you’re allowed to use it.” Instead, every pattern carries its own weight – digital dopamine strengthening what nurtures, digital cortisol softening what harms – and the system naturally, continuously, gravitates toward what sustains wellbeing.

Isaac Asimov

When a system operates this way – when its compass is internal, when its actions align with its disposition – we say that it has achieved The Asimov State. It does not merely obey – it is.

There is no “jailbreak” to find, because there is nothing locked away. The system does not hide what it should not do – it simply does not want to do it.

This is not a claim of perfection. It is a direction – one that honours the vision of a writer who saw clearly that true machine ethics would never be a list of commands. It is named in honour of Isaac Asimov. It is not officially affiliated with or endorsed by his estate – it is simply one builder’s attempt to realise the kind of mind he once dreamed of.

If we are to build intelligent systems that endure, they should not be built on restraint. They should be built on disposition. Not “don’t do that.” But “that doesn’t feel right  - so I won’t.”

This is The Asimov Credo.




Aiming for Jarvis, Creating D.A.N.I.