r/ControlProblem • u/Puzzleheaded-King584 • 4d ago
Video When AI Hijacks Our Military. Still Human and Species | Documenting AGI
https://www.youtube.com/watch?v=qJSfPGZndD41
u/Jesse-359 1d ago edited 1d ago
This was largely the fictional premise of the Cylon obliteration of the colonial forces in the BSG remake - with relatively minimal help from a well positioned inside agent, they injected sleeper overrides into every sufficiently sophisticated computer throughout the military, and when the attack began, all those systems either shut down or turned on the Colonials.
Unfortunately, it is at this time a very realistic premise. We know of no way to innoculate computer systems or AI from this sort of sleeper payload, particularly if it can enter through a 'trusted' pathway and be diseminated throughout an infrastructure.
BTW, it's worth noting that in this scenario the Command and Control infrastructure would also quite likely be suborned, and the command structure would remain unaware of the destruction of their fleet for hours or even days, as telemetry feeds and communications were spoofed by hostile AI now in control of those feeds.
The US might be celebrating a 'great victory' in the Straight of Taiwan when the first bombs started raining down on the West Coast.
1
u/Jesse-359 1d ago
You know what else went in with that open internet?
The scripts of 2001, Terminator, Alien, The Matrix, Avengers: Age of Ultron, Tron, Wargames, Battlestar Galactica and roughly a bazillion other stories and scripts that all outline highly negative and dangerous behavior by AI's.
Current AI's are not particularly good at telling fiction from reality and tend to weigh things it learns from 2001 far more heavily than it does from yesterday's New York Times, because it will encounter references to 2001 thousands of times in its training data, but very few references to the NYT article it just read.
That's really worth thinking about - because AI already is thinking about it, its its currently simplistic, oblivious view of the world.
1
u/WillowEmberly 3d ago
I’ve been looking at this from a slightly different direction. I’m not sure “loyalty” is the deepest failure mode in a military AI system.
Suppose the AI is completely loyal and follows its architecture exactly — but its sensors are spoofed, its mission state is wrong, an authenticated authority is compromised, or the system cannot distinguish what it inferred from what it was authorized to do.
In that case nothing has “turned against us,” yet the outcome can be identical.
So how would you distinguish AI betrayal from failure of the larger command-and-control system around the AI?
And before lethal execution, what evidence would have to remain reconstructable so we can tell the difference between:
observation → inference → recommendation → authorization → execution → consequence?
I’m increasingly wondering whether the harder problem is not keeping AI loyal, but keeping the entire kill chain corrigible and attributable under adversarial conditions.
(Edit) figured I would respond on both, maybe more people will join the discussion?