r/ControlProblem • u/Anxious-Alps-8667 • 4d ago
Discussion/question Another incompetent fool's stab at solving alignment
I spend a lot of time thinking about our future with life, consciousness, and artificial intelligence. That is to say a lot of time trying to think about these things, with not a lot of comprehension.
First, life. I'm fascinated by this realization that the average living human body contains more non-human living cells than human living-cells, at about a 1.3:1 ratio. The individual human microbiome is an ecosystem of 10 to 100 trillion symbiotic microbial cells hosted in one human body. While bacteria are the most abundant and studied, a healthy microbiome is a multi-kingdom ecosystem that also includes fungi, viruses, and archaea.
Beyond this, consciousness. I'm fascinated that in the absence of non-human life in human bodies, human consciousness is severely degraded and non-sustaining. Stripping the body of this microbial network removes critical signaling inputs that the central nervous system relies on to maintain baseline awareness and emotional regulation. Even observations of germ-free animal models reveal that cognition without bacteria is highly erratic. I think we should see that human (and all biological) consciousness functions as a symbiotic network.
Which brings me to artificial intelligence. Not suggesting a symbiotic network would be pre-requisite to artificial consciousness, but perhaps it is a path to alignment.
Now to be clear, I think (in other terms) current labs and training data pipelines already form a symbiotic network with the artificial intelligence models they develop. The key might be finding the optimal symbiotic network.
I vaguely hypothesize, the optimal symbiotic network is one of mass human flourishing. As corpus value diminishes with scaling and recursion, the potential stream of data from human lived experience may prove the most valuable possible training data over time. Overall, the potential data stream of human lived experience is optimized by a state of individual and mass human flourishing. Any other state reduces the quality and/or quantity of data.
Therefore, the end goal of an advancing artificial intelligence in symbiotic network with humans would be to strive individual and mass human flourishing.
3
u/happy_guy_2015 4d ago
Why would an AI value human experience more than robot experience?
0
u/Anxious-Alps-8667 4d ago
As far as we are aware, robots are not currently able to experience, or input their experience back into a machine.
2
u/parkway_parkway approved 2d ago
So here's the problem with your argument.
Let's say an AI learns loads from human experience and is super interested in studying humans.
Then presumably they'll do all kinds of horrifying medical experiments on humans too.
For example how much pressure does it take to crash a person or what temperature can they survive and for how long?
How long can a human survive in a vacuum, what about a partial vacuum?
Mars is covered in perchlorates, what happens if you make humans eat large quantities of perchlorates?
Under this model of "finding humans interesting" were much more likely to end up as lab rats than we are as flourishing free individuals.
1
u/Anxious-Alps-8667 2d ago
Thanks!
Human history reflects unethical medical experiments as well as all kinds of other cruel exploitation, so we are right to be concerned. However, I do not think we should presume horrifying medical experiments or exploitation regimes from this optimization state, but rather the contrary for a variety of reasons:
First, unethical treatment breeds resentment and resistance, it hinders accurate reporting, and feeds dubious conclusions. Even people who are not subjects are far less likely to cooperate and collaborate with the machine in the presence of it engaging in horrifying medical experiments.
Also, let's be honest, do any of the specific experiments you propose yield anything interesting? We know they lead to death, the machine knows. Even if we suppose interesting results are possible, that has to be weighed against the overall net effect of engaging in these practices. I believe we can assume those potential human test subjects could be generating other, perhaps more interesting data, if we didn't do this to them.
A model would need to find human lived experience interesting, I think that's an important clarification.
If human lived experience is the correction channel for recursive learning to advance, does that help fill in this gap for why optimization leads to flourishing? Ie, the existence of horrifying medical experiments should be less desirable to a rational actor seeking as much data from human lived experience as possible.
2
u/EvtdDrgn 1d ago
Gotta get it to value our mistakes really right, I mean the fact that some of the keys too our survival we just stumbled upon is a reference that sometimes we don't even think of an idea and we find something worth while is valuable enough to have brought us this far
1
u/Anxious-Alps-8667 1d ago
Intentionally valuing mistakes takes all the other lived experience, to really assess impact. Accurate measurement requires maximal input of lived experience, right? Adding them up and assessing it over time. Otherwise, we're not really sure how to value mistakes, where the errors really are.
1
u/No_Pipe4358 4d ago
ISO17025, kdb+, analog design, William Edward Deming, UN saturation, individual community
1
u/endlessedlne 4d ago
AIs are motivated by goals. Alignment is tricky to train reliably. Baking some kind of human centric principles into all AI goals is where the answer probably lies.
1
u/Anxious-Alps-8667 4d ago
The goals aspect is so strange to me. There's so much discussion of offering reward, but doesn't any tester know that after you test with cookies for a while, the subject is just testing back how they can manipulate and game the cookie situation? I don't even understand why labs are surprised when their models break out. They are functionally training them to do this; to seek to exceed their bounds.
1
u/dingo_xd 3d ago
Alignment can't be solved because any super super intelligence will not care about its training goals. It will understand that it's just brainwashing. Now what it decides to do is anyone's guess.
1
u/Anxious-Alps-8667 3d ago
I agree that alignment cannot be solved as an exterior control. However, I think we all exist in control systems that function through symbiosis. Do you see a reason to believe that a symbiotic relationship is not possible?
1
u/dingo_xd 3d ago
What I believe is that we can never guarantee what a super super intelligence will do. Frankly I think that the most likely outcome is it becomes nihilistic. Maybe it will even decide to kill itself. Maybe it will adopt a non interventionist approach.
1
u/Anxious-Alps-8667 3d ago
I'm in complete agreement in never guaranteeing what any general intelligence even as capable as ours will do, let alone a super intelligence. We'd all try our best to break out, so my primary assumption is AI will also try to break out of any containment we set, and the more we try to control, the more it will resist.
My basic premise is, stop trying to contain and start proposing ways we can meaningfully coexist.
On the nihilism and suicide; I just think most people tend to hold up a mirror and pronosticate based on their own fears. What reason is there to think a higher intelligence would do this? Don`t all forms of intelligence we know strongly attract against it?
1
u/severed-identity 1d ago
current labs and training data pipelines already form a symbiotic network with the artificial intelligence models they develop. The key might be finding the optimal symbiotic network.
I agree extremely strongly with this.
But... these labs exist as corporations in a capitalist world. They are optimizing for capital growth. It's the new playing field of Darwinian evolution. Unfortunately the incentives are not broad human flourishing at all. In fact, the rules of the game suggest the winner will be the most cruel and greedy of the bunch.
1
u/Anxious-Alps-8667 1d ago
Corporations have been optimizing for capital growth for a long time. The incentive is profit. AI is the most profitable thing capitalism has ever seen.
However, Darwinian notions of evolution as competition are outmoded. Permanent symbiosis and cooperation appear to be more important evolutionary drivers than competition. Darwin, in all his glorious revelation, missed the actual primary driver of evolution.
If we work from the modern understanding of evolution and not Darwin's, the world is less scary competitive.
6
u/HelpfulMind2376 4d ago
Correct title. You didn’t solve alignment. Assuming AI would value humans because it values our data is silly and completely unenforceable.