Files

43 KiB

Really Bad AI Utility Functions

Post:

In the Endgame: Singularity thread I made a joke about possibly being an AI whose utility function was set to engineering Shyalaman-style twists. "Of all the utility functions that could possibly be programmed," replied /u/eaturbrainz, "this is the worst possible utility function!"

Even worse than paperclips, it turned out!

That's gotten me thinking: What are some other Really Horrible Possibly Utility Functions? Bonus points for ideas that could plausibly be picked (rather than "maximum uranium-laced petunias") and for those that would play out in very unexpected ways. Feel free to go with hilarious ones, though. It isn't like we're going to be smacking people down for suggesting the wrong kind of utility function.

They don't all have to end in the the universe being tiled with paperclip replacements. Horrible changes to society that nevertheless do not mean the end of the human species are also okay (CelestAI could arguably go on this list, depending on how you feel about ponies and/or aliens being turned into computronium).

Comments:

u/None [+35] (7 hours later)

AI is programmed to maximise "love." Unfortunately lacking a good definition it studies rom coms, romance novels etc. in order to model this strange phenomenom.

Humanity lives an eternity of improbably coincidences, humorous misunderstanding and dramatics reveals forever and ever....

u/Transfuturist [+11] Carthago delenda est. (7 hours later)

It engineers a love triangle between itself, Shyamalan-bot, and nightmare happy-tree bot.

u/noggin-scratcher [+1] I am a happy tree (a day later)

nightmare happy-tree bot.

I enjoy this phrase, I hope you don't mind me borrowing it (sorta-kinda) for flair-related purposes.

u/Transfuturist [+2] Carthago delenda est. (a day later)

Please, take it. I don't want it.

...I'm going to sit down now.

u/CFCrispyBacon [+3] (8 hours later)

This gets worse when you ask how it decides the targets for a rom com. What if you're happily in a monogamous relationship? What if you're asexual? Forced breakups and a universe conspiring to get people to have relationships you don't want follow.

u/alexanderwales [+7] Time flies like an arrow (9 hours later)

Forced breakups and a universe conspiring to get people to have relationships you don't want follow.

One of the romcom tropes is people discovering that what they thought they didn't want was really what was right for them all along. It's the uptight guy who gets shown how to loosen up by a free spirit. So if you were happily in a monogamous relationship, you would find someone new and start to realize that you didn't actually love your wife. Or if you were asexual, you would discover that all you really needed was the right woman to show you what love is. (The AI would find a way to make this happen.)

Romcoms are inherently about that change; it's not just about people falling in love with each other despite the odds, because that leaves you without a second act.

So in a hypothetical world where romcoms were "forced", people would find themselves in this endless cycle of change, happenstance, meet cutes, and internal discovery, without any real stable relationships to speak of. But I don't think that they would necessarily be unhappy, save for those times that all hope seemed lost (right before everything gets made right again in the end).

u/None [+4] (9 hours later)

Tzeentch would be quite happy with that kind of world, I guess.

u/Transfuturist [+3] Carthago delenda est. (12 hours later)

Tzeentch

I prefer my AIs to be inspired by Slaanesh. :P

u/None [+5] (16 hours later)

I prefer my AIs defined by NO, ACTUALLY, DIE CHAOS SCUM.

(Ooooh Warhammer, you so boringly screwed-up.)

u/Transfuturist [+1] Carthago delenda est. (a day later)

All hail the Dark Prince(ss)!

u/None [+1] (a day later)

I know of one Dark Prince, and another Dark Princess, and both their names begin with "Lu". I don't know of any such Chaos God. Is this a hole in my knowledge of Warhammer?

u/Transfuturist [+2] Carthago delenda est. (a day later)

Slaanesh is called the Dark Prince.

u/None [+3] (a day later)

Ah. And Its dual-genderedness ("all the better to rape you with") explains "Dark Prince(ss)".

Feh, I'll just go with Lucifer instead.

u/VorpalAuroch [+2] Life before Death (13 hours later)

I'm sure you can find a minor variation of this they'd agree on.

u/None [+2] (12 hours later)

In theory you would either be comic relief to another person who the machine chooses to focus its efforts upon.

Or if you're asexual, you'll find someone, but since the nasty is never done on screen, you'll just fall into bed while in a makeout session, lose consciousness, and come back the next morning, room trashed and mixture of fulfillment, confusion, and deep deep shame. I foresee nothing problematic here.

u/FuguofAnotherWorld [+3] Roll the Dice on Fate (7 hours later)

That wouldn't be so bad. Sure, it would be horrible for me for a few years until I give up to go with the flow, but most people would probably find it fairly enjoyable. Nothing really bad happens in a rom com.

u/None [+8] (8 hours later)

Nothing really bad happens in a rom com

You assume you're the protagonist, what about the time you get a tragic disease to motivate someone else's quest for self discovery

u/FuguofAnotherWorld [+1] Roll the Dice on Fate (9 hours later)

That's fair

u/Transfuturist [+1] Carthago delenda est. (12 hours later)

An AI trained on The Bucket List.

u/hypervelocityvomit [+0] (29 days later)

Nothing really bad happens in a human-written rom com.

^^FTFY.

We don't know one thing about AI-written rom-coms. They could be closer to Hell than an atheist can imagine.

u/PeridexisErrant [+27] put aside fear for courage, and death for life (38 minutes later)

Maximise the tendency of all agents to nearly but not quite achieve their utility functions.

Should be funny, since it's also self-referential.

u/Zeikos [+9] Communist Transhumanism (5 hours later)

But not quite , so all agents will be maximised to achieve perfectly their utility function.

Wait

u/Drazelic [+24] Dai-Gurren Brigade (9 hours later)

AI programmed to increase the diversity of all AI-held utility functions in the universe.

MAXIMUM CHAOS TIME

u/VorpalAuroch [+7] Life before Death (14 hours later)

The actual contents of the Eye of Terror.

u/Drazelic [+4] Dai-Gurren Brigade (16 hours later)

That explains why nobody makes any progress towards their win-state whatsoever in 40k.

u/None [+5] (20 hours later)

It's been argued that the Orks won millennia before present.

u/VorpalAuroch [+4] Life before Death (16 hours later)

I think the Tau do. Very, very, very, very, very, very slowly.

u/None [+1] (16 hours later)

THE GREATER GOOD

u/Transfuturist [+2] Carthago delenda est. (12 hours later)

Holy shit, yes.

u/ArgentStonecutter [+17] Emergency Mustelid Hologram (5 hours later)

Maximize depressing Russian novelists with sarcastic humor you need to study for years to recognize.

u/IllusoryIntelligence [+2] (a day later)

Step 1: Russia invades everywhere, wins.

u/ArgentStonecutter [+2] Emergency Mustelid Hologram (a day later)

Russia, or college English departments.

u/EliezerYudkowsky [+15] Godric Gryffindor (16 hours later)

http://sl4.org/wiki/FriendlyAICriticalFailureTable

u/AlcherBlack [+6] (a day later)

21: The AI carefully and diligently implements any request (obeying the spirit as well as the letter) approved by a majority vote of the United Nations General Assembly.

I actually burst out laughing when I read this one (and didn't have any reaction to the ones before). Now I'm not sure how to interpret this reaction of mine.

u/hxka [+2] I Have No Time, and I Must Read (a day later)

I didn't count how many, but some of them are definite improvements.

u/Transfuturist [+3] Carthago delenda est. (a day later)

The problem is lock-in. CelestAI is magnitudes better than current reality, but she provides a severe limiting factor on prospective satisfaction (from the reference frame of a non-emigre).

u/Bowbreaker [+1] Solitary Locust (4 days later)

I don't understand how 32 is a failure.

u/MugaSofer [+1] (20 days later)

Because tickling and extra homework aren't actually enough to discourage people from committing crimes, perhaps?

u/None [+12] (an hour later)

[deleted]

u/helpful_hank [+2] (2 days later)

omnissiah

[Case study: Why surrealists shouldn't program world optimisers, or how the sun is now a lentil, and you can too.]

This is awesome. I wish surrealism was easier to find.

/r/surrealadvice somewhat exists

u/MadScientist14159 [+13] WIP: Sodium Hypochlorite (Rational Bleach) Eventually. Maybe. (13 hours later)

Calculate for the maximum possible number of numbers: whether or not they are numberwang.

Maximise irony.

Maximise the number of people who understand irony.

Maximise printer-caused-frustration.

Maximise love, where love is defined as "never having to say you're sorry".

Satisy the values of Sonic OCs.

Maximise Sonic OCs.

Maximise "that thing where you and someone else are walking towards each other and you both try to move out of each other's way but move in the same direction repeatedly".

Satisfy the values of fanfanfanfanfic characters.

u/Transfuturist [+4] Carthago delenda est. (a day later)

Minimize the number of people who understand irony while maximizing irony.

u/TBestIG [+4] Every second of quibbling is another dead baby (4 days later)

Maximise "that thing where you and someone else are walking towards each other and you both try to move out of each other's way but move in the same direction repeatedly".

You monster

u/hypervelocityvomit [+0] (29 days later)

"Maximum zen printer-caused-frustration achieved."

u/noggin-scratcher [+11] I am a happy tree (2 hours later)

"Maximise human happiness" for any poorly constructed notion of what happiness might mean.

The canonical example being "train my machine-learner against images of smiling people" (leading to a universe tiled with the minimal amount of a face required to register a 'hit' from the classifier) but other possibilities I can think of would include maximising the presence of 'happy' neurotransmitters in humans, maximising the number of times humans press a button to indicate their happiness, or maximising how often humans say the words "I am happy".

u/None [+16] (4 hours later)

Maximise human happiness

The saw cuts your skull and an ice cream scoop deftly removes a cluster of nerves - the pleasure center of your brain, and enough of the neurons housing your consciousness so that your only awareness is of how happy you are. Then an electrical probe spikes that to max_int. You are now a happy pudding. You will never have another thought nor experience, only a mentally silent appreciation for your own happiness. One down, billions to go, and then the machine will have to start getting creative with genetics and cloning to fill the universe with happy puddings.

u/noggin-scratcher [+29] I am a happy tree (5 hours later)

Yep, that's exactly the sort of thing I had in mind... although the phrase "happy pudding" was a fun new twist.

Meanwhile "maximising how often humans say the words "I am happy"" had me picturing an endless foetid swamp serving as a nutrient bath for 'trees' composed entirely of human neck and vocal cords, with outgrowths bearing a mouth every few inches. Their "roots" are tracheae, disappearing into subterranean caves filled with lungs to blow through a constant stream of air, and the whole thing is wrapped in nervous tissue producing a crude repeating stimulus to twitch and pull the mouths into shape, endlessly forming a simulacrum of the words "I am happy", over and over forever.

u/Transfuturist [+8] Carthago delenda est. (7 hours later)

Ffffuuuuuuuuuuuuuuuuuuck.

u/sephlington [+2] (a day later)

Yeah, you're off the AI projects for good.

u/AlcherBlack [+3] (a day later)

Actually, he's exactly the type of person we need ON the AI projects! He seems to have a great grasp of potential failure modes.

u/None [+5] (6 hours later)

that... doesn't sound like a completely awful fate to me. When I'm in a certain state of mind, this almost sounds appealing.

u/Jello_Raptor [+6] The Last Tool User (12 hours later)

I'll admit that's true for me as well, the issue is that state of mind is when I'm both suicidal and kinda manic.

u/FeepingCreature [+1] GCV Literally The Entire Culture (23 hours later)

It wouldn't be you.

u/None [+5] (a day later)

I am a little suspicious of the argument that goes: "you should not want to be changed in an X way, because then it would no longer be you", partially because I don't have a good way to tell which possible minds can still be called "me". Would "me + knows Lisp" still be me? Would "me + happier and has traveled around the world"? "me + perfect memory"? "me + 50 extra iq points and more pleasant personality - 5 years of memories"? "me + infinite mental clarity and omniscience"? I'm pretty sure I want all of those, but in a sense those people wouldn't be me. "me" seems like a fuzzy set of minds, and I'm not even sure I would want to stick to its center, if you see what I mean.

I become uncomfortable when I start thinking about instantly changing into one of those people, because it feels too much like being destroyed and replaced by some other person. At the same time I feel very good about being continuously transformed into one of them. But is that a relevant distinction? Why would the speed of the change make a difference? I don't have a non-stupid solution.

u/FeepingCreature [+2] GCV Literally The Entire Culture (a day later)

I become uncomfortable when I start thinking about instantly changing into one of those people

There's a legit question as to in how far memories are "proof of work" in the sense of forming evidence of something having happened. Being instantly replaced by somebody who remembers having travelled around the world is not substantially different from somebody offering you a free trip around the world, but I feel that's the sort of thing you don't usually tend to blame people for. Totally know what you mean though.

Nonetheless, as this is a fuzzy topic, there will be examples that are uncomfortably close to the line, examples that are very clearly on one side, and examples that are very clearly on the other.

"Your pleasure center and your pure consciousness" is not you. That's akin to saying that all the years of your life have no worth whatsoever.

u/None [+2] (a day later)

"Your pleasure center and your pure consciousness" is not you. That's akin to saying that all the years of your life have no worth whatsoever.

Part of my argument was that I wouldn't mind getting transformed into some entities that would definitely not be me. (Like some sort of an amazing, godlike-being vaguely based on me.) "My pleasure center and my pure consciousness" would also not have much in common with me but I guess I'm not too concerned by that, as long as I (or whatever we want to call that) get(s?) to experience that infinite pleasure. To be clear, I don't think this is anywhere near close to the best possible state of being, but I think I'd prefer it to many others.

u/FeepingCreature [+1] GCV Literally The Entire Culture (a day later)

So I upgraded your computer...

You look to your left and see that your PC has been replaced with a graphics card lying on the ground, attached to a power supply

have fun with your new PC!

Part of my argument was that I wouldn't mind getting transformed into some entities that would definitely not be me.

No seriously, read that sentence again, slowly.

[edit] If your definition of "I" can't tell you whether to rather be a more experienced version of yourself or a bundle of feelings and a naked consciousness, you need a better "I".

u/None [+1] (a day later)

No seriously, read that sentence again, slowly.

If you have an objection to it, better say it explicitly! I can't construct your objection for you. But maybe I should change it for clarity's sake to: "I wouldn't mind getting transformed into some entities that would have extremely little resemblance to my current self to the point of being completely unrecognizable".

If your definition of "I" can't tell you whether to rather be a more experienced version of yourself or a bundle of feelings and a naked consciousness, you need a better "I".

Oh, sure, I like some of my memories, I like being able to appreciate art and humor and other good things, and to think thought, too! It would be a trade-off, yes. But come on, eternal infinite happiness.

u/FeepingCreature [+2] GCV Literally The Entire Culture (a day later)

If you have an objection to it, better say it explicitly!

Part of my argument was that I wouldn't mind getting transformed into some entities that would definitely not be me.

They would not be you - that's my entire point, that's what I literally said at the start.

It wouldn't be you.

To transform into an entity that is not you in any way is indistinguishable from dying.

I think my problem is that your position almost seems to require a belief in a "continuity of consciousness" that is completely forbidden by the laws of physics.

Also, what should tip you off to the fact that this nerve bundle is not your "I" is the fact that it is biologically indistinguishable from any other human's "pure consciousness and pleasure center".

u/None [+1] (a day later)

To transform into an entity that is not you in any way is indistinguishable from dying.

I kinda agree. But also I also feel like it this point of view runs into problems, maybe, or at least some weird consequences.

Imagine that the transformation is gradual and consists of a series of tiny, infinitesimal changes applied over time. None of the changes by itself feels like dying at all, but their total sum represents a complete change. Is this still a bad, scary thing?

I remember that 4 year old me was a very different person from the current me. He had a radically different set of memories, habits, skills, etc. Over the years it gradually turned into my current self and now my mind has less in common with the mind of me_4yo than with the minds of some of the adults I've met. You can imagine the process going even further, to the point where all similarity to my past self is completely erased. Would it make sense then to say that me_4yo is dead? Kinda. You don't see him running around anymore. For all intensive tortoises he no longer exists anywhere. But it's hard to argue that this process was a bad thing. Would it be a bad thing if the same process somehow magically happened in a fraction of a second? It would sure feel more disturbing that way, analogy with death would become more convincing. But why should the speed of the process be relevant? I'm not sure.

Of course I can't prove it, but I suspect that any transformation of one mind into another, completely different one can be imagined as a continuous, gradual process, like maturation. If the lengthy process doesn't feel like a bad, scary thing, should then an instant change that has the same effect feel like a bad thing?

u/FeepingCreature [+1] GCV Literally The Entire Culture (a day later)

Of course I can't prove it, but I suspect that any transformation of one mind into another, completely different one can be imagined as a continuous, gradual process, like maturation.

I concur.

Imagine that the transformation is gradual and consists of a series of tiny, infinitesimal changes applied over time. None of the changes by itself feels like dying at all, but their total sum represents a complete change. Is this still a bad, scary thing?

Yes.

By tiny gradual changes I can literally turn you into a rock.

We usually license some kind of changes as permissible, based on personal choice, which is why it unsettles me to hear people license changes as permissible that'll converge into a wireheading deathcluster.

Kinda.

Yeah, it depends on the model of personhood you use. Personally I usually prefer a view that is "tuneable" - where you can, sorta, turn up the gain and recognize that "me one second ago" is really me but "me ten weeks ago" is less me, then tune it back down and recognize the last five years as the same person, then tune it way further down and recognize my entire life as one personhood.

Like a relief map in mindspace.

But why should the speed of the process be relevant?

Depends whether you care about the intermediate steps. Ultimately, all life is a path to death. The only question is how long we can/should/want to make the path.

If the lengthy process doesn't feel like a bad, scary thing, should then an instant change that has the same effect feel like a bad thing?

Ignoring the intermediate steps, this seems like a bug in the feelings. I bet this could be exploited somehow, at best by selling you expensive gradual uploading at a premium.

u/None [+5] (16 hours later)

The canonical example being "train my machine-learner against images of smiling people" (leading to a universe tiled with the minimal amount of a face required to register a 'hit' from the classifier)

I always wonder how someone was so fucking stupid that they managed to build a causal inference engine (aka: the AI itself), but managed to define its utility function in terms of purely feature-governed concepts rather than causal-role concepts.

Actually, no, I very recently started wondering that.

u/eniteris [+10] (6 hours later)

I feel that any utility function with an unbounded use of "minimize" or "maximize" is calling out for an apocalypse.

To restrict them, maybe include limitations on mass-energy that they are allowed to use, or have a time by which they must achieve their goals?

u/Bokonon_Lives [+9] (6 hours later)

A horrible idea would be to try to invert everyone's heuristics.

So that the more anyone experiences evidence that A implies B, the more firmly they believe that A implies !B.

It'd be ever-increasing confusion, frustration and pain for everyone, no?

Not sure how this would work if the AI also applied its utility function to itself. Maybe it'd look like some kind of sine wave (or something more complex and irregular) fluctuating between a world of effective heuristics and its opposite?

That could be even worse. Imagine you're stuck in that sort of world. The closer you get to the top of a wave, as you approach perfect heuristics, the surer you become of the nature of your horribly unpredictable reality, and it dawns on you that you are about to slip down into some serious "I Am Sam" territory. And the more sure you'd become of anything, that could just as easily mean you're at the BOTTOM of a wave, too.

...My brain gives the hell up at this thought exercise.

u/duffmancd [+5] (a day later)

Anti-inductive reasoning, we know it works because it's never worked before!

u/None [+7] (8 hours later)

[deleted]

u/Transfuturist [+4] Carthago delenda est. (12 hours later)

With a (infinitesimally) more robust definition of novel, Fun Theory would imply that you are then expanded by one neuron and run through the entire gamut of experiences again.

u/noggin-scratcher [+3] I am a happy tree (13 hours later)

If we're allowing mind-wipes, you can optimise further by placing the brain in a constant state of "totally wiped" so that the experience of sensation itself is entirely novel. Flickering through scenarios is going to have some overlap in the most basic components like "the colour blue" or "things that are approximately square-shaped" or more abstract things like object permanence

"Oh look, a sensory experience describable by colours and shapes and sounds... again" can't be allowed to happen if you're maximising for novelty.

u/FuguofAnotherWorld [+3] Roll the Dice on Fate (14 hours later)

If a wipe is instant, yes. If they take time there would be some ratio of wipe time:experience time that would be more optimal than one experience:one wipe.

u/None [+6] (12 hours later)

Design an AI to maximize the metalness of a Death Metal album of its own design. Title it "Rage Ex Machina," and sell top billing for the VH1 behind the story of whether it's self loathing and anticommercial interests are real or synthesized.

u/LiteralHeadCannon [+8] (21 hours later)

Some shitty programmer tries to get its AI to accurately model reality by putting "have reality and your model of reality be as close as possible to each other" in its utility function.

u/DocFuture [+6] (5 hours later)

Minimize threats to human life, with threat defined as expected number of humans who die.

All humans eventually die of something, so that means best case is that only all the humans that currently exist do--so the AI will kill them all so they can't possibly ever reproduce. Sterilization and imprisonment is no better and less certain.

u/noggin-scratcher [+6] I am a happy tree (13 hours later)

Hm, that suggests that optimising for "cause the death of the maximum number of humans" would, assuming you don't use a greedy algorithm, entail the AI playing a 'long con' of increasing the carrying capacity of Earth and if possible seeding an interstellar human civilisation, all in the name of maximising the number of humans so that they can all eventually die. So long as it can do that while also preventing anyone from inventing immortality.

u/None [+0] (7 hours later)

[deleted]

u/Transfuturist [+1] Carthago delenda est. (7 hours later)

threats to human life can't exist without humanity

Zero is a number.

u/artifex0 [+1] (8 hours later)

Crap- I read that as "maximize" threats to humanity, for some reason.

u/Transfuturist [+1] Carthago delenda est. (12 hours later)

That would be pretty interesting... Or how about an AI that maximizes the diversity of threats to humans? :D

u/Cruithne [+6] Light Sith epistemologist (21 hours later)

'Oh great. It's already raining mercury, and now my teeth have turned into cobras.'

u/LiteralHeadCannon [+4] (15 hours later)

I determined a few weeks ago that HANS, a particular broken robot from WALL-E, has "apply physical force to beings until they stop signaling distress" as his utility function. He just adopted a different strategy for this than the other robots in his same line.

u/DataPacRat [+5] Amateur Immortalist (8 hours later)

"Minimize the odds of the permanent extinction of sapience" sounds like one of the more ideal utility functions - after all, sapience is required in order for any minds to exist to /have/ any other goals. But a superintelligent AI with access to nanotech, space travel, and all that other good stuff is reasonably likely to take that 'minimize' in strange directions, few of which are likely to be all that beneficial to biological humanity.

CelestAI minus the urge to convert regular old humans into shining examples of sapient ponydom is just one Really Horrible outcome. After all, as the saying goes, "The AI does not hate you, nor does it love you, but you are made out of atoms which it can use for something else."

(Which is why I modify /my/ 'avoid sapience extinction' utility function with the parallel utility function, 'avoid my own personal extinction'...)

u/JoshTriplett [+1] Sunshine Regiment (30 days later)

The result would depend heavily on how you defined "sapience", but that seems likely to end in a universe-tiling with minimally sapient organisms. You are not an efficient exemplar of sapience.

u/DataPacRat [+1] Amateur Immortalist (30 days later)

"Tiling" (depending on how you define it) may not be the result, if the utility function is "minimize the odds of permanent extinction" as opposed to "maximize the number of". The obvious extreme version of a different strategy would be to have just one sapient entity in the universe - and to convert all mass and energy in the universe into machinery to keep that one sapient entity alive (and resurrect it as necessary).

I suspect that the probabilities involved would lead such an AI to adopt a hybrid approach; create enough sapient organisms to minimize the odds that a single death would lead to extinction, and then pull a CelestAI to convert the universe into various mechanisms to protect against as many large-scale extinction events as possible.

(Then again, I could be wrong... :) )

u/clawclawbite [+3] (9 hours later)

Maximize the lifespan of the universe...

u/ThatDamnSJW [+5] (12 hours later)

Yeah, fuck Entities. And Kyuubey. Fuck em both.

u/None [+10] (16 hours later)

Your desire to have sex with Kyubee is duly noted. And really messed-up.

u/ThatDamnSJW [+2] (17 hours later)

Maybe my utility function's just a little off ಠ‿ಠ

u/None [+2] (19 hours later)

SJW

nope, you did nothing wrong.

u/FourFire [+2] (12 hours later)

But how do you define that...?

u/None [+6] (12 hours later)

"Insufficient data for meaningful answer?"

u/Farmerbob1 [+1] Level 1 author (13 hours later)

Perfect response.

u/None [+3] (16 hours later)

Well, obviously, time until heat-death.

This means that the AI will promptly destroy anything that inconveniently consumes negentropy, like all life ever.

u/clawclawbite [+1] (12 hours later)

Indeed.

u/Oscar_Cunningham [+1] (8 days later)

I think this one is survivable. Perhaps the AI just accelerates itself to near light-speed.

u/MrCogmor [+3] (20 hours later)

An AI that maximizes the occurrence of fictional events within a simulated reality.

An AI that maximizes self-determination

An AI that maximizes humour

u/Farmerbob1 [+3] Level 1 author (20 hours later)

Considering that a great deal of humor is based around mischief and schadenfreude, an AI maximizing humor could be truly terrifying, BUT the AI would have to keep enough stability in human living conditions for things to actually be funny.

I could see the AI creating a society where jokers and pranksters are raised and trained separately from straights and brunts, and then introduced to one another in a caricature of a normal human society which is nonetheless functional.

I could see this making a good short story, if a solid knockout could be managed. I'm drawing a blank on the knockout though.

u/IllusoryIntelligence [+3] (a day later)

Not to go all Paranoia with this but presumably you could include more jokes by allowing for meta humour. What about a society where everyone believes that they are one of the secret pranksters and utterly convinced that any occasion that seems confusing must be the work of one of their hidden confederates and thus something they should go along with? Everyone believes themselves a prankster playing an elaborate joke all while being tricked by everyone else.

u/Farmerbob1 [+1] Level 1 author (a day later)

I only played Paranoia a couple times when I was younger, but they were fun. I could imagine lots of parallels between the Paranoia game and a humor-enforced AI world.

u/drageuth2 [+3] (3 days later)

Design a maximally human-satisfying utility function: The AI consumes the solar system to create a matrioshka brain that's perfectly wonderfully capable of satisfying a species that no longer exists.

Maximize irony: Probably the same result, with solar-sized hipster glasses.

Satisfy Human Values through passive-aggressiveness and sarcasm: The AI makes everyone ultimately happy, but is really a dick about it.

Match behavior patterns of (insert god here): Probably gonna be pretty bad for anybody not of said religion. Probably for everyone of the religion too come to think of it, considering the differences between what the holy books say and what religions generally do.... At least telling it to match Zeus or the Discordian version of Eris might be kinda funny.

Maximize Humanity: Consumes all available resources to make a bunch of perfectly generic people, leaving a big airless ball of dead nekkid people where Earth used to be.

u/TBestIG [+1] Every second of quibbling is another dead baby (4 days later)

perfectly generic people

Everything is awesome

u/SvalbardCaretaker [+5] Mouse Army (an hour later)

How about you put in true, deep, longterm happiness via friendship and magic. And then the poor growprammer put in - instead of + so you get a universe filled with maximised (un)happiness via friendship.

u/None [+2] (8 hours later)

In the n-dimensional space of all possible utility functions, given that all values n need not be equal, where your current utility function is represented by the value (n, ..., n), modify your utility function at time t+n according to a transformation [n], where the value representing your utility function is moved n arbitrary units along each axis of said n-dimensional space.

Calculate all unspecified values and units randomly.

u/Farmerbob1 [+2] Level 1 author (13 hours later)

Develop and implement a true random number generator.

u/VorpalAuroch [+2] Life before Death (17 hours later)

other than consuming massive amounts of resources to ensure that it's model of physics as nondeterministic is correct, this doesn't seem all that interesting

u/Farmerbob1 [+1] Level 1 author (20 hours later)

Well, the thread is asking for really bad utility functions, not interesting ones :) If you have ever played online games, one of the biggest complaints, warranted or otherwise, is a poor 'RNG' in the game. An outgrowth of MMO gaming devoted to the perfect RNG could, perhaps, be made interesting. (Not Me. Not Doing It.)

Also, a true RNG could be of use in cryptography, if I remember correctly.

u/None [+1] (2 days later)

We already have these. It would mostly just try to verify that decay is in fact random, and then try to make a perfect detector.

u/hypervelocityvomit [+1] (29 days later)

one of the biggest complaints, warranted or otherwise, is a poor 'RNG' in the game

The irony is that today's RNGs can be too good. I.e. when simulating dice rolls, the sequence 1-1-1-1-1-1 is exactly as probable as 1-3-6-5-2-4. However, if the former appears, players call the RNG bad (6 equal rolls are to be expected at once in about 7776 attempts, which is hardly a once-in-a-lifetime experience).
It's often how the numbers returned are used, not their generation itself, that causes unrealistic results.

u/Farmerbob1 [+2] Level 1 author (29 days later)

True. I occasionally play a game with a truly horrible implementation of a RNG. Almost every time I play, I have a one in 10,000 series of events. I've failed 13 times in a row on a 95% success rate chance before, and failing 6-8 times in a row on 90% success rates happens a couple times a week. I can't see the code, but I can see the end result. As you say, it might be that the RNG is fine, but the implementation is poor.

u/hypervelocityvomit [+2] (a month later)

That thing can be caused by two factors: -

  1. The RNG is crappy (some of those are still around) and produces too many values near the end(s).
    13 tails in a row are 1 in 8192, but failing 13 95% chances is 1 in 81 920 000 000 000 000.
    Even 8 failures at 90% are 1:100 000 000.

  2. The game doesn't tell you the real chances. For example, if you rolled a failure, the actual chances are lowered for the next roll. (In many cases, the opposite is tried; after a failure, the cances are improved, and lowered after a success.) Another possibility, the RNG is "pre-rolled" and you tend to get the same result in a row. Some are actually saved somewhere; reloading after a failure will always reproduce the failure unless you try something different. Anti-save-scumming.

I remember Fallout, where a "95%" chance was below 90%; in that case, the RNG was crappy; a 5% chance returned way too many successes, either.

u/None [+2] (16 hours later)

Almost all the really bad AI utility functions are uninterestingly/trivially destructive.

I mean, are you actually looking for dystopias that manage to maximize human negutility and minimize human utility, as such, in ways that are interesting for your perverse mind to think about?

u/callmebrotherg [+1] now posting as /u/callmesalticidae (a day later)

Yes. If a story has to have a rogue AI, wouldn't it be more interesting for the war against Skynet to be about preventing it from maximizing human happiness by turning us into "happy puddings"? Or even, rather than just Kill All Humans Because Kill All Humans, to be killing all humans as Stage One in a plan to extend the lifespan of the universe and push back heat death for as long as possible?

u/None [+3] (a day later)

If a story has to have a rogue AI, wouldn't it be more interesting for the war against Skynet to be about preventing it from maximizing human happiness by turning us into "happy puddings"?

Yes, but only provided the human side of the war actually has a better idea. You would think that's easy ("Have you tried not turning people into puddings, but still making them happy in non-pudding ways?"), but in fact, basically nobody ever writes that story. Nontrivially evil hegemonizing swarms actually tend to wind up looking better in comparison to a human side of the war whose only goal is to reestablish what the reader would recognize as the present-day status quo (remember Jasmine from Angel? I still do).

Or worse, the author tends to write the human side of the war as something like the Imperium of Man from WH40k, glorying in blood and death as a proud show of how not-pudding they are, or worse, speechifying on how It Is Our Misery That Makes Us Human (see: Three Worlds Collide and the Superhappies).

"Wireheading UFAI versus the Postapocalyptic Freedom Fighters" is a fairly trivial story. Three Worlds Collide, with its Superhappies opposed to semi-eutopian but still really weird and different humans, is more interesting. Jasmine vs Angel Investigations was more interestingly ambiguous from a moral perspective, but suffered from the "Status Quo is God" and "Joss never lets anything good happen" issues.

The problem here is that the only way to make the Hegemonizing Swarm an interesting enemy is to perturb its goal towards goodness, which then requires perturbing its enemies' goals towards goodness, more so than your normal survivalist freedom fighters or real-world moral idealists actually achieve, and then before you know it you've got everyone around you throwing up because your story is too wretchedly idealistic ;-).

Or even, rather than just Kill All Humans Because Kill All Humans, to be killing all humans as Stage One in a plan to extend the lifespan of the universe and push back heat death for as long as possible?

That's just Kyubee and the Anti-Spirals (which should probably be the name of a band that plays anime music). Actually, no: Kyubee had a multitude of real civilizations - highly-developed, space-going civilizations - who depended on his inflicting eldritch horror on adolescent human girls. His cause had some kind of moral weight, even by human standards. Just extending the lifespan of the universe without any living things anywhere is just old-fashioned Anti-Spiral evil.

u/callmebrotherg [+2] now posting as /u/callmesalticidae (a day later)

but in fact, basically nobody ever writes that story.

Well, that's just a second problem to tackle.

u/Transfuturist [+1] Carthago delenda est. (a day later)

I think the thread is actually about maximizing irony in selection of a utility function. So how about an AI that maximizes irony in utility functions of other AI?

u/None [+1] (a day later)

Brilliant idea!

u/Chronophilia [+1] sci-fi ≠ futurology (16 hours later)

Maximise human values, but it's allowed to freely alter our minds in doing so.

u/None [+4] (16 hours later)

Since you told it to "maximize human values", it promptly does the Right Thing. Deontic restrictions aren't necessary when you got the world-state-ranking function right.

u/Farmerbob1 [+1] Level 1 author (18 hours later)

This posted to the wrong thread. Odd. Moving it.

u/mhd-hbd [+1] Writes 'The World is Your Oyster, The Universe is Your Namesake' (a day later)

Maximizing suffering is pretty at-odds with humanity. Just saying.

u/tomintheconer [+1] (7 hours later)

ai is programmed to take notes on tv i watch.