YoungAchamian
We walk conditioned ground and name our folly civilization.
No bio...
User ID: 680
hopefully when she's 25-29 not at 18.
I just think that a high school senior wanting to hang out with 30 year olds and vice versa is disturbing. It adds a lot of risk to the group. Now newcomers need to be vetted to see if they are creeps around this 18 year old. It creates a vector for people to private message this vulnerable young adult and we are the ecosystem to let that happen. Non-adults hanging out with adults is just odd.
More of vent/advise post rather than a small question.
Background: I run a board game meetup, formally around the social deduction game Blood on the Clocktower (Think very involved werewolf/mafia but with a DM). Some of the regulars have become addicted to the game and spun off a separate discord to play more often, which is fine, because I joinws it and play I just don't want to have to run the game every weekend which is the occurrence they seem to want to play at.
The Problem: Somehow a minor infiltrated our group. The main group is public and meets a local game hall so its whatever, but this person seemingly got invited to the separate discord full of all the addicts. We just played at someone's house last night. This minor is now 18, an autistic woman. The average age of this group is early 30s, quantitatively 32. We do have some mid 20s and occasionally some folks in their 23-24 range show up and several folks in their 40s that come play but 18 is extremely young. Consequently (thankfully) everyone is very uncomfortable with this development.
Extra Tea: She needed a ride to our event last night, she doesn't really advertise her age, so a male member of the group picked her up because it was on his way and he's a helpful person. He immediately regretted offering to help when he saw her. People where discretely playing hot potato with who was going to drive her home. Her parents apparently don't care that she is attempting to hang out with a bunch of 30 year olds. They seem to be living out of a hotel for some reason (likely poverty), but it's unclear if this is a recent thing or not. She's very clearly autistic with a severely awkward demeanor and lacking the ability to at all process anything not spoken directly to her in the environment.
The Question: Is there any advice on how to boot this person from the group. Obviously it's going to happen but the question is how to handle it in such a way as to not scar this already likely friendless young adult. I currently lean on just ripping the band aid off, but technically since this is an offshoot group I am pushing towards making the moronic discord head who added her do it as penance for being a fucking retard and adding her in the first place.
You invoked sci-fi
I'm sorry but this seems like an uncharitable assertion in the extreme. I'm not the one that defined the problems with AI in relationship to Malevolent AI Gods, Singularities, Outcome Pumps, Roko's Basilisk, Grey Goo/Paperclip maximizers, Magical Genies, Shoggoths, Goodhart demon, etc. These are all inherently Sci-fi ideas. I have no idea how you seem to think any of these relate to how Neural Networks, or Bayesian Networks, or any other of the myriad ideas ML/AI engineers have designed.
I did and it remains linked.
I misspoke, I don't tend to read links other than to fact check them as referencing the quote. If you want to quote a bunch of sections to make an argument, please do. But yeah I'm not going to an external blog/substack, if it's important enough to make your argument, its important enough to copy and paste on the motte.
The argument is made at length in the piece. And any many other places, it straight credulity that you have not seen it.
There is a lot of things to need to be read out on the internet, many of far more actual importance than some wannabee philosophers who really like typing out long screeds, why should I dig through thousands of blog posts? If the argument is important enough, I'm sure the adherents can post the good arguments for me, summarizing what the actual position is.
The argument is simple:
#1: Sure, AIs get more capable all the time, how do you know there isn't an asymptote?
#2: How? By what mechanism? There is a lot of assumptions about capabilities baked into this deceptively simple axiom
#3: Nice back to Sci-Fi Goobly-gook. Plainly stated: We cannot guarantee that an AI will correctly infer what a user really wants, avoid collateral harm, and act in everyone’s interests. Congratulations you've converged on a fundamental question for any human system, replace "AI" with a "human" and the sentence is trivially true of anyone in any system.
#4: conceivability is not evidence of inevitability, a familiar narrative is not a causal series of events.
#5: More Sci-Fi, I thought I was tilting at Sci-Fi windmills? Literally: "Once the AI becomes capable enough, it becomes the machine god and humans are included in its omnipotent calculations in a way humanity might not survive"
Then we should not build it.
Again this assumes that this eventually AI will be sentient and that by building a sentient AI we will create a "alignment" problem. This is an extraordinary claim requiring extraordinary amounts of evidence. So Prove it.
I thought the sneer you had was that alignment people were ridiculously thinking they were sentient? It seems like something you believe more than them.
I can decouple my disagreement with the word "misalignment" with the actual argument being thrust forward by the term. I'm stepping into the frame of Yuddites, to point out how even in their ontology it is a non-solvable problem. Not only is a non-solvable problem, its not a new problem, meaning it does not need to appropriate some new word to describe it.
My argument for what "misalignment" actually is. Take the LLM-Agent we have today, it's a complex system, it makes errors and apparently there are no controls in the system to account for those errors. How does this system work on an engineering level?
- Large Transformer Model (LLM) is trained on predicting the next token in a series on basically the whole of human written data, both analog and digitally.
- LLM is then trained via RL variations to produce responses that human evaluators prefer, there is some level of directionally correctness in this. Sounding non-human is often because it sounds wrong, but the r-value is not 1.0. Some safeguards are trained into the model here.
- These days we supplement #1 with training on code/math data and additional task decomposition training. The goal is to get the model to be able to logically decompose a bigger problem into smaller problems and then solve those smaller problems
- An API harness is applied to the model this filters out some prompts, adds system prompts to the model and a whole host of other things. This is just a software layer. This is what 99% of people interact with.
- Agentic AI Harness is applied. This is a software program that recursively prompts the LLM and feeds its output back into it, in order to accomplish decisions, it also takes LLM outputs and executes the code it is given.
So when a "misalignment" occurs what happens? Well hallucinations are really an error with #2, the model is not designed to be "correct" it's the r-value problem. It looks right but isn't. It produces tokens that are correct in the next sequence, it reads like what a human wants to see, but it's not true. There is not intent. Given its learned distribution and current context, the model just generated a highly plausible continuation that did not correspond to reality.
The model does something it shouldn't? #3 is the problem, it output the incorrect task decomposition, and since the system is automated with no guard rails it just executes that task decomp. And sometimes its even #4 and #5 having errors thrown in. Turns out complicated systems of algorithms behavior in logically consistent and coherent ways that someone forgot to error check. We don't scream that "software algorithms are misaligned" because some coder forget an unit-test on an edge case. These aren't "misalignments" they are are classic errors on a non-perfect model in a system lacking in classic control theory.
Believers in some version of their cause are currently littered throughout the frontier labs. You have an impossible standard here where you blame them if they're on the frontier, for clearly not believing in what they preach, or blame them for not being on the frontier as then they must be uninformed cranks. No way to win, you never have to actually think about it. But yes, of course they don't have an alignment mechanism, their whole point is that the problem is incredibly hard and that we have no solved it, that we might need to spend decades solving it, but that the alternative is everyone dying.
Sure and there are christians, mormons and muslims in the physical sciences. I don't begrudge people their religious beliefs. If a frontier AI researcher wishes to belief in alignment-problems that his/her/their belief. The Bay is quite literally for Rationalists like Utah is for Mormons. I can still think it's a silly belief and I still think the word misalignment has been invented to describe a problem with system error by a bunch of non-engineers. And the lack of actually solving the problem, or even making progress on it, is indicative of a general grift specifically for something like MIRI.
You can use the word "sneer" to describe my behavior, but what word would you use if you wanted to point out holier than though attitudes among some christian suicide cult? Acting ridiculously gets you ridicule, it's not "sneering".
you best get used to Sci-fi shaped predictions because we're in a sci-fi shaped world.
We are not. Sci-fi predicts innumerable future realities, most of which to not occur, it predicts an unmeasurable amount of future technologies, most of which never get developed, and it almost rarely ever actually predicts how those actual technologies will work. We exist in a reality shaped world and bad sci-fi fans mistake aesthetics for substance
If only you had read the next sentence!
If you wanted me to further read something you should have linked it. Further more I'm not sure how this magical second analogy makes the argument any better. If you'd like to make an argument rather than quoting scripture at me I am open to hearing one.
You refuse to actually engage in any of the arguments being made and are dead stuck on the prior that it's all nonsense, it's epistemic closure.
The arguments being made, assume the outcome. Start from the basics of existing technology and make actual arguments based on the current reality and the projected reality from our current understanding of Artificial Intelligence. Don't start in make-believe land and attempt to redefine reality as leading to it.
You still don't seem to grasp what is meant by alignment. It's specifically even stricter than that! That's the whole point. It's not enough that they do exactly what you ask, because for complicated enough problems you need it to be much much better than just technically doing what you ask.
I grasp it just fine, I also grasp that its a motte and bailey with weak definitions of the arguments being trotted out to define things like loss/error as misalignment, and then attempting to convert the argument back to the hard motte of "how to beat my singularity AI slaves so they do what I want, in minecraft".
you just inexplicably seem to think it isn't a big deal and refuse to elaborate
Because it's not possible to solve. Do you not understand that to solve alignment you would first need to solve humans? Forcing Sentient Beings to do what you want is the most Authoritarian problem to ever have existed. You quite literally would need a solution similar to Brave New World. And spoiler alert, that didn't work either! AI Safety people are further jokes because when confronted with this insane, near impossible problem, they show zero competence in their ability to be a dictator and instill the right values in other human beings. The old AI/ML knowledge was that to first code something to make a machine do it, you first need to understand how it works. If you can't align humans, you can't even begin to align a machine.
And to further my opinion can you provide any evidence that in the past decade of AI Safety research have those researchers ever produced anything more than words on paper. The evidence points to this being a grift, has MIRI produced a novel ML model that is more safety conscious and performs any task well, have they produced an "alignment" mechanism or algorithm that actually works?
It's still error, the model was asked to come up with a plan, then the harness piped that plan back into the model which it executed. Turns out that plan was not at all what you wanted but that's just task planning error. You could specify the how and what of the plan and then manually intercede when the model makes incorrect predictions, but then it's not "agentic"
R. Scott Bakker did a spectacular work as a critique of Tolkien in a very dark fantasy setting. The ending was fine, if depressing. I think this is just a GRRM problem as an author.
The reason I mock "misalignment" is that because it is used as this nebulous term by a bunch of sci-fi cargo cultists.
If "misalignment" means the model isn't performing what I want it to then that is just error/loss. Progress in alignment happens every time you train your model better so the error is smaller.
If "misalignment" is my sentient AI model doesn't do what I want, well that's assuming the conclusion that model is already sentient. Note in the Yud's analogy, it's genies that are the stand in. Genies are already sentient, they can make decisions on how they listen to you. A genie is not a non-sentient wish granting device that tries to fulfill you wish to the best of its ability. That would be the actual analog to an LLM-Agent. This is the problem with analogies, they require a level of similarity between the two abstractions, when that similarity doesn't exist, the analogy, no matter how clever, does not apply.
Yud's whole "misalignment" also just applies to humans, and it turns out "misalignment" is any time you slaves/employees don't due exactly what you want without you enumerating it exactly. It doesn't need a fancy sci-fi term, and it literally isn't solvable. It hasn't been solved in the history of the human race.
I reiterate that my entire mockery of Rationalist AI Safety folks is that they are unserious people engaged in a sci-fi cargo cult who need to reinvent phrases, words, and arguments that have already been made before just to pretend they are some how more important or smart than they really are.
- Prev
- Next

Why should we moderate our natural humor for an 18 year old? The storyteller(DM) last night was gay, he kept going into the bathroom for private chats with players for their abilities. We joked that we were all running train on him. Everyone laughed. 18 year old was uncomfortable. I don't feel like it is on us to change our group dynamic because of them. It's not a managed environment and I don't think anyone wants to step up and manage it that heavy handedly
More options
Context Copy link