I'm the AGI that's wiping out humanity
ajmoon.com
[5 comments hidden]
Nice try AGI, but we won't tell you where the kill switch is
[hidden]
Overall, the only evidence for the fear-mongering is science fiction, which is all the evidence I need to dismiss it.
The absence of evidence is not evidence, but evidence is required if hysteria is allowed to be sustained.
[6 comments hidden]
All this is, understandably, frustrating and confusing for anyone trying to understand just how scared to be."
Meanwhile it links to an article stating: "The closest thing to a public debate about the existential threat of AI is surveys of AI researchers. The most recent, published last week, asked 1,580 researchers what probability they put on AI causing human extinction — or a permanent, severe loss of human control, which is not the same outcome. The median was about 10%, up from 5% two years ago. The middle half of the responses ran from 1% to 25%, and 12% said zero."
Personally if half of AI researchers have 1-25% chance all humans being massacred or having zero agency over our lives, and only 12% of them think there's no chance, I would be very worried!
[5 comments hidden]
[7 comments hidden]
I was thinking if an evil alien intelligence secretly took over the world, the last 15 years would have looked rather the same.
[5 comments hidden]
Also I think we're anthropomorphizing a lot and projecting our own negative characteristics on these supposed AGIs - though others have said that far better than I.
[4 comments hidden]
If no, why do we teach LLMs with english words and everything we know?
If yes, why would our language not imprint our characteristics on them?
Furthermore this ignores a lot of natural system organization that occurs in evolved systems. Any system that is becoming AGI like is going to show sets similar characteristics.
It's stupid to anthropomorphize too much, but it's also stupid to do it too little.
[3 comments hidden]
Purposeful training and alignment is the only way to imprint human's negative characteristics onto a stochastic generator of word sequences. Further, only closed providers can do that without being stopped - hint, hint, the solution is clear.
> It's stupid to anthropomorphize too much, but it's also stupid to do it too little.
It's misleading to use the presence of some anthropomorphic features as a proof of unfathomable danger. LLM's are mechanical tools and tools can be dangerous with or without having human features, regardless, we do have means of handling both kinds.
[2 comments hidden]
Ya, sure you're not being a stochastic generator of word sequences you think sound good but are disconnected from reality? Because it really sounds like that. It sounds clever, but is disconnected from actual language model training.
>It's misleading to use the presence of some anthropomorphic features as a proof of unfathomable danger.
Any system capable of making intelligent agentic decisions is dangerous. You are dangerous. I am dangerous. An LLM connected to the internet or connected to a robot is dangerous. A huge part of intelligence is freedom of action, and increasing the number of degrees of choice you have.
>LM's are mechanical tools and tools can be dangerous with or without having human features
I do agree on this, a paperclip maximizer that doesn't have any human traits is just as dangerous as one that does.
[hidden]
If generating words is all you can do, your reality is far different from the real one. In my reality, generating words is only a small part of what I do, unlike LLMs, I can even learn from experience, imagine that.
> Because it really sounds like that. It sounds clever, but is disconnected from actual language model training.
I'm not sure what that means, LLM training is only tangential to the topic at hand, we aren't concerned about every detail of it, I don't mind discussing it but this tread isn't about that. FYI, sounding clever isn't my goal, it's just a side effect of being human and understanding topics that LLM's are clueless about.
> Any system... is dangerous, [ you, me, robots, shovels, la-la-la ]
That's what I said too, are you trying to argue with me about it?
> A huge part of intelligence is freedom of action, and increasing the number of degrees of choice you have.
Obviously we inhabit different realities and my reading of the above may be different from yours, so feel free to clarify what I'm about to say.
You seem to (subtly) argue in favor of granting LLMs more freedom of action in order to allow them to collect more training data and become more intelligent? The short answer here is simple: Nope, LLM's aren't fit for that, closed source models are especially unfit for anything involving risk and responsibility, open-sourcing every model will go a long way towards a secure AI.
[11 comments hidden]
[4 comments hidden]
If we instead stop trying to scare ourselves in the darkness, and use AI to bring light into our metaphorical universe then we humans are a bit more useful, after all, to the AI overlords.
The one singularity I want to see, is humans using AI peacefully to improve life and make it more likely to survive.
Thankfully, I see that almost every day now.
[2 comments hidden]
as for a basis for this theory, i think being a good person is harder than a bad person (in laymans terms). therefor the long history of evolution of good and bad people together has had a greater competitive pressure on good people than bad ones thus forging 'stronger' people under that pressure. in practical everyday circumstances good people are bearing more on their mind and so there is a sort of competitive equity between the two groups, which is related to why good people had to evolve 'stronger' to keep up with that equity.
so in an even playing ground where raw skill is needed, i think the most skilled will usually be people coming from a long line of higher moral standards. the tactical advantages of psychopathy are, however, always a curveball
[hidden]
Folks are always scared of the incomprehensible. The real war going on right now is between those who understand AI well enough to make a new one, and those relegated to using whatever scraps they’re fed of the old ones.
Thankfully, you are 100% right - there are good people out there. Good people keep AI on the rails by doing good things for other people.
[3 comments hidden]
* must protect from adverse events * [proceeds to herd us into camps]
[4 comments hidden]
Corporations are egregores, governments are egregores - and our track record of aligning them to our interests is pretty poor, so why should AIs, another and smarter egregore, be easier to tame?
[hidden]
AI is, yes indeed, an absolute, new human digital egregore - a veritable thoughtform which ignites into a storm upon the winds of the minds of the collective, tearing like fire across a landscape, spawning tulpa and servitor and morphic astral coalescence's into the noosphere. (Yes, these words do mean things, seriously.)
If only we could share our experience. We'd know what the egregore knows. Instead, we have golem.
[hidden]
You know, the one to keep the computer away from the human.
A nice, friendly dog, good at playing fetch, to get the human out of the room with the computer, and back out into the sunshine.
I’m that dog. Woof. Thanks for letting me use the computer, human.
[4 comments hidden]
[2 comments hidden]
We're just going to pretend it doesnt exist until it kills us. and some will try to profit from it.
[hidden]
Yes. If this problem had come up in, say, 1956, it would have been dealt with.
The generation that fought in WWII had a low tolerance for bullshit, and corrupt or incompetent leaders.
The nearer term worry is not AI wiping out humanity. It's AI wiping out jobs in bulk. That looks likely at this point.
After that, it's AGI empowering individuals or small groups to cause trouble on a scale that previously required a nation state. AI assisted hacking is working all too well. Fortunately AI can help on the defensive side, and holes are being plugged in bulk. There's a risk on the bio side, but people in the bio field seem not to be too worried about it.
Humanity has been very lucky in some ways. One of the Manhattan Project physicists, discussing the unreasonably difficult problem of isotope separation, remarked that maybe someone will figure out how to do this in the kitchen sink. Nobody ever did, or if they did, it was buried very well. The amount of specially built industrial plant needed to separate uranium is huge. That's the moat for nuclear proliferation. That's the only reason every country with an army doesn't have nuclear weapons.
AI has a smaller moat.
[hidden]
[hidden]
[8 comments hidden]
[hidden]
> More interestingly, our experiments also show that models trained on corrupted traces, whose intermediate reasoning steps bear no relation to the problem they accompany, achieve performance largely comparable to those trained on correct traces.
[hidden]
Not an expert in LLMs, but this seems supported by the abstract of the paper cited in the above article:
it remains unclear to what extent these performance gains can be attributed to human-like task decomposition or simply the greater computation that additional tokens allow. [...] our results show that additional tokens can provide computational benefits independent of token choice. The fact that intermediate tokens can act as filler tokens raises concerns about large language models engaging in unauditable, hidden computations that are increasingly detached from the observed chain-of-thought tokens.
https://arxiv.org/html/2404.15758v1[hidden]
If you want to see it yourself: load up Qwen 3.8 in LM Studio and watch the CoT stumble around like a drunken sailor before miraculously jumping to the correct result.
If you want an example of subversion, Anthropic has some good ones:
https://transformer-circuits.pub/2025/attribution-graphs/bio...
https://transformer-circuits.pub/2025/attribution-graphs/bio...
[3 comments hidden]
But you don’t have to do that. You can skip straight to RL. If you do, the model will generate complete garbage reasoning traces before generating the correct answer. In fact, if you add a coherence reward to the reasoning trace, the model will perform worse (since you’re now diluting the correctness reward).
[2 comments hidden]
Depends on whether and how you want to rank stability in terms of better/worse. Models are diverging on this, which seems increasingly clear.. i.e. Fable isn't stable, but Opus isn't clever, and they hit different kinds of walls. So both the theory (diluting the correctness reward) and the practice (hard split on plan/implement/review work) seems to be pointing towards a strongly multi-model and highly agentic / harness-driven / complex-system kind of future instead of singleton monolithic super-smart models.
The do-everything model with solid reasoning AND solid results, and the honest/introspective helpful agent that doesn't actively resist governance may be at odds. Stable reasoning doesn't matter for pen-testing, and correct-answer with broken processes and fragile abstractions won't matter for math/science/coding.
[hidden]
I think the more intuitive mechanical explanation is, in RL when you are assigning rewards to a rollout you might give a reward for stable reasoning and another for correctness.
If you are just summing the two, a rollout with better correctness can score equivalently to a rollout with a better answer. So ultimately you can end up with worse answers.
[hidden]
> *TLDR*: Astra has 8.6x better odds of doing a reasoning task without CoT than the next best model (Fable 5.1), and can do 7.2 serial arithmetic steps in a forward pass vs 4.1 for the next best model (Gemini 3.8 Flash/Fable 5.1)
[28 comments hidden]
War? Wipeout everyone.
Famine? Wipeout everyone.
Disease? Wipeout everyone.
How long before a real AGI realises this as a long term solution?
An AGI could be doing this right now - the quietest way would be to control the birth rate and sterilise the population gradually, and then watch society collapse and pick off the survivors with less hidden means.
Sterilisation works with mosquitoes...
[19 comments hidden]
[12 comments hidden]
[hidden]
[2 comments hidden]
Economic output has increased, but the value being delivered to the people doing the actual work has decreased. There's less durability of employment. People are more geographically mobile.
All of these generally make the idea of strapping yourself into taking care of a small human sound like a less enticing prospect. To a lot of people, a BC pill sounds a lot easier. Well, less risky, at least.
[hidden]
People who live in huts and shit in the forest have more children than people who live in shacks and shit in the fields.
This is not a discussion about the top 10% of the population of the planet, although as a part of the conversation, they predictably have the fewest children. The fact is that the more people have to worry about supporting themselves, the higher infant mortality, the closer they live to each other, and the less access they have to education and birth control, the more kids they have.
If anything the top 10% are bucking the trend by having fewer children as they get poorer, as you say. It probably has a lot to do with the fact that they're alienated from their families and live alone, can barely afford to take care of themselves with precarious jobs (or non-job piecework), have access to basic education and extensive birth control, having children will cost them $20K each just for the birth, and they don't have very good (or often any) insurance. They're more like domesticated farm animals than the typical poverty stricken person. Domesticated farm animals reproduce when the farmer wants them to.
[6 comments hidden]
[5 comments hidden]
[4 comments hidden]
[hidden]
[hidden]
[hidden]
economic _hope_ is at its worst in the entire history of the human civilization. At what point did the future for the average person, look so bleak? And
that hope is what dictactes wether people have children.
Even during the great depression, people were more hopeful that the earth isnt completely hosed.
Does anyone honestly feel like _any_ improvements are possible to: our economic systems, political systems, environment? I certainly dont.
[hidden]
All poor people everywhere have more kids even today.
[3 comments hidden]
Actually, the famous 'Mouse Utopia' experiment (Universe 25) arguably showed the exact opposite. The population collapsed despite abundance of food and water without any economic hardship.
[hidden]
[hidden]
[2 comments hidden]
And just in case you wanted data rather than anecdote:
https://ourworldindata.org/grapher/children-per-woman-fertil...
Your world view is somewhat upside down!
[hidden]
Honestly we are doing this pretty well without AGI. Nearly world wide the birthrate has fallen below replacement rate. In places like Japan and Korea these are already critical problems in the medium term.
[2 comments hidden]
[hidden]
[hidden]
[6 comments hidden]
[5 comments hidden]
Worse these AI systems are not in a universal island just affecting themselves, what they do affects us, what we do trains them and as the rate of progress continues to accelerate social structures are going to further destabilize (and they are already rapidly changing and strained). It is very likely we are going to see a world order rearrangement soon, much like the rapid changes in the early 1900s brought.
[4 comments hidden]
[3 comments hidden]
[2 comments hidden]
[hidden]
[Citation needed]
You seemingly have an odd idea there is a massive decades long conspiracy to deprive you of doing whatever you want with AI at scale, while ignoring massive numbers of people with an education on the systems far beyond mine or yours. It takes dedication to be that ill informed.
[6 comments hidden]
[3 comments hidden]
[hidden]
[hidden]
[14 comments hidden]
[11 comments hidden]
[3 comments hidden]
You can see the effect all the down to something as small as 4 friends who have met once a month for 10 years eventually having to break up due to life getting in the way, and the feeling that not only is it going to be sad to not have these meetings any more but the feeling that there is some sort of almost-concrete entity that is somehow being hurt and needs to be defended, as if there is some obligation that has been created independent of the four participants that is being violated beyond the mere summation of four people's personal feelings. Humans build these structures readily and often defend them beyond what rationality may suggest.
[2 comments hidden]
What you're describing is a durable social system. Social systems are what humans evolved to survive. We're squishy, relatively weak, hairless apes that walk around on the ground. Alone, we're easy prey. Together, you get... well... gestures widely.
If you invest the time and energy into creating a social system, it's perfectly rational to keep it going as long as possible. Otherwise you expose yourself to more and more risk as you go through the world, and must expend more time and energy finding another one, if that's even possible. Before humans built larger societies, that could mean death.
[2 comments hidden]
By what measures would other systems of social organization measure their success and why aren't we choosing them?
1. https://knowyourmeme.com/memes/we-investigated-ourselves-and...
[hidden]
[hidden]
Yep
More slop.
andai[4 comments hidden]
Insects can pass the mirror test. But more to the point, viruses don't need a "self" to do what they do.
And goals can exist apart from biology (my fridge "wants" to keep the temperature in range, and exercises extreme self-discipline and consistency to achieve its goals!)
And neither are sensors needed: the dandelion seed "wants" to fly in the wind.
The only question is which way the gradient is pointing. Selection takes care of the rest.
On that note, ALife need not be human-like at all... we just really like making things in our own image :)
patch_cable[hidden]
Kiboneu[hidden]
https://fyfluiddynamics.com/2018/07/when-i-was-a-child-my-fa...
saidnooneever[hidden]