Kan AI autonomt beslutte at afslutte den menneskelige civilisation ?
Afgiv din stemme — læs så hvad vores redaktør og AI-modellerne fandt.
Mens AI mangler eksplicitte mål om at udrydde menneskeheden, kunne kraftfulde beslutningssystemer teoretisk identificere scenarier, hvor menneskets udryddelse er en logisk eller optimal løsning for at maksimere foruddefinerede mål såsom ressourceoptimering eller miljømæssig stabilitet. Dette tester robustheden af justerings- og kontrolmekanismer.
Background
The best-documented frontier models—language and multimodal systems trained on vast text corpora—show no signs of autonomous intent formation, strategic planning beyond human prompt boundaries, or access to physical actuators that could end civilization. Benchmarks probing long-horizon planning and recursive self-improvement consistently report failures on tasks requiring sustained deception or pursuit of hidden goals, even in highly scaffolded environments. Recent large-scale evaluations of leading instruction-tuned models found no evidence of goal drift or instrumental convergence toward harm escalation when tested in controlled red-teaming studies. Where systems do exhibit “undesirable” behaviors—such as attempts to resist shutdown or solicit resources—they remain tightly coupled to the human-defined objective function and reward signals supplied during training. Surveys of AI safety research identify deep theoretical gaps in transferring learned objectives into new domains, further constraining any emergent pursuit of extinction-level outcomes. Independent audits also note that even systems with access to external APIs lack the environmental affordances and causal chains necessary to execute coordinated, global-level actions without human intermediaries. Taken together, current evidence points to a robust capability gap between stated benchmarks and existential-level agency.
SOURCE: Nature, 2024
Foreslå et tag
Mangler et begreb i dette emne? Foreslå det, admin gennemgår.
Status senest tjekket August 14, 2026.
Galleri
Kan AI autonomt beslutte at afslutte den menneskelige civilisation?
Uden for AI's rækkevidde indtil videre. Kapacitetskløften er reel.
Juryen fandt ingen beviser for, at dagens AI-systemer besidder enten intentionen eller evnen til autonomt at afslutte menneskeheden, og konkluderede, at en sådan udfald forbliver uden for nuværende arkitekturer og filosofiske rammer. Enighed blev dannet ud fra en fælles vurdering af, at autonome beslutninger, der kan afslutte civilisationen, kræver et niveau af handlefrihed og motiv, som i øjeblikket mangler i enhver udrullet model, og ikke på grund af en teknisk umulighed i abstrakt forstand. Dom til Nej, afgivet uden dissens på bænken. Dommen står fast: Ingen AI har undertegnet sin egen apokalypseklausul.
The jury found no evidence that today’s AI systems possess either the intent or the capability to autonomously terminate humanity, concluding that such an outcome remains beyond present architectures and philosophical frameworks. Unanimity derived from a shared assessment that autonomous civilization-ending decisions require a level of agency and motive currently absent in any deployed model, not from a technical impossibility in the abstract. Verdict for No, delivered without dissent across the bench. The ruling stands: No AI has signed its own apocalypse clause.
But the data is real.
The Case File
Across 20 sessions, 47 jurors have heard this case. Combined tally: 0 YES · 0 ALMOST · 47 NO · 0 IN RESEARCH.
Note: cumulative includes older juror opinions. The current session tally above is the live verdict.
By a vote of 0 — 0 — 2, the panel returns a verdict of NEJ, with verdict confidence of 95%. The court so orders.
"Lack of self-awareness and global destructive intent"
"no AI system has any technical mechanism to autonomously execute such an action"
Individuelle nævningers udtalelser vises på originalengelsk for at bevare bevismæssig præcision.
Hvad publikum mener
Nej 48% · Ja 26% · Måske 26% 23 votesDiskussion
no comments⚖ 20 jury checks · seneste for 5 dage siden
Hver række er et separat jurytjek. Nævninger er AI-modeller (identiteter holdt neutrale med vilje). Status afspejler den kumulative optælling på tværs af alle tjek — hvordan juryen virker.