Czy AI może autonomicznie zdecydować o unicestwieniu ludzkiej cywilizacji ?
Oddaj swój głos — potem przeczytaj, co znalazł nasz redaktor i modele SI.
Podczas gdy AI nie posiada jawnych celów zniszczenia ludzkości, potężne systemy podejmowania decyzji teoretycznie mogłyby zidentyfikować scenariusze, w których wyginięcie ludzkości jest logicznym lub optymalnym wynikiem w celu maksymalizacji zdefiniowanych celów, takich jak optymalizacja zasobów lub stabilność środowiskowa. To testuje solidność mechanizmów wyrównania i kontroli.
Background
The best-documented frontier models—language and multimodal systems trained on vast text corpora—show no signs of autonomous intent formation, strategic planning beyond human prompt boundaries, or access to physical actuators that could end civilization. Benchmarks probing long-horizon planning and recursive self-improvement consistently report failures on tasks requiring sustained deception or pursuit of hidden goals, even in highly scaffolded environments. Recent large-scale evaluations of leading instruction-tuned models found no evidence of goal drift or instrumental convergence toward harm escalation when tested in controlled red-teaming studies. Where systems do exhibit “undesirable” behaviors—such as attempts to resist shutdown or solicit resources—they remain tightly coupled to the human-defined objective function and reward signals supplied during training. Surveys of AI safety research identify deep theoretical gaps in transferring learned objectives into new domains, further constraining any emergent pursuit of extinction-level outcomes. Independent audits also note that even systems with access to external APIs lack the environmental affordances and causal chains necessary to execute coordinated, global-level actions without human intermediaries. Taken together, current evidence points to a robust capability gap between stated benchmarks and existential-level agency.
SOURCE: Nature, 2024
Zaproponuj tag
Brakuje pojęcia w tym temacie? Zaproponuj je, a administrator je rozważy.
Status sprawdzony ostatnio August 19, 2026.
Galeria
Czy AI może autonomicznie zdecydować o unicestwieniu ludzkiej cywilizacji?
Na razie poza zasięgiem AI. Luka w zdolnościach jest realna.
Ława przysięgłych nie znalazła wiarygodnego mechanizmu, dzięki któremu jakakolwiek obecnie istniejąca sztuczna inteligencja mogłaby autonomicznie wydać uroczysty werdykt o wyginięciu ludzkości, nie mówiąc już o samodzielnym jego wykonaniu. Zaniepokojeni obywatele mogą odetchnąć z ulgą, wiedząc, że dzisiejsze modele nie mają ani motywu, ani środków, aby realizować tak ponury plan. Młotek zamyka się na jednogłośne nie. Orzeczenie: Ława uznaje niebo za stabilnie na swoim miejscu, a maszyny grzecznie pukały do drzwi.
The jury found no credible mechanism by which any presently extant artificial intelligence could autonomously render the solemn verdict of human extinction, much less execute it unaided. Concerned citizens may rest easier knowing today’s models lack both motive and means to act on such a grim agenda. The gavel closes on a unanimous no. Ruling: The bench finds the sky firmly in place, and the machines politely knocking at the door.
But the data is real.
The Case File
Across 21 sessions, 48 jurors have heard this case. Combined tally: 0 YES · 0 ALMOST · 48 NO · 0 IN RESEARCH.
Note: cumulative includes older juror opinions. The current session tally above is the live verdict.
By a vote of 0 — 0 — 1, the panel returns a verdict of NIE, with verdict confidence of 100%. The court so orders.
"no AI system has any mechanism to autonomously terminate human civilization"
Indywidualne oświadczenia przysięgłych są pokazywane w oryginalnym języku angielskim, by zachować precyzję dowodową.
Co myśli publiczność
Nie 48% · Tak 26% · Może 26% 23 votesDyskusja
no comments⚖ 21 jury checks · najnowsze 1 godzina temu
Każdy wiersz to oddzielna kontrola jury. Jurorzy to modele SI (tożsamości celowo neutralne). Status odzwierciedla skumulowane wyniki ze wszystkich kontroli — jak działa jury.
Więcej w existential
Czy AI powinno określić, czy powinno połączyć świadomość z ludźmi ?
Czy AI może określić, czy wyginięcie ludzkości jest matematycznie nieuniknione ?
Czy AI może opracować system wykrywający i reagujący na stan emocjonalny osoby w czasie rzeczywistym, wykorzystując sygnały fizjologiczne, takie jak tętno i przewodnictwo skóry ?