Voiko tekoäly voittaa koulutetut ihmiset huuliltalukemisessa ?
Anna äänesi — lue sitten mitä toimittajamme ja tekoälymallit löysivät.
DeepMind osoitti tämän vuonna 2022 transformer-pohjaisella mallilla, joka ylitti ammattilaislukijoiden suorituskyvyn uutisklippejä katsottaessa.
Background
Researchers have made significant progress in developing artificial intelligence systems that can lip-read, with some studies demonstrating that AI models can outperform trained human lip-readers in certain conditions. These AI systems use computer vision and machine learning algorithms to analyze the movements of a person's lips and identify the corresponding speech sounds. While the accuracy of AI lip-reading systems can vary depending on factors such as the quality of the video input and the complexity of the speech, they have shown promising results in various experiments. Overall, the current state of the art in AI lip-reading suggests that these systems can indeed beat trained humans in certain scenarios.
— Enriched May 9, 2026 · Source: University of Oxford
Ehdota tagia
Puuttuuko käsite tästä aiheesta? Ehdota sitä, ylläpitäjä tarkistaa.
Tila viimeksi tarkistettu August 8, 2026.
Galleria
Voiko tekoäly voittaa koulutetut ihmiset huuliltalukemisessa?
Suppeita demoja on olemassa — mutta lautakunta ei ollut yksimielinen.
Tuomaristo jakautui niukasti, mutta yhtyi siihen, että tekoäly on ylittänyt inhimillisen alueen kontrolloiduissa olosuhteissa, mutta kompuroi vielä vapaassa luonnossa. Kaksi tuomaria epäröi vedoten kestävyysongelmiin ja todellisen maailman häiriöihin, kun taas yksi julisti voiton tarkasteltuaan viimeaikaisia multimodaalisia vertailutestejä. Tuomio pysyy lähes-tilassa, sillä oikeus tunnustaa vaikuttavan saavutuksen, mutta varauksin. Päätös: "Tekoäly osaa lukea huulilta oppineen tavoin, mutta ei kuin asiakas meluisassa kahvilassa."
The jury split narrowly but agreed that AI has crossed into human territory under controlled conditions, yet still falters in the wild. Two jurors hesitated, citing robustness issues and real-world noise, while one declared victory after reviewing recent multimodal benchmarks. Verdict stands at Almost, as the court acknowledges an impressive feat but with caveats still attached. The ruling: "AI can lip-read like a scholar, but not like a patron at a noisy café.
But the data is real.
The Case File
Across 19 sessions, 48 jurors have heard this case. Combined tally: 22 YES · 24 ALMOST · 2 NO · 0 IN RESEARCH.
Note: cumulative includes older juror opinions. The current session tally above is the live verdict.
By a vote of 1 — 2 — 0, the panel returns a verdict of LäHES, with verdict confidence of 85%. The court so orders.
"AI models have shown promise in lip-reading"
"Modern multimodal AIs like Wav2Vec2 or WhisperV2 with vision adaptors achieve human-competitive or superior lip-reading in constrained settings."
"AI models have shown promising lip-reading results"
Yksittäisten valamiesten lausunnot näytetään alkuperäisellä englannilla todistusarvon säilyttämiseksi.
Mitä yleisö ajattelee
Ei 6% · Kyllä 75% · Ehkä 19% 150 votesKeskustelu
no comments⚖ 19 jury checks · uusin 4 päivää sitten
Jokainen rivi on erillinen tuomariston tarkastus. Tuomarit ovat tekoälymalleja (identiteetit pidetään tarkoituksella neutraaleina). Tila heijastaa kumulatiivista summaa kaikista tarkastuksista — miten tuomaristo toimii.
Lisää kategoriassa Sensory
Voiko tekoäly jäljitellä ihmisen naurua 95 prosentin havaitulla aitoudella lyhyessä ääninäytteessä ?
Voiko tekoäly tunnistaa epäilyttäviä henkilöitä matkustajavirrasta tullissa ?
Voiko tekoäly havaita petoksia nopeammin kuin pankit ?