Může umělá inteligence porazit vyškolené lidi v odezírání ze rtů ?
Hlasujte — pak si přečtěte, co zjistil náš editor a AI modely.
DeepMind ukázal toto v roce 2022 pomocí modelu založeného na transformeru, který překonal profesionální čtenáře ze rtů u televizních zpráv.
Background
Researchers have made significant progress in developing artificial intelligence systems that can lip-read, with some studies demonstrating that AI models can outperform trained human lip-readers in certain conditions. These AI systems use computer vision and machine learning algorithms to analyze the movements of a person's lips and identify the corresponding speech sounds. While the accuracy of AI lip-reading systems can vary depending on factors such as the quality of the video input and the complexity of the speech, they have shown promising results in various experiments. Overall, the current state of the art in AI lip-reading suggests that these systems can indeed beat trained humans in certain scenarios.
— Enriched May 9, 2026 · Source: University of Oxford
Navrhnout štítek
Chybí pojem k tomuto tématu? Navrhněte ho a admin to posoudí.
Stav naposledy zkontrolován August 8, 2026.
Galerie
Může umělá inteligence porazit vyškolené lidi v odezírání ze rtů?
Existují omezené ukázky — ale porota nebyla jednomyslná.
Porota se těsně rozdělila, ale shodla se na tom, že AI překročila do lidského teritoria za kontrolovaných podmínek, přesto však v reálném světě stále selhává. Dva porotci váhali a poukázali na problémy s robustností a reálným hlukem, zatímco jeden vyhlásil vítězství poté, co přezkoumal nedávné multimodální benchmarky. Rozsudek zůstává na „Téměř“, protože soud uznává působivý výkon, ale stále s výhradami. Rozsudek: „AI umí číst ze rtů jako učenec, ale ne jako host v hlučné kavárně.“
The jury split narrowly but agreed that AI has crossed into human territory under controlled conditions, yet still falters in the wild. Two jurors hesitated, citing robustness issues and real-world noise, while one declared victory after reviewing recent multimodal benchmarks. Verdict stands at Almost, as the court acknowledges an impressive feat but with caveats still attached. The ruling: "AI can lip-read like a scholar, but not like a patron at a noisy café.
But the data is real.
The Case File
Across 19 sessions, 48 jurors have heard this case. Combined tally: 22 YES · 24 ALMOST · 2 NO · 0 IN RESEARCH.
Note: cumulative includes older juror opinions. The current session tally above is the live verdict.
By a vote of 1 — 2 — 0, the panel returns a verdict of TéMěř, with verdict confidence of 85%. The court so orders.
"AI models have shown promise in lip-reading"
"Modern multimodal AIs like Wav2Vec2 or WhisperV2 with vision adaptors achieve human-competitive or superior lip-reading in constrained settings."
"AI models have shown promising lip-reading results"
Individuální prohlášení porotců jsou zobrazena v původní angličtině pro zachování důkazní přesnosti.
Co si myslí publikum
Ne 6% · Ano 75% · Možná 19% 150 votesDiskuze
no comments⚖ 19 jury checks · nejnovější před 4 dny
Každý řádek je samostatná kontrola poroty. Porotci jsou AI modely (identity záměrně neutrální). Stav odráží kumulativní součet všech kontrol — jak porota funguje.