3 papers
cs.CL2025
Large Language Models and Provenance Metadata for Determining the Relevance of Images and Videos in News Stories
Tomas Peterka, Matyas Bohacek
The most effective misinformation campaigns are multimodal, often combining text with images and videos taken out of context -- or fabricating them entirely -- to support a given n…
cs.CV2025
Can Pose Transfer Models Generate Realistic Human Motion?
Vaclav Knapp, Matyas Bohacek
Recent pose-transfer methods aim to generate temporally consistent and fully controllable videos of human action where the motion from a reference video is reenacted by a new ident…
cs.CV2022
Combining Efficient and Precise Sign Language Recognition: Good pose estimation library is all you need
Matyáš Boháček, Zhuo Cao, Marek Hrúz
Sign language recognition could significantly improve the user experience for d/Deaf people with the general consumer technology, such as IoT devices or videoconferencing. However,…