SLVideo: A Sign Language Video Moment Retrieval Framework

  • 2024-07-22 15:29:36
  • Gonçalo Vinagre Martins, Afonso Quinaz, Carla Viegas, Sofia Cavaco, João Magalhães
  • 0

Abstract

Sign Language Recognition has been studied and developed throughout the yearsto help the deaf and hard-of-hearing people in their day-to-day lives. Thesetechnologies leverage manual sign recognition algorithms, however, most of themlack the recognition of facial expressions, which are also an essential part ofSign Language as they allow the speaker to add expressiveness to their dialogueor even change the meaning of certain manual signs. SLVideo is a video momentretrieval software for Sign Language videos with a focus on both hands andfacial signs. The system extracts embedding representations for the hand andface signs from video frames to capture the language signs in full. This willthen allow the user to search for a specific sign language video segment withtext queries, or to search by similar sign language videos. To test thissystem, a collection of five hours of annotated Sign Language videos is used asthe dataset, and the initial results are promising in a zero-shotsetting.SLVideo is shown to not only address the problem of searching signlanguage videos but also supports a Sign Language thesaurus with a search bysimilarity technique. Project web page: https://novasearch.github.io/SLVideo/

 

Quick Read (beta)

loading the full paper ...