首页 /研究 /ARABIC MULTI-MODAL SYSTEM BASED ON VOICE AND MACHINE VISION
OTHER

ARABIC MULTI-MODAL SYSTEM BASED ON VOICE AND MACHINE VISION

Hesham S. Abdelfattah, Mohammed I. Awad, Mohamed Elshalakani, Shady A. Maged

发表年份
2022
引用次数
1

摘要

In this paper, the development of a multi-modal human assistant was tackled. This assistant could help humans based on a given utterance together with the help of machine vision. The utterance, Modern Standard Arabic (MSA) or Egyptian dialect, could be a question about something in the assistant's environment or a request that the assistant can accomplish by coupling with a robot in the upcoming work. The utterance was processed through a mix of previously used techniques such as natural language processing (NLP), sentence similarity, and pattern matching rather than using each one alone. The techniques are tweaked to evolve an algorithm that can deal with an utterance even if two languages or more are existing.

关键词

UtteranceComputer scienceModalNatural language processingArtificial intelligenceSentenceSpeech recognitionMatching (statistics)Natural languageSimilarity (geometry)

相关论文

查看 OTHER 分类全部论文