Home /Research /Style-transfer based Speech and Audio-visual Scene understanding for Robot Action Sequence Acquisition from Videos
OTHER

Style-transfer based Speech and Audio-visual Scene understanding for Robot Action Sequence Acquisition from Videos

Chiori Hori, Puyuan Peng, David Harwath, Xinyu Liu, Kei Ota, Siddarth Jain, Radu Corcodel, Devesh K. Jha, Diego Romeres, Jonathan Le Roux

Year
2023
Citations
5

Keywords

Computer scienceAudio visualAction (physics)RobotSpeech recognitionStyle (visual arts)Artificial intelligenceSequence (biology)Computer visionHuman–computer interaction

Related papers

Browse all OTHER papers