AimigoTutor - tutoring application using multi-modal capabilities

Video captioning has been an up-and-coming research topic. Thanks to the recent advances in the performance of deep neural networks, especially with transformers, video captioning is seeing a huge potential improvement in accuracy and versatility. Most state-of-the-art video captioning models employ...

Full description

Saved in:

Bibliographic Details
Main Author:	Nguyen, Viet Hoang
Other Authors:	Hanwang Zhang
Format:	Final Year Project
Language:	English
Published:	Nanyang Technological University 2024
Subjects:	Computer and Information Science Multi-modal
Online Access:	https://hdl.handle.net/10356/175732
Tags:	Add Tag No Tags, Be the first to tag this record!
Institution:	Nanyang Technological University
Language:	English

Be the first to leave a comment!

AimigoTutor - tutoring application using multi-modal capabilities

Similar Items