AimigoTutor - tutoring application using multi-modal capabilities
Video captioning has been an up-and-coming research topic. Thanks to the recent advances in the performance of deep neural networks, especially with transformers, video captioning is seeing a huge potential improvement in accuracy and versatility. Most state-of-the-art video captioning models employ...
Main Author: | Nguyen, Viet Hoang |
---|---|
Other Authors: | Hanwang Zhang |
Format: | Final Year Project (FYP) |
Language: | English |
Published: |
Nanyang Technological University
2024
|
Subjects: | |
Online Access: | https://hdl.handle.net/10356/175732 |
Similar Items
-
Online multi-face tracking with multi-modality cascaded matching
by: Weng, Zhenyu, et al.
Published: (2024) -
Crowdsourcing platform for private tutoring (Tutor Society)
by: Zaini, Farah Nadia
Published: (2016) -
Q-instruct: improving low-level visual abilities for multi-modality foundation models
by: Wu, Haoning, et al.
Published: (2024) -
Online tutor management system using mobile application (EziTutor)
by: Muhammad Khairil Akmal, Mohd Khairuddin
Published: (2018) -
Website for tutoring agency
by: Sim, Nicole Kian Min
Published: (2019)