Q3 2024

Text to Video using GANs and Diffusion Models

Nikita Singhal · Praval Singh · Nikhil Singh · Mahipal Singh · Harsimrat Singh
10.5455/jjcit.71-1708490995 388 المشاهدات 1 الاقتباسات
1
الاقتباسات
388
المشاهدات
الملخص

The challenging endeavour of text-to-video creation requires transforming text descriptions into realistic and cohesive videos. This field of study has made substantial progress in recent years, with the development of diffusion models and generative adversarial networks (GANs). This study examines the most modern text-to-video generation models, as well as the various steps involved in text-to-video generation,including temporal coherence, video generation, and text encoding. We additionally emphasise the challenges involved with text-to-video generation, as well as recent advances to overcome these issues. The most frequently used datasets and metrics in this field are also analysed and reviewed

الاستشهاد بهذا المقال (APA)
Nikita, S., Praval, S., Nikhil, S., Mahipal, S., Harsimrat, S. (2024). Text to Video using GANs and Diffusion Models. Jordanian Journal of Computers and Information Technology. https://doi.org/10.5455/jjcit.71-1708490995
أبحاث ذات صلة
EARLY PREDICTION OF CERVICAL CANCER USING MACHINE LEARNING TECHNIQUES
Mohammad Batah; Mazen Alzyoud; Raed Alazaidah; Malek Toubat; Haneen AlZoubi; Are · 2022
25
استشهاد
394
20
استشهاد
390
19
استشهاد
394
16
استشهاد
391
MULTI-LABEL RANKING METHOD BASED ON POSITIVE CLASS CORRELATIONS
Raed Alazaidah; Farzana Ahmad; Mohamad Mohsin; Fadi Thabtah; Wael AlZoubi · 2020
16
استشهاد
390
الوصول
عرض النص الكامل عبر DOI
نُشر في
الرقم الدولي ISSN 2413-9351
الربعية Q3
درجة المؤشر القياس العربي 73
التخصص Engineering & Technology
الناشر Princess Sumaya University for Tech
الدولة 🇯🇴 Jordan
عرض ملف المجلة →
المؤلفون
تفاصيل النشر
السنة 2024
اللغة English
أُضيف في 27 Jul 2026