All systems operational
Q2 2025

Neural Volumetric Representations for Real-Time 3D Scene Reconstruction using Multi-Modal Learning Algorithm

Pidatala Babu · Preeti Jha
10.34028/iajit/22/6/12 384 Views 0 Citations
0
Citations
384
Views
Abstract

Deep Learning (DL) is a subfield of Machine Learning (ML) models used in various complex fields. DL algorithms are mostly widely used to reconstruct 3D images collected from multiple online sources. It is a very challenging task for the existing algorithms to reconstruct 2D images into 3D pictures without losing high-quality pixels because of the complex scenes with different lighting situations, dynamic components, and occlusions. This paper presents a novel real-time 3D scene reconstruction using neural volumetric representations combined with a Multi-Modal Learning Algorithm (MMLA). The proposed MMLA focuses on solving issues like volumetric representations of scenes, which are improved by combining numerous modalities such as RGB images, depth sensors, and Inertial Measurement Unit (IMU) data. The MMLA combines the DeepVoxels model and Neural Radiance Fields (NeRF) model, which it calls the Neural Rendering technique, to learn complex patterns in 3D scenes. The pre-trained model EfficientNet accurately obtained the 3D- reconstruction patterns and understood the spatial structures that transfer to the proposed MMLA. The proposed MMLA performance is analyzed using the ShapeNet dataset, which consists of 2D images. Finally, the experimental results show that the proposed MMLA outperforms the superior performance in terms of Mean Squared Error (MSE) of 0.167, Root Mean Squared Error (RMSE) of 0.50, and Mean Absolute Error (MAE) of 1.1. These results may differ from other datasets

Cite this Article (APA)
Pidatala, B., Preeti, J. (2025). Neural Volumetric Representations for Real-Time 3D Scene Reconstruction using Multi-Modal Learning Algorithm. The International Arab Journal of Information Technology. https://doi.org/10.34028/iajit/22/6/12
Related Papers
Perception of Natural Scenes: Objects Detection and Segmentations using Saliency Map with AlexNet
Muhammad Waqas Ahmed; Abdulwahab Alazeb; Naif Al Mudawi; Touseef Sadiq; Bayan Al · 2025
21
cites
390
Agile Proactive Cybercrime Evidence Analysis Model for Digital Forensics
Mohammad Al-Mousa; Waleed Amer; Mosleh Abualhaj; Sultan Albilasi; Ola Nasir; Gha · 2025
16
cites
387
14
cites
396
Heart Disease Diagnosis Using Decision Trees with Feature Selection Method
Alaa Sheta; Walaa El-Ashmawi; Abdelkarim Baareh · 2024
13
cites
383
Access
View Full Text via DOI
Published in
ISSN 1683-3198
Quartile Q2
AMS Score 100
Field Computer Science & AI
Publisher Zarqa University / Colleges of Comp
Country 🇯🇴 Jordan
View Journal Profile →
Authors
Publication Details
Year 2025
Language English
Added 30 Jul 2026