arXiv · 2406.17916
Camera Model Identification Using Audio and Visual Content from Videos
Abstract
The identification of device brands and models plays a pivotal role in the realm of multimedia forensic applications. This paper presents a framework capable of identifying devices using audio, visual content, or a fusion of them. The fusion of visual and audio content occurs later by applying two fundamental fusion rules: the product and the sum. The device identification problem is tackled as a classification one by leveraging Convolutional Neural Networks. Experimental evaluation illustrates that the proposed framework exhibits promising classification performance when independently using audio or visual content. Furthermore, although the fusion results don't consistently surpass both individual modalities, they demonstrate promising potential for enhancing classification performance. Future research could refine the fusion process to improve classification performance in both modalities consistently. Finally, a statistical significance test is performed for a more in-depth study of the classification results.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Ioannis Tsingalis, Christos Korgialas, Constantine Kotropoulos. 2024-06-25. Camera Model Identification Using Audio and Visual Content from Videos. https://arxiv.org/abs/2406.17916
Cite the original work for its findings. Save a collection to share your selection of sources.