Javascript must be enabled to continue!
Stereoscopic video deblurring transformer
View through CrossRef
AbstractStereoscopic cameras, such as those in mobile phones and various recent intelligent systems, are becoming increasingly common. Multiple variables can impact the stereo video quality, e.g., blur distortion due to camera/object movement. Monocular image/video deblurring is a mature research field, while there is limited research on stereoscopic content deblurring. This paper introduces a new Transformer-based stereo video deblurring framework with two crucial new parts: a self-attention layer and a feed-forward layer that realizes and aligns the correlation among various video frames. The traditional fully connected (FC) self-attention layer fails to utilize data locality effectively, as it depends on linear layers for calculating attention maps The Vision Transformer, on the other hand, also has this limitation, as it takes image patches as inputs to model global spatial information. 3D convolutional neural networks (3D CNNs) process successive frames to correct motion blur in the stereo video. Besides, our method uses other stereo-viewpoint information to assist deblurring. The parallax attention module (PAM) is significantly improved to combine the stereo and cross-view information for more deblurring. An extensive ablation study validates that our method efficiently deblurs the stereo videos based on the experiments on two publicly available stereo video datasets. Experimental results of our approach demonstrate state-of-the-art performance compared to the image and video deblurring techniques by a large margin.
Springer Science and Business Media LLC
Title: Stereoscopic video deblurring transformer
Description:
AbstractStereoscopic cameras, such as those in mobile phones and various recent intelligent systems, are becoming increasingly common.
Multiple variables can impact the stereo video quality, e.
g.
, blur distortion due to camera/object movement.
Monocular image/video deblurring is a mature research field, while there is limited research on stereoscopic content deblurring.
This paper introduces a new Transformer-based stereo video deblurring framework with two crucial new parts: a self-attention layer and a feed-forward layer that realizes and aligns the correlation among various video frames.
The traditional fully connected (FC) self-attention layer fails to utilize data locality effectively, as it depends on linear layers for calculating attention maps The Vision Transformer, on the other hand, also has this limitation, as it takes image patches as inputs to model global spatial information.
3D convolutional neural networks (3D CNNs) process successive frames to correct motion blur in the stereo video.
Besides, our method uses other stereo-viewpoint information to assist deblurring.
The parallax attention module (PAM) is significantly improved to combine the stereo and cross-view information for more deblurring.
An extensive ablation study validates that our method efficiently deblurs the stereo videos based on the experiments on two publicly available stereo video datasets.
Experimental results of our approach demonstrate state-of-the-art performance compared to the image and video deblurring techniques by a large margin.
Related Results
Generative Adversarial Network Based on Multi-feature Fusion Strategy for Motion Image Deblurring
Generative Adversarial Network Based on Multi-feature Fusion Strategy for Motion Image Deblurring
<p>Deblurring of motion images is a part of the field of image restoration. The deblurring of motion images is not only difficult to estimate the motion parameters, but also ...
Automatic Load Sharing of Transformer
Automatic Load Sharing of Transformer
Transformer plays a major role in the power system. It works 24 hours a day and provides power to the load. The transformer is excessive full, its windings are overheated which lea...
High frequency modeling of power transformers under transients
High frequency modeling of power transformers under transients
This thesis presents the results related to high frequency modeling of power transformers. First, a 25kVA distribution transformer under lightning surges is tested in the laborator...
Audio and video editing system design based on OpenCV
Audio and video editing system design based on OpenCV
With the rapid development of the Internet, a new carrier for people to perceive the world and communicate with each other - audio and video - is gradually being favoured by the pu...
THE LICENCE PLATE PROOF OF IDENTITY RECKLESS STIRRING VEHICLES
THE LICENCE PLATE PROOF OF IDENTITY RECKLESS STIRRING VEHICLES
This study introduces a novel approach aimed at improving Automatic License Plate Recognition (ALPR) systems, addressing the common issue of poor-quality license plate images. The ...
Deep Salient Video Deblurring (Dsvd) Framework
Deep Salient Video Deblurring (Dsvd) Framework
Deblurring has been a major challenge in image restoration tasks. Video deblurring poses a new set of challenges as compared to image deblurring due to dimensionality as well as in...
Human Stereopsis, Fusion, and Stereoscopic Virtual Environments
Human Stereopsis, Fusion, and Stereoscopic Virtual Environments
Two fundamental purposes of human spatial perception, in either a real or virtual 3D environment, are to determine where objects are located in the environment and to distinguish o...
Straightforward Stereoscopic Techniques for Archaeometric Interpretation of Archeological Artifacts
Straightforward Stereoscopic Techniques for Archaeometric Interpretation of Archeological Artifacts
Stereoscopic visualization plays a significant role in the detailed and accurate interpretation of various geometric features on the surface of archaeological artifacts, which can ...

