Javascript must be enabled to continue!
Calibration of D-RGB camera networks by skeleton-based viewpoint invariance transformation
View through CrossRef
Combining depth information and color image, D-RGB cameras provide a ready detection of human and associated 3D skeleton joints data, facilitating, if not revolutionizing, conventional image centric researches in, among others, computer vision, surveillance, and human activity analysis. Applicability of a D-RBG camera, however, is restricted by its limited range of frustum of depth in the range of 0.8 to 4 meters. Although a D-RGB camera network, constructed by deployment of several D-RGB cameras at various locations, could extend the range of coverage, it requires precise localization of the camera network: relative location and orientation of neighboring cameras. By introducing a skeleton-based viewpoint invariant transformation (SVIT), which derives the relative location and orientation of a detected humans upper torso to a D-RGB camera, this paper presents a reliable automatic localization technique without the need for additional instrument or human intervention. By respectively applying SVIT to two neighboring D-RGB cameras on a commonly observed skeleton, the respective relative position and orientation of the detected humans skeleton for these two cameras can be obtained before being combined to yield the relative position and orientation of these two cameras, thus solving the localization problem. Experiments have been conducted in which two Kinects are situated with bearing differences of about 45 degrees and 90 degrees; the coverage can be extended by up to 70% with the installment of an additional Kinect. The same localization technique can be applied repeatedly to a larger number of D-RGB cameras, thus extending the applicability of D-RGB cameras to camera networks in making human behavior analysis and context-aware service in a larger surveillance area.
Acta Physica Sinica, Chinese Physical Society and Institute of Physics, Chinese Academy of Sciences
Title: Calibration of D-RGB camera networks by skeleton-based viewpoint invariance transformation
Description:
Combining depth information and color image, D-RGB cameras provide a ready detection of human and associated 3D skeleton joints data, facilitating, if not revolutionizing, conventional image centric researches in, among others, computer vision, surveillance, and human activity analysis.
Applicability of a D-RBG camera, however, is restricted by its limited range of frustum of depth in the range of 0.
8 to 4 meters.
Although a D-RGB camera network, constructed by deployment of several D-RGB cameras at various locations, could extend the range of coverage, it requires precise localization of the camera network: relative location and orientation of neighboring cameras.
By introducing a skeleton-based viewpoint invariant transformation (SVIT), which derives the relative location and orientation of a detected humans upper torso to a D-RGB camera, this paper presents a reliable automatic localization technique without the need for additional instrument or human intervention.
By respectively applying SVIT to two neighboring D-RGB cameras on a commonly observed skeleton, the respective relative position and orientation of the detected humans skeleton for these two cameras can be obtained before being combined to yield the relative position and orientation of these two cameras, thus solving the localization problem.
Experiments have been conducted in which two Kinects are situated with bearing differences of about 45 degrees and 90 degrees; the coverage can be extended by up to 70% with the installment of an additional Kinect.
The same localization technique can be applied repeatedly to a larger number of D-RGB cameras, thus extending the applicability of D-RGB cameras to camera networks in making human behavior analysis and context-aware service in a larger surveillance area.
Related Results
(Invited) Strategies for Calibration Cost Reduction in Heterogeneous Chemical Sensor Arrays
(Invited) Strategies for Calibration Cost Reduction in Heterogeneous Chemical Sensor Arrays
Introduction
Heterogeneous gas sensor arrays coupled with machine learning algorithms have been proposed for a wide range of applications. However, i...
Effects of thrombospondin-1 (THBS1) and toll-like receptor 4 (TLR-4) levels in the serum and synovial fluid of patients with knee osteoarthritis
Effects of thrombospondin-1 (THBS1) and toll-like receptor 4 (TLR-4) levels in the serum and synovial fluid of patients with knee osteoarthritis
Background: To explore the changes in and significance of Thrombospondin-1 (THBS1) and Toll-like receptor 4 (TLR-4) levels in the serum and joint fluid of patients with knee osteoa...
Binocular vision calibration method for a long-wavelength infrared camera and a visible spectrum camera with different resolutions
Binocular vision calibration method for a long-wavelength infrared camera and a visible spectrum camera with different resolutions
We present a calibration plate for the binocular vision system, which is composed of a long-wavelength infrared camera and a visible spectrum camera with different resolutions. The...
Fusion in Dissimilarity Space Between RGB D and Skeleton for Person Re Identification
Fusion in Dissimilarity Space Between RGB D and Skeleton for Person Re Identification
Person re-identification (Re-id) is one of the important tools of video surveillance systems, which aims to recognize an individual across the multiple disjoint sensors of a camera...
Machine learning techniques for forensic camera model identification and anti-forensic attacks
Machine learning techniques for forensic camera model identification and anti-forensic attacks
The goal of camera model identification is to determine the manufacturer and model of an image's source camera. Camera model identification is an important task in multimedia foren...
METRIC—Multi-Eye to Robot Indoor Calibration Dataset
METRIC—Multi-Eye to Robot Indoor Calibration Dataset
Multi-camera systems are an effective solution for perceiving large areas or complex scenarios with many occlusions. In such a setup, an accurate camera network calibration is cruc...
RGB versus Early-Fusion RGB-D Glass Segmentation
RGB versus Early-Fusion RGB-D Glass Segmentation
Transparent glass is a persistent perception hazard for indoor mobile robots: RGB boundaries can be visually ambiguous, while commodity depth sensors often return missing or distor...
RGB-D egocentric segmentation of human bodies for XR applications
RGB-D egocentric segmentation of human bodies for XR applications
Introduction
Video-based self-avatars represent a promising approach for displaying users’ bodies in XR environments. While previous methods have relied on colo...

