US20260187951A1 · App 19/296,131
METHOD FOR PROVIDING FULL-BODY MOTION INTERACTION USING MULTIPLE MOBILE CAMERAS AND APPARATUS THEREFOR
Publication
Application
Classifications
IPC Classifications
CPC Classifications
Applicants
ELECTRONICS AND TELECOMMUNICATIONS RESEARCH INSTITUTE
Inventors
Sung-Jin HONG, Kyung-Kyu KANG, Youn-Hee GIL, Hye-Sun KIM, Seong-Min BAEK, Cho-Rong YU
Abstract
Disclosed herein are a method for providing full-body motion interaction using multiple mobile cameras and an apparatus for the same. The method, performed by the apparatus, includes estimating a joint outside the field of view of a user camera by using self-joint detection information captured by multiple user cameras mounted on a user, searching for the camera of an additional user located in the same space as the user based on landmark information extracted from images captured by the multiple user cameras, selecting candidate joint information for the joint outside the field of view of the user camera in other-user joint detection information captured by the camera of the additional user, reconstructing full-body joints of the user by combining estimated joint information with the candidate joint information, and providing interaction for a full-body motion of the user based on the reconstructed full-body joints.
Get a summary, plain-language explanation, or ask your own question.
Figures
Description
CROSS REFERENCE TO RELATED APPLICATION
[0001]This application claims the benefit of Korean Patent Application No. 10-2024-0197422, filed Dec. 26, 2024, which is hereby incorporated by reference in its entirety into this application.
BACKGROUND OF THE INVENTION
1. Technical Field
[0002]The present disclosure relates generally to technology for providing full-body motion interaction using multiple mobile cameras, and more particularly to technology for reconstructing 3D full-body motions of a user and providing interaction using multiple mobile cameras in order to more accurately provide interaction with a virtual or real object in real/virtual environments.
2. Description of Related Art
[0003]Accurately estimating 3D joints of users is a crucial factor in providing interaction with objects to users who experience a virtual or extended reality (XR) environment. In general, a method of estimating joints using mobile sensors attached near the user's head is used, but this method has a limitation that it is difficult to accurately estimate full-body joints when the joint area to be estimated falls outside the field of view of a camera or is occluded by obstacles.
[0004]Also, many methods of using fisheye lenses have been proposed to expand the scope of interaction. However, a fisheye lens causes significant distortion as the distance from the center of the lens increases, which results in a decrease in the accuracy of joint estimation. In order to solve this problem, high-end Head Mounted Displays (HMDs) incorporate additional depth cameras, thereby providing accurate interaction.
[0005]Also, in order to improve the estimation accuracy of invisible or occluded joints or joints that are inaccurate due to distortion, a method of combining a wide-angle lens having a narrower field of view than a fisheye lens with multiple Inertial Measurement Unit (IMU) sensors has been proposed to estimate full-body joints of a user. However, this method imposes many constraints on the usage environment due to the inconvenience of wearing additional IMU sensors and severe noise in environments with a large amount of metal.
[0006]Currently, mobile XR products with fields of view of normal lenses are only used facing the user's front and provide only a limited range of hand-based interaction, and they lack support for full-body joint interaction.
Documents of Related Art
[0007](Patent Document 1) Korean Patent Application Publication No. 10-2024-0072397, published on May 24, 2024 and titled “Method and apparatus for providing user augmented reality interaction based on mobile devices”.
SUMMARY OF THE INVENTION
[0008]An object of the present disclosure is to provide full-body motion interaction by more accurately estimating 3D full-body joints using multiple RGBD cameras with standard lenses, rather than using fisheye or wide-angle lenses.
[0009]Another object of the present disclosure is to use cameras with standard lenses installed around the head of a user, thereby more accurately estimating joints even in an area that is outside the field of view of a camera or heavily occluded.
[0010]A further object of the present disclosure is to support interaction with a virtual or real object using a reconstructed 3D full-body motion in an XR environment.
[0011]In order to accomplish the above objects, a method for providing full-body motion interaction using multiple mobile cameras, performed by an apparatus for providing full-body motion interaction, according to the present disclosure includes estimating a joint outside the field of view of a user camera by using self-joint detection information captured by multiple user cameras mounted on a user, searching for a camera of an additional user located in the same space as the user based on landmark information extracted from images captured by the multiple user cameras, selecting candidate joint information for the joint outside the field of view of the user camera in other-user joint detection information captured by the camera of the additional user, reconstructing full-body joints of the user by combining estimated joint information with the candidate joint information, and providing interaction for a full-body motion of the user based on the reconstructed full-body joints.
[0012]Here, estimating the joint may include identifying body regions of the user visible in the field of view of the user camera, setting a reference point for each of the identified body regions, and estimating the direction of the joint outside the field of view of the user camera based on principal component analysis considering the reference point.
[0013]Here, the principal component analysis may comprise inferring a position of at least one adjacent joint from the reference point and setting a weight for the position of the adjacent joint.
[0014]Here, the weight may be set higher as the position is closer to the reference point, and may be set lower as the position is more distant from the reference point.
[0015]Here, the adjacent joint may correspond to a joint directly connected to a joint corresponding to the reference point in terms of body structure.
[0016]Here, searching for the camera of the additional user may include extracting the landmark information based on the images captured by the multiple user cameras, setting a global coordinate system by combining the landmark information, determining the position and orientation of the user in the global coordinate system, and detecting the camera of the additional user based on the position and orientation of the user.
[0017]Here, selecting the candidate joint information may include converting an image captured by the camera of the additional user into the global coordinate system and generating the other-user joint detection information by classifying joint information of the user in the image captured by the camera of the additional user based on the position and orientation of the user.
[0018]Here, the candidate joint information may be selected to correspond to other-user joint detection information of a camera of an additional user selected in consideration of joint detection reliability of each camera, and when multiple additional users' cameras having similar reliability are found, the candidate joint information may be selected by further considering a distance from the user in the global coordinate system.
[0019]Here, the multiple user cameras may be mounted around the head of the user and may include a front camera for capturing in a direction in front of the user, a left camera for capturing in a direction toward the ground from the left side of the user, and a right camera for capturing in the direction toward the ground from the right side of the user.
[0020]Here, the front camera, the left camera, and the right camera may operate in a unified coordinate system and correspond to RGBD cameras with standard lenses.
[0021]Also, an apparatus for providing full-body motion interaction using multiple mobile cameras according to an embodiment of the present disclosure includes a processor for estimating a joint outside the field of view of a user camera by using self-joint detection information captured by multiple user cameras mounted on a user, searching for a camera of an additional user located in the same space as the user based on landmark information extracted from images captured by the multiple user cameras, selecting candidate joint information for the joint outside the field of view of the user camera in other-user joint detection information captured by the camera of the additional user, reconstructing full-body joints of the user by combining estimated joint information with the candidate joint information, and providing interaction for a full-body motion of the user based on the reconstructed full-body joints; and memory for storing the full-body joints.
[0022]Here, the processor may identify body regions of the user visible in the field of view of the user camera, set a reference point for each of the identified body regions, and estimate the direction of the joint outside the field of view of the user camera based on principal component analysis considering the reference point.
[0023]Here, the principal component analysis may comprise inferring a position of at least one adjacent joint from the reference point and setting a weight for the position of the adjacent joint.
[0024]Here, the weight may be set higher as the position is closer to the reference point, and may be set lower as the position is more distant from the reference point.
[0025]Here, the adjacent joint may correspond to a joint directly connected to a joint corresponding to the reference point in terms of body structure.
[0026]Here, the processor may extract the landmark information based on the images captured by the multiple user cameras, set a global coordinate system by combining the landmark information, determine the position and orientation of the user in the global coordinate system, and detect the camera of the additional user based on the position and orientation of the user.
[0027]Here, the processor may convert an image captured by the camera of the additional user into the global coordinate system and generate the other-user joint detection information by classifying joint information of the user in the image captured by the camera of the additional user based on the position and orientation of the user.
[0028]Here, the candidate joint information may be selected to correspond to other-user joint detection information of a camera of an additional user selected in consideration of joint detection reliability of each camera, and when multiple additional users' cameras having similar reliability are found, the candidate joint information may be selected by further considering a distance from the user in the global coordinate system.
[0029]Here, the multiple user cameras may be mounted around the head of the user and may include a front camera for capturing in a direction in front of the user, a left camera for capturing in a direction toward the ground from the left side of the user, and a right camera for capturing in the direction toward the ground from the right side of the user.
[0030]Here, the front camera, the left camera, and the right camera may operate in a unified coordinate system and correspond to RGBD cameras with standard lenses.
BRIEF DESCRIPTION OF THE DRAWINGS
[0031]The above and other objects, features, and advantages of the present disclosure will be more clearly understood from the following detailed description taken in conjunction with the accompanying drawings, in which:
[0032]
[0033]
[0034]
[0035]
[0036]
[0037]
[0038]
[0039]
[0040]
[0041]
DESCRIPTION OF THE PREFERRED EMBODIMENTS
[0042]The present disclosure will be described in detail below with reference to the accompanying drawings. Repeated descriptions and descriptions of known functions and configurations which have been deemed to unnecessarily obscure the gist of the present disclosure will be omitted below. The embodiments of the present disclosure are intended to fully describe the present disclosure to a person having ordinary knowledge in the art to which the present disclosure pertains. Accordingly, the shapes, sizes, etc. of components in the drawings may be exaggerated in order to make the description clearer.
[0043]In the present specification, each of expressions such as “A or B”, “at least one of A and B”, “at least one of A or B”, “A, B, or C”, “at least one of A, B, and C”, and “at least one of A, B, or C” may include any one of the items listed in the expression or all possible combinations thereof.
[0044]Hereinafter, a preferred embodiment of the present disclosure will be described in detail with reference to the accompanying drawings.
[0045]
[0046]Referring to
[0047]The full-body motion interaction provision apparatus 110 estimates a joint outside the field of view of a user camera by using self-joint detection information captured by multiple user cameras mounted on a user.
[0048]Here, the multiple user cameras are mounted around the head of the user and may include a front camera for capturing in a direction in front of the user, a left camera for capturing in a direction toward the ground from the left side of the user, and a right camera for capturing in the direction toward the ground from the right side of the user.
[0049]Here, the front camera, the left camera, and the right camera operate in a unified coordinate system, and may correspond to RGBD cameras with standard lenses.
[0050]Here, the body regions of the user visible in the field of view of the user camera may be identified, a reference point for each of the identified body regions may be set, and the direction of the joint outside the field of view of the user camera may be estimated based on principal component analysis considering the reference point.
[0051]Here, the principal component analysis may comprise inferring the position of at least one adjacent joint from the reference point and setting a weight for the position of the adjacent joint.
[0052]Here, the weight may be set higher as the position is closer to the reference point, and may be set lower as the position is more distant from the reference point.
[0053]Here, the adjacent joint may correspond to a joint directly connected to a joint corresponding to the reference point in terms of body structure.
[0054]Also, the full-body motion interaction provision apparatus 110 searches for the camera of an additional user located in the same space as the user based on landmark information extracted from the images captured by the multiple user cameras.
[0055]Here, the landmark information may be extracted based on the images captured by the multiple user cameras, a global coordinate system may be set by combining the landmark information, the position and orientation of the user may be determined in the global coordinate system, and the camera of the additional user may be detected based on the position and orientation of the user.
[0056]Also, the full-body motion interaction provision apparatus 110 selects candidate joint information for the joint outside the field of view of the user camera in other-user joint detection information captured by the camera of the additional user.
[0057]Here, an image captured by the camera of the additional user is converted into the global coordinate system, and the joint information of the user in the image captured by the camera of the addition user is classified based on the position and orientation of the user, whereby the other-user joint detection information may be generated.
[0058]Here, the candidate joint information is selected to correspond to other-user joint detection information of the camera of an additional user selected in consideration of joint detection reliability of each camera, and when multiple additional users' cameras having similar reliability are found, the candidate joint information may be selected by further considering the distance from the user in the global coordinate system.
[0059]Also, the full-body motion interaction provision apparatus 110 reconstructs the full-body joints of the user by combining estimated joint information with the candidate joint information and provides interaction for a full-body motion of the user based on the reconstructed full-body joints.
[0060]The user terminals 120-1 to 120-N may correspond to terminals worn or held by respective users that use the system according to the present disclosure.
[0061]Here, the user terminals 120-1 to 120-N may include multiple user cameras and a haptic device.
[0062]For example, the user terminals 120-1 to 120-N may provide the images captured by the multiple user cameras to the full-body motion interaction provision apparatus 110 through the network and may receive interaction for the full-body motion of the user from the full-body motion interaction provision apparatus 110 and implement the same with the haptic device.
[0063]
[0064]Referring to
[0065]Here, the multiple user cameras are mounted around the head of the user and may include a front camera for capturing in a direction in front of the user, a left camera for capturing in a direction toward the ground from the left side of the user, and a right camera for capturing in the direction toward the ground from the right side of the user.
[0066]Here, the front camera, the left camera, and the right camera may operate in a unified coordinate system and correspond to RGBD cameras with standard lenses.
[0067]That is, the present disclosure is for reconstructing body joints of a user using RGBD cameras attached to the body of the user and providing an interface in a virtual environment, and may operate with the configuration illustrated in
[0068]For example, three RGBD cameras may be attached around the head of the user. The front camera may be placed to face the front of the user, as illustrated in
[0069]Here, the body regions of the user visible in the field of view of the user camera may be identified, a reference point may be set for each of the identified body regions, and the direction of the joint outside the field of view of the user camera may be estimated based on principal component analysis considering the reference point.
[0070]Here, principal component analysis may comprise inferring the position of at least one adjacent joint from the reference point and setting a weight for the position of the adjacent joint.
[0071]Here, the weight may be set higher as the position is closer to the reference point, and may be set lower as the position is more distant from the reference point.
[0072]Here, the adjacent joint may correspond to a joint directly connected to a joint corresponding to the reference point in terms of body structure.
[0073]For example, the user's own joints and the joints of others may be detected using the three cameras attached to the body of the user.
[0074]Referring to
[0075]Also, referring to
[0076]Here, through ‘self-directed joint estimation’ in
[0077]Therefore, in the present disclosure, the body part outside the FOV of the camera may be estimated through the process illustrated in
[0078]Referring to
[0079]First, segmenting the body parts at step S810 may comprise identifying the body regions of a user visible from the first-person view. For example, a shoulder, an upper arm, a forearm, a hand, and the like may be identified.
[0080]Subsequently, setting the anchor joint at step S820 may comprise setting a reference point using the joint observed in the FOV of the camera. For example, the parts corresponding to the actual joint positions 710 in
[0081]Subsequently, the principal component analysis based on the body part weight at step S830 may comprise inferring the positions of adjacent joints outside the FOV of the camera by using the set reference points.
[0082]Here, an adjacent joint may indicate a directly connected joint in terms of body structure. For example, the adjacent joints of an elbow may correspond to a wrist and a shoulder, and the adjacent joints of a shoulder may correspond to a neck and an elbow.
[0083]Also, the principal component analysis based on the body part weight at step S830 may comprise performing principal component analysis on the region extracted from the body part segment by setting the position of the reference point (the anchor joint) as the starting point and determining the direction in which the invisible joint is located from the visible joint (the starting point).
[0084]Here, the body region that is not adjacent to the starting point may be assigned a low weight. For example, when the position of the hand is estimated, the body region of the upper arm may be excluded from the principal component analysis.
[0085]Also, if it is far away from the anchor joint, the adjacent body region may also be assigned a low weight. For example, when principal component analysis is performed to estimate an elbow that is invisible from a shoulder joint, a region corresponding to part of an upper arm that is observed at the position far away from the shoulder may be assigned a low weight.
[0086]Subsequently, estimating the user's own joint at step S840 may comprise estimating the position of the joint outside the FOV of the camera.
[0087]Here, the direction of the joint, estimated through the principal component analysis performed by setting the reference point (the anchor point) as the starting point, is combined with the length of the joint, whereby the position of the joint outside the FOV of the camera may be estimated. If two or more adjacent joints are detected (e.g., if a shoulder and a hand are detected but an elbow is not detected), as illustrated in
[0088]Here, the length of the joint may be estimated using standard body size information or through the observed joint length information.
[0089]Also, in the method for providing full-body motion interaction using multiple mobile cameras according to an embodiment of the present disclosure, the full-body motion interaction provision apparatus searches for the camera of an addition user located in the same space as the user based on landmark information extracted from images captured by the multiple user cameras at step S220.
[0090]Here, the landmark information may be extracted based on the images captured by the multiple user cameras, a global coordinate system may be set by combining the landmark information, the position and orientation of the user may be determined in the global coordinate system, and the camera of the additional user may be detected based on the position and orientation of the user.
[0091]Here, the landmark information is information about a background or fixed object, and it may be used as the information for estimating the position and orientation of each of the users in the same space.
[0092]For example, referring to
[0093]Here, the estimated landmark information may be delivered to the process of ‘user information sharing’, along with joint information of others.
[0094]For example, referring to
[0095]First, the user position alignment at step S1110 may comprise setting a reference coordinate system (a global coordinate system) by combining the pieces of landmark information estimated by the user and determining the position and orientation of the user in the set coordinate system.
[0096]Subsequently, the user joint classification at step S1120 may comprise receiving other-user joint detection information, which is detected when the user is in the same space, converting the other-user joint detection information into the reference coordinate system, and classifying the joint information corresponding to each user.
[0097]Here, the joint of each user may be classified using clustering based on the distance between joints expressed in the global coordinate system or the feature similarity between the body regions of the user.
[0098]Also, in the method for providing full-body motion interaction using multiple mobile cameras according to an embodiment of the present disclosure, the full-body motion interaction provision apparatus selects candidate joint information for the joint outside the field of view of the user camera in the other-user joint detection information captured by the camera of the additional user at step S230.
[0099]Here, the image captured by the camera of the additional user is converted into the global coordinate system, and the joint information of the user is classified in the image captured by the camera of the additional user based on the position and orientation of the user, whereby the other-user joint detection information may be generated.
[0100]Here, the candidate joint information may be selected to correspond to other-user joint detection information of the camera of an additional user selected in consideration of joint detection reliability of each camera, and when multiple additional users' cameras having similar reliability are found, the candidate joint information may be selected by further considering the distance from the user in the global coordinate system.
[0101]For example, referring to
[0102]Here, the reliability that is measured when joints are detected by each camera may be used, and when cameras have similar reliability, information of the camera that is closer to the user may be selected and used as the candidate joint.
[0103]Also, through ‘user joint reconstruction’ illustrated in
[0104]For example, when it is difficult to reconstruct joints because consecutive joints are outside the FOV of the camera as illustrated in
[0105]However, when reconstructing joints, the joint detected through ‘self-joint detection’ may have higher priority than information acquired through ‘other-user joint detection’, and the full-body joints of the user may be reconstructed using the angle between the joints detected by the additional user and the length information of the joints.
[0106]Also, when a single user is present, similar joint information accumulated in a database may be collected using information acquired through ‘self-directed joint estimation’ and information acquired through ‘user joint exploration’. By assuming that the information collected in this way is joint information detected by an additional user, this information may be used to reconstruct the joints of the user.
[0107]Also, in the method for providing full-body motion interaction using multiple mobile cameras according to an embodiment of the present disclosure, the full-body motion interaction provision apparatus constructs the full-body joints of the user by combining estimated joint information with the candidate joint information and provides interaction for a full-body motion of the user based on the reconstructed full-body joints at step S240.
[0108]That is, the full-body joint information of the user, which is finally reconstructed through the process illustrated in
[0109]Through the above-described method for providing full-body motion interaction using multiple mobile cameras, full-body motion interaction may be provided by more accurately estimating 3D full-body joints using multiple RGBD cameras with standard lenses, rather than using fisheye or wide-angle lenses.
[0110]Also, using cameras with standard lenses installed around the head of a user, joints may be more accurately estimated even in an area that is outside the field of view of a camera or heavily occluded.
[0111]Also, interaction with a virtual or real object may be supported using a reconstructed 3D full-body motion in an XR environment.
[0112]
[0113]Referring to
[0114]Accordingly, an embodiment of the present disclosure may be implemented as a non-transitory computer-readable medium in which methods implemented using a computer or instructions executable in a computer are recorded. When the computer-readable instructions are executed by a processor, the computer-readable instructions may perform a method according to at least one aspect of the present disclosure.
[0115]The processor 1310 estimates a joint outside the field of view of a user camera by using self-joint detection information captured by multiple user cameras mounted on a user.
[0116]Here, the multiple user cameras may be mounted around the head of the user and may include a front camera for capturing in a direction in front of the user, a left camera for capturing in a direction toward the ground from the left side of the user, and a right camera for capturing in the direction toward the ground from the right side of the user.
[0117]Here, the front camera, the left camera, and the right camera may operate in a unified coordinate system and correspond to RGBD cameras with standard lenses.
[0118]Here, the body regions of the user visible in the field of view of the user camera may be identified, a reference point may be set for each of the identified body regions, and the direction of the joint outside the field of view of the user camera may be estimated based on principal component analysis considering the reference point.
[0119]Here, the principal component analysis may comprise inferring the position of at least one adjacent joint from the reference point and setting a weight for the position of the adjacent joint.
[0120]Here, the weight may be set higher as the position is closer to the reference point, and may be set lower as the position is more distant from the reference point.
[0121]Here, the adjacent joint may correspond to a joint directly connected to a joint corresponding to the reference point in terms of body structure.
[0122]Also, the processor 1310 searches for a camera of an additional user located in the same space as the user based on landmark information extracted from the images captured by the multiple user cameras.
[0123]Here, the landmark information may be extracted based on the images captured by the multiple user cameras, a global coordinate system may be set by combining the landmark information, the position and orientation of the user may be determined in the global coordinate system, and the camera of the additional user may be detected based on the position and orientation of the user.
[0124]Also, the processor 1310 selects candidate joint information for the joint outside the field of view of the user camera in other-user joint detection information captured by the camera of the additional user.
[0125]Here, the image captured by the camera of the additional user is converted into the global coordinate system, and the joint information of the user is classified in the image captured by the camera of the additional user based on the position and orientation of the user, whereby the other-user joint detection information may be generated.
[0126]Here, the candidate joint information is selected to correspond to other-user joint detection information of the camera of an additional user selected in consideration of the joint detection reliability of each camera, and when multiple additional users' cameras having similar reliability are found, the candidate joint information may be selected by further considering the distance from the user in the global coordinate system.
[0127]Also, the processor 1310 reconstructs the full-body joints of the user by combining estimated joint information with the candidate joint information and provides interaction for a full-body motion of the user based on the reconstructed full-body joints.
[0128]The memory 1330 stores various kinds of information generated in the above-described apparatus for providing full-body motion interaction according to an embodiment of the present disclosure.
[0129]According to an embodiment, the memory 1330 may be separate from the apparatus for providing full-body motion interaction, and may support the function for providing full-body motion interaction. Here, the memory 1330 may operate as separate mass storage, and may include a control function for performing operations.
[0130]Meanwhile, the apparatus for providing full-body motion interaction includes memory installed therein, whereby information may be stored therein. In an embodiment, the memory is a computer-readable medium. In an embodiment, the memory may be a volatile memory unit, and in another embodiment, the memory may be a nonvolatile memory unit. In an embodiment, the storage device is a computer-readable medium. In different embodiments, the storage device may include, for example, a hard-disk device, an optical disk device, or any other kind of mass storage device.
[0131]Using the above-described apparatus for providing full-body motion interaction using multiple mobile cameras, full-body motion interaction may be provided by more accurately estimating 3D full-body joints using multiple RGBD cameras with standard lenses, rather than using fisheye or wide-angle lenses.
[0132]Also, using cameras with standard lenses installed around the head of a user, joints may be more accurately estimated even in an area that is outside the field of view of a camera or heavily occluded.
[0133]Also, interaction with a virtual or real object may be supported using a reconstructed 3D full-body motion in an XR environment.
[0134]According to the present disclosure, full-body motion interaction may be provided by more accurately estimating 3D full-body joints using multiple RGBD cameras with standard lenses, rather than using fisheye or wide-angle lenses.
[0135]Also, the present disclosure uses cameras with standard lenses installed around the head of a user, thereby more accurately estimating joints even in an area that is outside the field of view of a camera or heavily occluded.
[0136]Also, the present disclosure may support interaction with a virtual or real object using a reconstructed 3D full-body motion in an XR environment.
[0137]As described above, the method for providing full-body motion interaction using multiple mobile cameras and the apparatus for the same according to the present disclosure are not limitedly applied to the configurations and operations of the above-described embodiments, but all or some of the embodiments may be selectively combined and configured, so the embodiments may be modified in various ways.
Claims
What is claimed is:
1. A method for providing full-body motion interaction, performed by an apparatus for providing full-body motion interaction, comprising:
estimating a joint outside a field of view of a user camera by using self-joint detection information captured by multiple user cameras mounted on a user;
searching for a camera of an additional user located in a same space as the user based on landmark information extracted from images captured by the multiple user cameras;
selecting candidate joint information for the joint outside the field of view of the user camera in other-user joint detection information captured by the camera of the additional user; and
reconstructing full-body joints of the user by combining estimated joint information with the candidate joint information and providing interaction for a full-body motion of the user based on the reconstructed full-body joints.
2. The method of
identifying body regions of the user visible in the field of view of the user camera;
setting a reference point for each of the identified body regions; and
estimating a direction of the joint outside the field of view of the user camera based on principal component analysis considering the reference point.
3. The method of
4. The method of
5. The method of
6. The method of
extracting the landmark information based on the images captured by the multiple user cameras;
setting a global coordinate system by combining the landmark information;
determining a position and orientation of the user in the global coordinate system; and
detecting the camera of the additional user based on the position and orientation of the user.
7. The method of
converting an image captured by the camera of the additional user into the global coordinate system; and
generating the other-user joint detection information by classifying joint information of the user in the image captured by the camera of the additional user based on the position and orientation of the user.
8. The method of
9. The method of
10. The method of
11. An apparatus for providing full-body motion interaction, comprising:
a processor for estimating a joint outside a field of view of a user camera by using self-joint detection information captured by multiple user cameras mounted on a user, searching for a camera of an additional user located in a same space as the user based on landmark information extracted from images captured by the multiple user cameras, selecting candidate joint information for the joint outside the field of view of the user camera in other-user joint detection information captured by the camera of the additional user, reconstructing full-body joints of the user by combining estimated joint information with the candidate joint information, and providing interaction for a full-body motion of the user based on the reconstructed full-body joints; and
memory for storing the full-body joints.
12. The apparatus of
13. The apparatus of
14. The apparatus of
15. The apparatus of
16. The apparatus of
17. The apparatus of
18. The apparatus of
19. The apparatus of
20. The apparatus of