WO2019211429A1 - A method and an apparatus for reducing an amount of data representative of a multi-view plus depth content - Google Patents

A method and an apparatus for reducing an amount of data representative of a multi-view plus depth content Download PDF

Info

Publication number
WO2019211429A1
WO2019211429A1 PCT/EP2019/061357 EP2019061357W WO2019211429A1 WO 2019211429 A1 WO2019211429 A1 WO 2019211429A1 EP 2019061357 W EP2019061357 W EP 2019061357W WO 2019211429 A1 WO2019211429 A1 WO 2019211429A1
Authority
WO
WIPO (PCT)
Prior art keywords
view
pixel
coordinates
color
point
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/EP2019/061357
Other languages
French (fr)
Inventor
Neus SABATER
Didier Doyen
Guillaume Boisson
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
InterDigital CE Patent Holdings SAS
Original Assignee
InterDigital CE Patent Holdings SAS
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by InterDigital CE Patent Holdings SAS filed Critical InterDigital CE Patent Holdings SAS
Publication of WO2019211429A1 publication Critical patent/WO2019211429A1/en
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Classifications

    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N19/00Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
    • H04N19/50Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using predictive coding
    • H04N19/597Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using predictive coding specially adapted for multi-view video sequence encoding
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N19/00Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
    • H04N19/50Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using predictive coding
    • H04N19/503Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using predictive coding involving temporal prediction
    • H04N19/507Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using predictive coding involving temporal prediction using conditional replenishment

Definitions

  • the present disclosure relates to multi-view plus depth contents. More particularly, the disclosure aims at reducing the amount of effective data contained in a multi-view plus depth content.
  • a camera array or a plenoptic camera can capture a scene from different viewpoints.
  • the amount of data representing the content is huge and comprises a lot of redundancies.
  • the 3D-HEVC extension of the MPEG HEVC standard allows a global compression of the multi view plus depth content including the use of depth information in the prediction loop.
  • the 3D- HEVC standard considers the whole set of data and compresses redundant information.
  • the proposed method consists in removing pixels in a multi-view plus depth content which are redundant since they are not providing any new information. To be considered as redundant a pixel shall have at the same time similar RGB values and depth value.
  • the present invention proposes a method to determine if a given pixel of one view is redundant with another one in another view.
  • Some processes implemented by elements of the invention may be computer implemented. Accordingly, such elements may take the form of an entirely hardware embodiment, an entirely software embodiment (including firmware, resident software, micro code, etc.) or an embodiment combining software and hardware aspects that may all generally be referred to herein as a "circuit", "module” or “system'. Furthermore, such elements may take the form of a computer program product embodied in any tangible medium of expression having computer usable program code embodied in the medium.
  • a tangible carrier medium may comprise a storage medium such as a floppy disk, a CD-ROM, a hard disk drive, a magnetic tape device or a solid- state memory device and the like.
  • a transient carrier medium may include a signal such as an electrical signal, an electronic signal, an optical signal, an acoustic signal, a magnetic signal or an electromagnetic signal, e.g. a microwave or RF signal.
  • Figure 1 is a flowchart of a method for reducing redundancies in a multi-view plus depth content according to an embodiment
  • Figure 2 represents how a pixel P(u,v) is de-projected in the world coordinate system, how the point Pw is re-projected in the projection view Vp and how the projection pixel Pp(u’,v’)i is de-projected from the projection view Vp into the world coordinate system
  • Figure 3 is a schematic block diagram illustrating an example of a device capable of executing the method of the disclosure.
  • aspects of the present principles can be embodied as a system, method or computer readable medium. Accordingly, aspects of the present principles can take the form of an entirely hardware embodiment, an entirely software embodiment, (including firmware, resident software, micro-code, and so forth) or an embodiment combining software and hardware aspects that can all generally be referred to herein as a“circuit”,“module”, or“system”. Furthermore, aspects of the present principles can take the form of a computer readable storage medium. Any combination of one or more computer readable storage medium(a) may be utilized.
  • a multi-view plus depth content is acquired by a 4x4 camera rig.
  • the multi-view plus depth content may be acquired by a plenoptic camera.
  • the cameras are calibrated intrinsically and extrinsically using for example Sparse Bundle Adjustment.
  • the cameras of the camera rig are calibrated, for example, to fit a distorted pinhole projection model.
  • Figure 1 is a flowchart of a method for reducing redundancies in a multi-view plus depth content according to an embodiment. Such a method may be executed, for example, in a camera, a smartphone, or any other device capable of processing a multi-view plus depth content.
  • the multi-view plus depth content comprises n views, each view being acquired by one of the CN cameras constituting the camera rig.
  • a first reference view VI is selected from the n views of the multi-view plus depth content.
  • the first reference view VI corresponds to a view acquired by a camera Ci located on the top left of the camera rig.
  • the reference view VI is selected among the views of the matrix of views calculated from the lenslet content.
  • Each of the n views of the multi-view plus depth content comprises P pixels of coordinates (u,v).
  • the set of coordinates (Xw) is obtained using a pose matrix, the inverse intrinsic matrix of the camera Ci from which the reference view VI is acquired, and the depth associated to the pixel (Zuv) of the first reference view V 1.
  • P (R T) G D3 ⁇ 4 3 x4 is the pose matrix of the camera C in the world coordinate system
  • Q (R _1 — R _1 T) G M 3x4 is its extrinsic matrix.
  • K be the intrinsic matrix of the camera Ci; where / is the distance from the optical center to the sensor expressed in pixels, ) is the principal point.
  • the projection view Vp is defined among the set of n views of the multi view plus depth content different from the first reference view V 1.
  • a pixel Pproj (up, vp) in the projection view Vp of the multi-view plus depth content is obtained using the intrinsic matrix and the Q matrix of the camera C2 from which the projection view Vp is acquired.
  • the projection of the point Pw may not correspond to an actual pixel of the projection view Vp but most probably in between a plurality of pixels of the projection view
  • a projection pixel Pp(u’,v’) in the projection view Vp is selected which is the one that reduces the distance between the result of the projection Pproj(up, vp) of the point Pw in the projection view Vp and one of the neighboring pixels: o Pproj 1 (int(up), int(vp))
  • the distance is calculated as the square root of the sum of square difference between coordinates.
  • a step 106 the projection pixel Pp(u’,v’) is de-projected in the world coordinate system as in step 102 but using matrix of the camera C2.
  • a set of coordinates (X’w) of a second point P’w in the world coordinate system is obtained which corresponds to the pixel Pp(u’,v’).
  • Figure 2 represents how a pixel P(u,v) is de-projected in the world coordinate system, how the point Pw is re-projected in the projection view Vp and how the projection pixel Pp(u’,v’)i is de-projected from the projection view Vp into the world coordinate system.
  • camera Ci is the camera which acquired the reference view VI
  • camera C2 is the camera which acquired the projection view Vp.
  • a step 107 the sets of coordinates (Xw) and (X’w) of the first point Pw and the second point P’w are compared. In an embodiment, if the square root of the sum of square distance between each component of coordinates (Xw) and (X’w) is below a threshold, the two points Pw and P’w, and consequently the two corresponding pixels P(u,v) in the reference view VI and Pp(u’,v’) in the projection view Vp, are considered as redundant.
  • a similar comparison is applied on a parameter representative of the colors associated with the pixels P 1 (u,v) and Pp(u’ ,v’ ) .
  • the difference between the two parameters representative of a color of the pixels Pl(u,v) and Pp(u’,v’) may be expressed as sum of square difference for each component (e.g. R, G and B). In an embodiment it may be the sum of absolute values of the three differences.
  • the parameters representative of the color of a pixel may be RGB values or in any other color reference system (e.g. Lab, ).
  • Step 108 since the second pixel Pp(u’,v’) is redundant with the first pixel Pl(u,v) of the reference view VI, the second pixel Pp(u’,v’) is removed from the projection view Vp. Steps 104 to 108 are executed for the pixel Pl(u,v) of the first reference view VI in all the remaining views of the multi-view plus depth content except the reference view V 1.
  • steps 104 to 108 are executed for the pixel Pl(u,v) of the first reference view VI in all the remaining views of the multi-view plus depth content , steps 102 to 108 are executed for the other pixels of the reference view V 1.
  • a new reference view V2 is selected in a step 109.
  • the new reference view is selected among the remaining views of multi-view plus depth content in which pixels which are redundant with pixels of the first reference view V 1 are removed.
  • a third reference view V3 is selected in step 109.
  • the third reference view is selected among the remaining views of multi-view plus depth content in which pixels which are redundant with pixels of the second reference view V2 are removed.
  • Step 109 is executed until all the n views of the multi-view plus depth content are processed. .
  • Figure 3 is a schematic block diagram illustrating an example of a device capable of executing the method described in reference to figure 1.
  • a device may be a camera, a camera rig, a smartphone, etc.
  • the apparatus 300 comprises a processor 301, a storage unit 302, an input device 303, a display device 304, and an interface unit 305 which are connected by a bus 306. Of course, constituent elements of the apparatus 300 may be connected by a connection other than a bus connection.
  • the processor 301 controls operations of the apparatus 300.
  • the storage unit 302 stores at least one program to be executed by the processor 301, and various data such as a multi-view plus depth content, parameters used by computations performed by the processor 301, intermediate data of computations performed by the processor 301, and so on.
  • the processor 301 may be formed by any known and suitable hardware, or software, or a combination of hardware and software.
  • the processor 301 may be formed by dedicated hardware such as a processing circuit, or by a programmable processing unit such as a CPU (Central Processing Unit) that executes a program stored in a memory thereof.
  • CPU Central Processing Unit
  • the storage unit 302 may be formed by any suitable storage or means capable of storing the program, data, or the like in a computer-readable manner. Examples of the storage unit 302 include non-transitory computer-readable storage media such as semiconductor memory devices, and magnetic, optical, or magneto-optical recording media loaded into a read and write unit.
  • the program causes the processor 301 to perform the method as described with reference to figure 1.
  • the input device 303 may be formed by a keyboard, a pointing device such as a mouse, or the like for use by the user to input commands, etc.
  • the output device 304 may be formed by a display device to display, for example, a Graphical User Interface (GUI).
  • GUI Graphical User Interface
  • the input device 303 and the output device 304 may be formed integrally by a touchscreen panel, for example.
  • the interface unit 305 provides an interface between the apparatus 300 and an external apparatus.
  • the interface unit 305 may be communicable with the external apparatus via cable or wireless communication.
  • a method for reducing an amount of data representative of a multi-view plus depth content comprising: a) selecting a reference view in said multi-view plus depth content, b) obtaining a first parameter representative of a color as well as a first set of coordinates of a first point corresponding to a pixel in said reference view when said first pixel is projected in a coordinate system,
  • An apparatus capable of reducing an amount of data representative of a multi-view plus depth content, said apparatus comprising at least a processor configured to : a) select a reference view in said multi-view plus depth content,
  • the method and the apparatus comprising selecting a second reference view in the multi-view plus depth content on which the removing d) was performed, the method and the apparatus further comprising: e) obtaining a third parameter representative of a color as well as a third set of coordinates of a third point corresponding to a pixel in said second reference view when said pixel is projected in said coordinate system,
  • a computer program characterized in that it comprises program code instructions for the implementation of the method according to claim 1 when the program is executed by a processor.

Landscapes

  • Engineering & Computer Science (AREA)
  • Multimedia (AREA)
  • Signal Processing (AREA)
  • Processing Or Creating Images (AREA)
  • Image Processing (AREA)

Abstract

The present disclosure relates to multi-view plus depth contents. When considering multi-view acquisition, the amount of data representing the content is huge and comprises a lot of redundancies. The 3D-HEVC standard considers the whole set of data and compresses redundant information. Thus, the amount of data representing a multi-view plus depth content, though compressed, remains huge. The proposed method consists in removing pixels in a multi- view plus depth content which are redundant since they are not providing any new information.

Description

A METHOD AND AN APPARATUS FOR REDUCING AN AMOUNT OF DATA
REPRESENTATIVE OF A MULTI- VIEW PLUS DEPTH CONTENT
TECHNICAL FIELD The present disclosure relates to multi-view plus depth contents. More particularly, the disclosure aims at reducing the amount of effective data contained in a multi-view plus depth content.
BACKGROUND
A camera array or a plenoptic camera can capture a scene from different viewpoints. When considering multi-view acquisition, the amount of data representing the content is huge and comprises a lot of redundancies.
Compression schemes for multi-view plus depth content exist for content delivery. The 3D-HEVC extension of the MPEG HEVC standard allows a global compression of the multi view plus depth content including the use of depth information in the prediction loop. The 3D- HEVC standard considers the whole set of data and compresses redundant information.
Thus, the amount of data representing a multi-view plus depth content, though compressed, remains huge.
The present disclosure has been devised with the foregoing in mind.
SUMMARY OF INVENTION The proposed method consists in removing pixels in a multi-view plus depth content which are redundant since they are not providing any new information. To be considered as redundant a pixel shall have at the same time similar RGB values and depth value. The present invention proposes a method to determine if a given pixel of one view is redundant with another one in another view.
Some processes implemented by elements of the invention may be computer implemented. Accordingly, such elements may take the form of an entirely hardware embodiment, an entirely software embodiment (including firmware, resident software, micro code, etc.) or an embodiment combining software and hardware aspects that may all generally be referred to herein as a "circuit", "module" or "system'. Furthermore, such elements may take the form of a computer program product embodied in any tangible medium of expression having computer usable program code embodied in the medium.
Since elements of the present invention can be implemented in software, the present invention can be embodied as computer readable code for provision to a programmable apparatus on any suitable carrier medium. A tangible carrier medium may comprise a storage medium such as a floppy disk, a CD-ROM, a hard disk drive, a magnetic tape device or a solid- state memory device and the like. A transient carrier medium may include a signal such as an electrical signal, an electronic signal, an optical signal, an acoustic signal, a magnetic signal or an electromagnetic signal, e.g. a microwave or RF signal.
BRIEF DESCRIPTION OF THE DRAWINGS
Embodiments of the invention will now be described, by way of example only, and with reference to the following drawings in which:
Figure 1 is a flowchart of a method for reducing redundancies in a multi-view plus depth content according to an embodiment,
Figure 2 represents how a pixel P(u,v) is de-projected in the world coordinate system, how the point Pw is re-projected in the projection view Vp and how the projection pixel Pp(u’,v’)i is de-projected from the projection view Vp into the world coordinate system, Figure 3 is a schematic block diagram illustrating an example of a device capable of executing the method of the disclosure.
DETAILED DESCRIPTION
As will be appreciated by one skilled in the art, aspects of the present principles can be embodied as a system, method or computer readable medium. Accordingly, aspects of the present principles can take the form of an entirely hardware embodiment, an entirely software embodiment, (including firmware, resident software, micro-code, and so forth) or an embodiment combining software and hardware aspects that can all generally be referred to herein as a“circuit”,“module”, or“system”. Furthermore, aspects of the present principles can take the form of a computer readable storage medium. Any combination of one or more computer readable storage medium(a) may be utilized.
In an embodiment, a multi-view plus depth content is acquired by a 4x4 camera rig. In another embodiment, the multi-view plus depth content may be acquired by a plenoptic camera. In both embodiments, the cameras are calibrated intrinsically and extrinsically using for example Sparse Bundle Adjustment. In case the multi-view plus depth content is acquired by the camera rig, the cameras of the camera rig are calibrated, for example, to fit a distorted pinhole projection model.
Figure 1 is a flowchart of a method for reducing redundancies in a multi-view plus depth content according to an embodiment. Such a method may be executed, for example, in a camera, a smartphone, or any other device capable of processing a multi-view plus depth content. The multi-view plus depth content comprises n views, each view being acquired by one of the CN cameras constituting the camera rig.
In a step 101, a first reference view VI is selected from the n views of the multi-view plus depth content. For example, the first reference view VI corresponds to a view acquired by a camera Ci located on the top left of the camera rig. In an embodiment where the multi-view plus depth content is acquired by a plenoptic camera, the reference view VI is selected among the views of the matrix of views calculated from the lenslet content. Each of the n views of the multi-view plus depth content comprises P pixels of coordinates (u,v). In a step 102, a set of coordinates (Xw), in a coordinate system such as the world coordinate system, of a point Pw corresponding to a pixel Pl(u,v) of the first reference view VI, is obtained. The set of coordinates (Xw) is obtained using a pose matrix, the inverse intrinsic matrix of the camera Ci from which the reference view VI is acquired, and the depth associated to the pixel (Zuv) of the first reference view V 1. Considering a camera Ci of the camera rig, P = (R T) G D¾3 x4 is the pose matrix of the camera C in the world coordinate system, and Q = (R_1 — R_1 T) G M3x4 is its extrinsic matrix.
Let K = be the intrinsic matrix of the camera Ci; where / is the distance
Figure imgf000006_0001
from the optical center to the sensor expressed in pixels, ) is the principal point. Now, if Xw is a 3D point in the world coordinate system and X denotes the corresponding pixel in the coordinate system of the camera Ci, then, Xw = P and X =
Figure imgf000006_0002
Figure imgf000006_0003
Then, given a 3D point Xw in the world coordinate system, its projection in pixel coordinates in the camera Ci image plane is determined by
Figure imgf000006_0004
and W\ M2 ® M2 denotes the distortion warping operator, defined
Figure imgf000007_0001
in the plane z = 1 m. s
Note that, using the homogeneous coordinates:
Figure imgf000007_0002
t y
z
A given point in pixel coordinates in the camera Ci image plane is de-projected in
Figure imgf000007_0003
the world coordinate system Xw
Figure imgf000007_0004
In a step 103, the projection view Vp is defined among the set of n views of the multi view plus depth content different from the first reference view V 1.
In a step 104, knowing the coordinates (Xw) of the first point Pw in the world coordinate system, a pixel Pproj (up, vp) in the projection view Vp of the multi-view plus depth content is obtained using the intrinsic matrix and the Q matrix of the camera C2 from which the projection view Vp is acquired. The projection of the point Pw may not correspond to an actual pixel of the projection view Vp but most probably in between a plurality of pixels of the projection view
Vp. In a step 105, a projection pixel Pp(u’,v’) in the projection view Vp is selected which is the one that reduces the distance between the result of the projection Pproj(up, vp) of the point Pw in the projection view Vp and one of the neighboring pixels: o Pproj 1 (int(up), int(vp))
o Pproj2 (int(up)+l, int(vp))
o Pproj 3 (int(up), int(vp)+l)
o Pproj 4 (int(up)+ 1 , int(vp)+ 1 ) In an embodiment, the distance is calculated as the square root of the sum of square difference between coordinates.
In a step 106, the projection pixel Pp(u’,v’) is de-projected in the world coordinate system as in step 102 but using matrix of the camera C2. In other words, a set of coordinates (X’w) of a second point P’w in the world coordinate system is obtained which corresponds to the pixel Pp(u’,v’).
Figure 2 represents how a pixel P(u,v) is de-projected in the world coordinate system, how the point Pw is re-projected in the projection view Vp and how the projection pixel Pp(u’,v’)i is de-projected from the projection view Vp into the world coordinate system. On figure 2, camera Ci is the camera which acquired the reference view VI and camera C2 is the camera which acquired the projection view Vp.
In a step 107, the sets of coordinates (Xw) and (X’w) of the first point Pw and the second point P’w are compared. In an embodiment, if the square root of the sum of square distance between each component of coordinates (Xw) and (X’w) is below a threshold, the two points Pw and P’w, and consequently the two corresponding pixels P(u,v) in the reference view VI and Pp(u’,v’) in the projection view Vp, are considered as redundant.
A similar comparison is applied on a parameter representative of the colors associated with the pixels P 1 (u,v) and Pp(u’ ,v’ ) . The difference between the two parameters representative of a color of the pixels Pl(u,v) and Pp(u’,v’) may be expressed as sum of square difference for each component (e.g. R, G and B). In an embodiment it may be the sum of absolute values of the three differences. The parameters representative of the color of a pixel may be RGB values or in any other color reference system (e.g. Lab, ...).
In a step 108, since the second pixel Pp(u’,v’) is redundant with the first pixel Pl(u,v) of the reference view VI, the second pixel Pp(u’,v’) is removed from the projection view Vp. Steps 104 to 108 are executed for the pixel Pl(u,v) of the first reference view VI in all the remaining views of the multi-view plus depth content except the reference view V 1.
Once steps 104 to 108 are executed for the pixel Pl(u,v) of the first reference view VI in all the remaining views of the multi-view plus depth content , steps 102 to 108 are executed for the other pixels of the reference view V 1.
When all the pixels of the reference view VI are processed during steps 102 to 108, a new reference view V2 is selected in a step 109. The new reference view is selected among the remaining views of multi-view plus depth content in which pixels which are redundant with pixels of the first reference view V 1 are removed. When all the pixels of the new reference view V2 are processed during steps 102 to 108, a third reference view V3 is selected in step 109. The third reference view is selected among the remaining views of multi-view plus depth content in which pixels which are redundant with pixels of the second reference view V2 are removed.
Step 109 is executed until all the n views of the multi-view plus depth content are processed. .
Figure 3 is a schematic block diagram illustrating an example of a device capable of executing the method described in reference to figure 1. Such a device may be a camera, a camera rig, a smartphone, etc.
The apparatus 300 comprises a processor 301, a storage unit 302, an input device 303, a display device 304, and an interface unit 305 which are connected by a bus 306. Of course, constituent elements of the apparatus 300 may be connected by a connection other than a bus connection. The processor 301 controls operations of the apparatus 300. The storage unit 302 stores at least one program to be executed by the processor 301, and various data such as a multi-view plus depth content, parameters used by computations performed by the processor 301, intermediate data of computations performed by the processor 301, and so on. The processor 301 may be formed by any known and suitable hardware, or software, or a combination of hardware and software. For example, the processor 301 may be formed by dedicated hardware such as a processing circuit, or by a programmable processing unit such as a CPU (Central Processing Unit) that executes a program stored in a memory thereof.
The storage unit 302 may be formed by any suitable storage or means capable of storing the program, data, or the like in a computer-readable manner. Examples of the storage unit 302 include non-transitory computer-readable storage media such as semiconductor memory devices, and magnetic, optical, or magneto-optical recording media loaded into a read and write unit. The program causes the processor 301 to perform the method as described with reference to figure 1. The input device 303 may be formed by a keyboard, a pointing device such as a mouse, or the like for use by the user to input commands, etc. The output device 304 may be formed by a display device to display, for example, a Graphical User Interface (GUI). The input device 303 and the output device 304 may be formed integrally by a touchscreen panel, for example.
The interface unit 305 provides an interface between the apparatus 300 and an external apparatus. The interface unit 305 may be communicable with the external apparatus via cable or wireless communication.
A method for reducing an amount of data representative of a multi-view plus depth content, said method comprising: a) selecting a reference view in said multi-view plus depth content, b) obtaining a first parameter representative of a color as well as a first set of coordinates of a first point corresponding to a pixel in said reference view when said first pixel is projected in a coordinate system,
c) obtaining a second parameter representative of a color as well as a second set of coordinates of a second point corresponding to a pixel of another view of said multi-view plus depth content when said pixel is projected said coordinate system, said pixel in said other view corresponding to a projection of said first point in said other view,
d) removing said pixel in said other view when the first parameter representative of a color as well as the first set of coordinates of the first point corresponds to the second parameter representative of a color as well as the second set of coordinates of the second point.
An apparatus capable of reducing an amount of data representative of a multi-view plus depth content, said apparatus comprising at least a processor configured to : a) select a reference view in said multi-view plus depth content,
b) obtain a first parameter representative of a color as well as a first set of coordinates of a first point corresponding to a pixel in said reference view when said first pixel is projected in a coordinate system,
c) obtain a second parameter representative of a color as well as a second set of coordinates of a second point corresponding to a pixel of another view of said multi-view plus depth content when said pixel is projected said coordinate system, said pixel in said other view corresponding to a projection of said first point in said other view,
d) remove said pixel in said other view when the first parameter representative of a color as well as the first set of coordinates of the first point corresponds to the second parameter representative of a color as well as the second set of coordinates of the second point. The method and the apparatus wherein the obtaining b), the obtaining c) and the removing d) are performed for all the pixels of the reference view.
The method and the apparatus wherein the obtaining c) and the removing d) are performed for all the views of the multi-view plus depth content.
The method and the apparatus comprising selecting a second reference view in the multi-view plus depth content on which the removing d) was performed, the method and the apparatus further comprising: e) obtaining a third parameter representative of a color as well as a third set of coordinates of a third point corresponding to a pixel in said second reference view when said pixel is projected in said coordinate system,
f) obtaining a fourth parameter representative of a color as well as a fourth set of coordinates of a fourth point corresponding to a pixel of another view of said multi-view plus depth content on which the removing d) was performed when said pixel is projected said coordinate system, said pixel in said other view corresponding to a projection of said third point in said other view, g) removing said pixel in said other view when the third parameter representative of a color as well as the third set of coordinates of the third point corresponds to the fourth parameter representative of a color as well as the fourth set of coordinates of the fourth point.
A computer program characterized in that it comprises program code instructions for the implementation of the method according to claim 1 when the program is executed by a processor.
Although the present invention has been described hereinabove with reference to specific embodiments, the present invention is not limited to the specific embodiments, and modifications will be apparent to a skilled person in the art which lie within the scope of the present invention.
Many further modifications and variations will suggest themselves to those versed in the art upon making reference to the foregoing illustrative embodiments, which are given by way of example only and which are not intended to limit the scope of the invention, that being determined solely by the appended claims. In particular the different features from different embodiments may be interchanged, where appropriate.

Claims

1. A method comprising:
a) selecting a reference view in multi-view plus a depth content,
b) obtaining a first parameter representative of a color including a first set of coordinates of a first point corresponding to a first pixel in said reference view; wherein said first pixel is projected in a coordinate system,
c) obtaining a second parameter representative of a color including a second set of coordinates of a second point corresponding to a pixel of another view of said multi-view plus depth content; wherein said pixel is projected in said coordinate system, said pixel in said other view corresponding to a projection of said first point in said other view,
d) removing said pixel from said other view when the first parameter representative of a color as well as the first set of coordinates of the first point corresponds to the second parameter representative of a color; said second parameter including the second set of coordinates of the second point.
2. An apparatus comprising at least a processor configured to: a) select a reference view in multi-view plus depth content,
b) obtain a first parameter representative of a color as well as a first set of coordinates of a first point corresponding to a pixel in said reference view when said first pixel is projected in a coordinate system,
c) obtain a second parameter representative of a color as well as a second set of coordinates of a second point corresponding to a pixel of another view of said multi-view plusdepth content when said pixel is projected in said coordinate system, said pixel in said other view corresponding to a projection of said first point in said other view, d) remove said pixel in said other view when the first parameter representative of a color as well as the first set of coordinates of the first point corresponds to the second parameter representative of a color as well as the second set of coordinates of the second point.
3. The method of claim 1 or the apparatus according to claim 2 wherein the obtaining b), the obtaining c) and the removing d) are performed for all the pixels of the reference view.
4. The method or the apparatus according to claim 3 wherein the obtaining c) and the removing d) are performed for all the views of the multi-view plus depth content.
5. The method according to claim 4 comprising or the apparatus according to claim 4 configured for selecting a second reference view in the multi-view plus depth content on which the removing d) was performed.
6. The method according to claim 5 comprising at least one of the following steps or the apparatus according to claim 4 configured for performing at least one of the following steps: e) obtaining a third parameter representative of a color including a third set of coordinates of a third point corresponding to a pixel in said second reference view when said pixel is projected in said coordinate system,
f) obtaining a fourth parameter representative of a color including a fourth set of coordinates of a fourth point corresponding to a pixel of another view of said multi-view plus depth content on which the removing d) was performed when said pixel is projected said coordinate system, said pixel in said other view corresponding to a projection of said third point in said other view, g) removing said pixel in said other view when the third parameter representative of a color as well as the third set of coordinates of the third point corresponds to the fourth parameter representative of a color as well as the fourth set of coordinates of the fourth point.
7. The method of claim 1 or apparatus of claim 2 wherein a projection view is obtained among a set of views of the multi-view capture plus depth content that is different from the first reference view.
8. The method or the apparatus of claim 7, wherein the projection view is calculated as a point that does not correspond to an actual pixel but as a calculated approximation of a pixel calculated by using values of a plurality of pixels.
9. The method or the apparatus of claim 7 wherein a projection pixel is selected in the said projection view that reduces a distance between the result of a point Pw in the projection view and one of a neighboring pixel.
10. The method or the apparatus of claim 9 wherein said distance is calculated as the square root of the sum of square difference between said coordinates of said pixels.
11. A method of decoding data to recover a multi-view capture having depth content comprising: a) determining according to data which reference view in said multi-view capture having a depth content is used to generate said data,
b) obtaining a first parameter representative of a color as well as a first set of coordinates of a first point corresponding to a pixel in said reference view; wherein said first pixel is projected in a coordinate system,
c) obtaining a second parameter representative of a color including a second set of coordinates of a second point corresponding to a pixel of another view of said multi-view capture having depth content; wherein said pixel is projected in said coordinate system, said pixel in said other view corresponding to a projection of said first point in said other view, a) generating from said data additional pixels needed to reconstitute said first view, wherein said first view includes a first parameter representative of a color using a first set of coordinates; and a related second parameter representative of a color having a second set of coordinates of a second point corresponding;
b) generating additional pixels and adding them to said other view when the first parameter representative of a color as well as the first set of coordinates of the first point corresponds to the second parameter representative of a color said second parameter including the second set of coordinates of the second point.
12. An apparatus for decoding data to recover a multi-view capture having depth content comprising at least one processor configured to perform: determining according to data which reference view in said multi-view capture having a depth content is used to generate said data, obtaining a first parameter representative of a color as well as a first set of coordinates of a first point corresponding to a pixel in said reference view; wherein said first pixel is projected in a coordinate system,
obtaining a second parameter representative of a color including a second set of coordinates of a second point corresponding to a pixel of another view of said multi-view capture having depth content; wherein said pixel is projected in said coordinate system, said pixel in said other view corresponding to a projection of said first point in said other view, generating from said data additional pixels needed to reconstitute said first view, wherein said first view includes a first parameter representative of a color using a first set of coordinates; and a related second parameter representative of a color having a second set of coordinates of a second point corresponding;
generating additional pixels and adding them to said other view when the first parameter representative of a color as well as the first set of coordinates of the first point corresponds to the second parameter representative of a color said second parameter including the second set of coordinates of the second point.
13. A computer program comprising program code instructions for the implementation of the method according to any of claims 1 and 3 to 1 1 when the program is executed by a processor.
PCT/EP2019/061357 2018-05-04 2019-05-03 A method and an apparatus for reducing an amount of data representative of a multi-view plus depth content Ceased WO2019211429A1 (en)

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
EP18305564.9A EP3565253A1 (en) 2018-05-04 2018-05-04 A method and an apparatus for reducing an amount of data representative of a multi-view plus depth content
EP18305564.9 2018-05-04

Publications (1)

Publication Number Publication Date
WO2019211429A1 true WO2019211429A1 (en) 2019-11-07

Family

ID=62217901

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/EP2019/061357 Ceased WO2019211429A1 (en) 2018-05-04 2019-05-03 A method and an apparatus for reducing an amount of data representative of a multi-view plus depth content

Country Status (2)

Country Link
EP (1) EP3565253A1 (en)
WO (1) WO2019211429A1 (en)

Cited By (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2023507815A (en) * 2019-12-19 2023-02-27 インターデジタル ヴイシー ホールディングス, インコーポレイテッド Encoding and decoding method and apparatus

Non-Patent Citations (1)

* Cited by examiner, † Cited by third party
Title
KIRAN NANJUNDA IYER ET AL: "Multiview video coding using depth based 3D warping", MULTIMEDIA AND EXPO (ICME), 2010 IEEE INTERNATIONAL CONFERENCE ON, IEEE, PISCATAWAY, NJ, USA, 19 July 2010 (2010-07-19), pages 1108 - 1113, XP031761757, ISBN: 978-1-4244-7491-2 *

Cited By (5)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
JP2023507815A (en) * 2019-12-19 2023-02-27 インターデジタル ヴイシー ホールディングス, インコーポレイテッド Encoding and decoding method and apparatus
JP7494303B2 (en) 2019-12-19 2024-06-03 インターデジタル ヴイシー ホールディングス, インコーポレイテッド Encoding and decoding method and device
JP2024112926A (en) * 2019-12-19 2024-08-21 インターデジタル ヴイシー ホールディングス, インコーポレイテッド Encoding and decoding method and device
US12301871B2 (en) 2019-12-19 2025-05-13 Interdigital Vc Holdings, Inc. Encoding and decoding methods and apparatus
JP7717902B2 (en) 2019-12-19 2025-08-04 インターデジタル ヴイシー ホールディングス, インコーポレイテッド Encoding and decoding method and apparatus

Also Published As

Publication number Publication date
EP3565253A1 (en) 2019-11-06

Similar Documents

Publication Publication Date Title
US20150243048A1 (en) System, method, and computer program product for performing one-dimesional searches in two-dimensional images
CN106447607B (en) A kind of image split-joint method and device
CN111738923A (en) Image processing method, device and storage medium
TW200903347A (en) Edge mapping using panchromatic pixels
WO2018068420A1 (en) Image processing method and apparatus
US11074742B2 (en) Image processing apparatus, image processing method, and storage medium
CN113298187B (en) Image processing method and device and computer readable storage medium
CN111127303A (en) Background blurring method and device, terminal equipment and computer readable storage medium
CN112106352A (en) Image processing method and device
WO2018102880A1 (en) Systems and methods for replacing faces in videos
US9406140B2 (en) Method and apparatus for generating depth information
CN110599586A (en) Semi-dense scene reconstruction method and device, electronic equipment and storage medium
US20150043807A1 (en) Depth image compression and decompression utilizing depth and amplitude data
CN114119701A (en) Image processing method and device
US20240202886A1 (en) Video processing method and apparatus, device, storage medium, and program product
US10375368B2 (en) Image data conversion
WO2020000333A1 (en) Image processing method and apparatus
CN104756151A (en) System and method to enhance and process a digital image
JP6221333B2 (en) Image processing apparatus, image processing circuit, and image processing method
CN107169930A (en) Image processing method and device
EP3565253A1 (en) A method and an apparatus for reducing an amount of data representative of a multi-view plus depth content
US12387342B2 (en) Image processing method, image processing device, and recording medium for obtaining a composite image by synthesizing a high-frequency image and one or more color images
CN113824894A (en) Exposure control method, device, equipment and storage medium
WO2017209213A1 (en) Image processing device, image processing method, and computer-readable recording medium
CN110785772A (en) An image processing method, device, system and storage medium

Legal Events

Date Code Title Description
121 Ep: the epo has been informed by wipo that ep was designated in this application

Ref document number: 19720669

Country of ref document: EP

Kind code of ref document: A1

NENP Non-entry into the national phase

Ref country code: DE

122 Ep: pct application non-entry in european phase

Ref document number: 19720669

Country of ref document: EP

Kind code of ref document: A1