METHOD FOR SEGMENTING 3D OBJECTS FROM COMPRESSED VIDEOS

Patent №

US 7,142,602

Granted

2006-11-28

Filed 2003

Owner

MITSUBISHI ELECTRIC RESEARCH LABORATORIES, INC.

Lab

AI components

3

vision · kr · planning

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

10442417

A method segments a video into objects, without user assistance. An MPEG compressed video is converted to a structure called a pseudo spatial/temporal data using DCT coefficients and motion vectors. The compressed video is first parsed and the pseudo spatial/temporal data are formed. Seeds macro-blocks are identified using, e.g., the DCT coefficients and changes in the motion vector of macro-blocks.A video volume is “grown” around each seed macro-block using the DCT coefficients and motion distance criteria. Self-descriptors are assigned to the volume, and mutual descriptors are assigned to pairs of similar volumes. These descriptors capture motion and spatial information of the volumes. Similarity scores are determined for each possible pair-wise combination of volumes. The pair of volumes that gives the largest score is combined iteratively. In the combining stage, volumes are classified and represented in a multi-resolution coarse-to-fine hierarchy of video objects.

VisionKnowledge representationPlanningH04N 19/48G06T 7/11G06T 7/187G06V 10/26H04N 19/87G06T 2207/10016G06T 2207/20048G06T 2207/20101

AI classification

Vision1.00
Planning0.65
Knowledge representation0.62
AI hardware0.18
Machine learning0.17
Natural language0.00
Speech0.00
Evolutionary computation0.00

Ownership

MITSUBISHI ELECTRIC RESEARCH LABORATORIES, INC.

assignment · 141050076

Assignors

PORIKLI, FATIH M., SUN, HUIFANG, DIVAKARAN, AJAY

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC