Patent №
US 9,305,600
Granted
2016-04-05
Filed 2013
Owner
PROVOST FELLOWS AND SCHOLARS OF THE COLLEGE OF THE HOLY AND UNDIVIDED TRINITY OF QUEEN ELIZABETH, NEAR DUBLIN
Lab
—
AI components
3
vision · speech · kr
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
13748727
The invention provides a method and system for the automated post production of a single video file, the method comprising the steps of gathering video data from a plurality of camera sources; gathering audio data from a plurality of microphone sources; using an automated tracking offline algorithm to track a sound emitting from a moving target object in a 3D space, to provide localization data of said target object to identify an optimum camera source to provide video data of said target object; and composing a composite video sequence of said moving target from a plurality of identified optimum camera sources in a single video file. The algorithm relies on both video data from multiple camera views and audio data from multiple microphone arrays to infer the 3D position of the active speaker over the duration of the captured presentation.
AI classification
Ownership
PROVOST FELLOWS AND SCHOLARS OF THE COLLEGE OF THE HOLY AND UNDIVIDED TRINITY OF QUEEN ELIZABETH, NEAR DUBLIN
assignment · 297210333
Assignors
BOLAND, FRANK, KOKARAM, ANIL, KELLY, DAMIEN, PITIE, FRANCOIS
On an employer assignment, the assignors are typically the inventors.