International Conference on 3D Vision - 3D Lip Event Detection via Interframe Motion Divergence at Multiple Temporal Resolutions

3D Lip Event Detection via Interframe Motion Divergence at Multiple Temporal Resolutions
Authors: Jie Zhang and Robert B Fisher
Abstract: The lip is a dominant dynamic facial unit when a person is speaking. Detecting lip events is beneficial to speech analysis and support for the hearing impaired. This paper proposes a 3D lip event detection pipeline that automatically determines the lip events from a 3D speaking lip sequence. We define a motion divergence measure using 3D lip landmarks to quantify the interframe dynamics of a 3D speaking lip. Then, we cast the interframe motion detection in a multi-temporal-resolution framework that allows the detection to be applicable to different speaking speeds. The experiments on the S3DFM Dataset investigate the overall 3D lip dynamics based on the proposed motion divergence. The proposed 3D pipeline is able to detect opening and closing lip events across 100 sequences, achieving a state-of-the-art performance.
PDF (protected)

Paper registration

July 23 30, 2021

Paper submission

July 30, 2021

Supplementary

August 8, 2021

Tutorial submission

August 15, 2021

Tutorial notification

August 31, 2021

Rebuttal period

September 16-22, 2021

Paper notification

October 1, 2021

Camera ready

October 15, 2021

Demo submission

~~July 30~~ Nov 15, 2021

Demo notification

~~Oct 1~~ Nov 19, 2021

Tutorial

November 30, 2021

Main conference

December 1-3, 2021