Weakly supervised pain localization using multiple instance learning

Karan Sikka, Abhinav Dhall, Marian Bartlett

    Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

    50 Citations (Scopus)

    Abstract

    Automatic pain recognition from videos is a vital clinical application and, owing to its spontaneous nature, poses interesting challenges to automatic facial expression recognition (AFER) research. Previous pain vs no-pain systems have highlighted two major challenges: (1) ground truth is provided for the sequence, but the presence or absence of the target expression for a given frame is unknown, and (2) the time point and the duration of the pain expression event(s) in each video are unknown. To address these issues we propose a novel framework (referred to as MS-MIL) where each sequence is represented as a bag containing multiple segments, and multiple instance learning (MIL) is employed to handle this weakly labeled data in the form of sequence level ground-truth. These segments are generated via multiple clustering of a sequence or running a multi-scale temporal scanning window, and are represented using a state-of-the-art Bag of Words (BoW) representation. This work extends the idea of detecting facial expressions through 'concept frames' to 'concept segments' and argues through extensive experiments that algorithms like MIL are needed to reap the benefits of such representation. The key advantages of our approach are: (1) joint detection and localization of painful frames using only sequence-level ground-truth, (2) incorporation of temporal dynamics by representing the data not as individual frames but as segments, and (3) extraction of multiple segments, which is well suited to signals with uncertain temporal location and duration in the video. Experiments on UNBC-McMaster Shoulder Pain dataset highlight the effectiveness of our approach by achieving promising results on the problem of pain detection in videos.

    Original languageEnglish
    Title of host publication2013 10th IEEE International Conference and Workshops on Automatic Face and Gesture Recognition, FG 2013
    DOIs
    Publication statusPublished - 2013
    Event2013 10th IEEE International Conference and Workshops on Automatic Face and Gesture Recognition, FG 2013 - Shanghai, China
    Duration: 22 Apr 201326 Apr 2013

    Publication series

    Name2013 10th IEEE International Conference and Workshops on Automatic Face and Gesture Recognition, FG 2013

    Conference

    Conference2013 10th IEEE International Conference and Workshops on Automatic Face and Gesture Recognition, FG 2013
    Country/TerritoryChina
    CityShanghai
    Period22/04/1326/04/13

    Fingerprint

    Dive into the research topics of 'Weakly supervised pain localization using multiple instance learning'. Together they form a unique fingerprint.

    Cite this