-
Invariant Feature Regularization for Fair Face Recognition
Authors:
Jiali Ma,
Zhongqi Yue,
Kagaya Tomoyuki,
Suzuki Tomoki,
Karlekar Jayashree,
Sugiri Pranata,
Hanwang Zhang
Abstract:
Fair face recognition is all about learning invariant feature that generalizes to unseen faces in any demographic group. Unfortunately, face datasets inevitably capture the imbalanced demographic attributes that are ubiquitous in real-world observations, and the model learns biased feature that generalizes poorly in the minority group. We point out that the bias arises due to the confounding demog…
▽ More
Fair face recognition is all about learning invariant feature that generalizes to unseen faces in any demographic group. Unfortunately, face datasets inevitably capture the imbalanced demographic attributes that are ubiquitous in real-world observations, and the model learns biased feature that generalizes poorly in the minority group. We point out that the bias arises due to the confounding demographic attributes, which mislead the model to capture the spurious demographic-specific feature. The confounding effect can only be removed by causal intervention, which requires the confounder annotations. However, such annotations can be prohibitively expensive due to the diversity of the demographic attributes. To tackle this, we propose to generate diverse data partitions iteratively in an unsupervised fashion. Each data partition acts as a self-annotated confounder, enabling our Invariant Feature Regularization (INV-REG) to deconfound. INV-REG is orthogonal to existing methods, and combining INV-REG with two strong baselines (Arcface and CIFP) leads to new state-of-the-art that improves face recognition on a variety of demographic groups. Code is available at https://github.com/PanasonicConnect/InvReg.
△ Less
Submitted 23 October, 2023;
originally announced October 2023.
-
Equivariance and Invariance Inductive Bias for Learning from Insufficient Data
Authors:
Tan Wang,
Qianru Sun,
Sugiri Pranata,
Karlekar Jayashree,
Hanwang Zhang
Abstract:
We are interested in learning robust models from insufficient data, without the need for any externally pre-trained checkpoints. First, compared to sufficient data, we show why insufficient data renders the model more easily biased to the limited training environments that are usually different from testing. For example, if all the training swan samples are "white", the model may wrongly use the "…
▽ More
We are interested in learning robust models from insufficient data, without the need for any externally pre-trained checkpoints. First, compared to sufficient data, we show why insufficient data renders the model more easily biased to the limited training environments that are usually different from testing. For example, if all the training swan samples are "white", the model may wrongly use the "white" environment to represent the intrinsic class swan. Then, we justify that equivariance inductive bias can retain the class feature while invariance inductive bias can remove the environmental feature, leaving the class feature that generalizes to any environmental changes in testing. To impose them on learning, for equivariance, we demonstrate that any off-the-shelf contrastive-based self-supervised feature learning method can be deployed; for invariance, we propose a class-wise invariant risk minimization (IRM) that efficiently tackles the challenge of missing environmental annotation in conventional IRM. State-of-the-art experimental results on real-world benchmarks (VIPriors, ImageNet100 and NICO) validate the great potential of equivariance and invariance in data-efficient learning. The code is available at https://github.com/Wangt-CN/EqInv
△ Less
Submitted 6 September, 2022; v1 submitted 25 July, 2022;
originally announced July 2022.
-
Temporo-Spatial Collaborative Filtering for Parameter Estimation in Noisy DCE-MRI Sequences: Application to Breast Cancer Chemotherapy Response
Authors:
Xia Zhu,
Dipanjan Sengupta,
Andrew Beers,
Kalpathy-Cramer Jayashree,
Theodore L. Willke
Abstract:
Dynamic contrast-enhanced magnetic resonance imaging (DCE-MRI) is a minimally invasive imaging technique which can be used for characterizing tumor biology and tumor response to radiotherapy. Pharmacokinetic (PK) estimation is widely used for DCE-MRI data analysis to extract quantitative parameters relating to microvascu- lature characteristics of the cancerous tissues. Unavoidable noise corruptio…
▽ More
Dynamic contrast-enhanced magnetic resonance imaging (DCE-MRI) is a minimally invasive imaging technique which can be used for characterizing tumor biology and tumor response to radiotherapy. Pharmacokinetic (PK) estimation is widely used for DCE-MRI data analysis to extract quantitative parameters relating to microvascu- lature characteristics of the cancerous tissues. Unavoidable noise corruption during DCE-MRI data acquisition has a large effect on the accuracy of PK estimation. In this paper, we propose a general denoising paradigm called gather- noise attenuation and reduce (GNR) and a novel temporal-spatial collaborative filtering (TSCF) denoising technique for DCE-MRI data. TSCF takes advantage of temporal correlation in DCE-MRI, as well as anatomical spatial similar- ity to collaboratively filter noisy DCE-MRI data. The proposed TSCF denoising algorithm decreases the PK parameter normalized estimation error by 57% and improves the structural similarity of PK parameter estimation by 86% com- pared to baseline without denoising, while being an order of magnitude faster than state-of-the-art denoising methods. TSCF improves the univariate linear regression (ULR) c-statistic value for early prediction of pathologic response up to 18%, and shows complete separation of pathologic complete response (pCR) and non-pCR groups on a challenge dataset.
△ Less
Submitted 2 March, 2018;
originally announced March 2018.
-
Neural Person Search Machines
Authors:
Hao Liu,
Jiashi Feng,
Zequn Jie,
Karlekar Jayashree,
Bo Zhao,
Meibin Qi,
Jianguo Jiang,
Shuicheng Yan
Abstract:
We investigate the problem of person search in the wild in this work. Instead of comparing the query against all candidate regions generated in a query-blind manner, we propose to recursively shrink the search area from the whole image till achieving precise localization of the target person, by fully exploiting information from the query and contextual cues in every recursive search step. We deve…
▽ More
We investigate the problem of person search in the wild in this work. Instead of comparing the query against all candidate regions generated in a query-blind manner, we propose to recursively shrink the search area from the whole image till achieving precise localization of the target person, by fully exploiting information from the query and contextual cues in every recursive search step. We develop the Neural Person Search Machines (NPSM) to implement such recursive localization for person search. Benefiting from its neural search mechanism, NPSM is able to selectively shrink its focus from a loose region to a tighter one containing the target automatically. In this process, NPSM employs an internal primitive memory component to memorize the query representation which modulates the attention and augments its robustness to other distracting regions. Evaluations on two benchmark datasets, CUHK-SYSU Person Search dataset and PRW dataset, have demonstrated that our method can outperform current state-of-the-arts in both mAP and top-1 evaluation protocols.
△ Less
Submitted 21 July, 2017;
originally announced July 2017.
-
Video-based Person Re-identification with Accumulative Motion Context
Authors:
Hao Liu,
Zequn Jie,
Karlekar Jayashree,
Meibin Qi,
Jianguo Jiang,
Shuicheng Yan,
Jiashi Feng
Abstract:
Video based person re-identification plays a central role in realistic security and video surveillance. In this paper we propose a novel Accumulative Motion Context (AMOC) network for addressing this important problem, which effectively exploits the long-range motion context for robustly identifying the same person under challenging conditions. Given a video sequence of the same or different perso…
▽ More
Video based person re-identification plays a central role in realistic security and video surveillance. In this paper we propose a novel Accumulative Motion Context (AMOC) network for addressing this important problem, which effectively exploits the long-range motion context for robustly identifying the same person under challenging conditions. Given a video sequence of the same or different persons, the proposed AMOC network jointly learns appearance representation and motion context from a collection of adjacent frames using a two-stream convolutional architecture. Then AMOC accumulates clues from motion context by recurrent aggregation, allowing effective information flow among adjacent frames and capturing dynamic gist of the persons. The architecture of AMOC is end-to-end trainable and thus motion context can be adapted to complement appearance clues under unfavorable conditions (e.g. occlusions). Extensive experiments are conduced on three public benchmark datasets, i.e., the iLIDS-VID, PRID-2011 and MARS datasets, to investigate the performance of AMOC. The experimental results demonstrate that the proposed AMOC network outperforms state-of-the-arts for video-based re-identification significantly and confirm the advantage of exploiting long-range motion context for video based person re-identification, validating our motivation evidently.
△ Less
Submitted 12 June, 2017; v1 submitted 31 December, 2016;
originally announced January 2017.