Learning Transformations To Reduce the Geometric Shift in Object Detection

The performance of modern object detectors drops when the test distribution differs from the training one. Most of the methods that address this focus on object appearance changes caused by, e.g., different illumination conditions, or gaps between synthetic and real images. Here, by contrast, we tackle geometric shifts emerging from variations in the image capture process, or due to the constraints of the environment causing differences in the apparent geometry of the content itself. We introduce a self-training approach that learns a set of geometric transformations to minimize these shifts without leveraging any labeled data in the new domain, nor any information about the cameras. We evaluate our method on two different shifts, i.e., a camera's field of view (FoV) change and a viewpoint change. Our results evidence that learning geometric transformations helps detectors to perform better in the target domains.

Chat with Graph Search

Ask any question about EPFL courses, lectures, exercises, research, news, etc. or try the example questions below.

DISCLAIMER: The Graph Chatbot is not programmed to provide explicit or categorical answers to your questions. Rather, it transforms your questions into API requests that are distributed across the various IT services officially administered by EPFL. Its purpose is solely to collect and recommend relevant references to content that you can explore to help you answer your questions.

Learning Transformations To Reduce the Geometric Shift in Object Detection

Graph Chatbot

Chat with Graph Search

Coronal jets identification using Deep Learning as Image and Video Object Detection

CLIP the Gap: A Single Domain Generalization Approach for Object Detection

Rigidity-Aware Detection for 6D Object Pose Estimation

CLIP the Gap: A Single Domain Generalization Approach for Object Detection

Coronal jets identification using Deep Learning as Image and Video Object Detection

Rigidity-Aware Detection for 6D Object Pose Estimation