Multitask AET with Orthogonal Tangent Regularity for Dark Object Detection

About

Dark environment becomes a challenge for computer vision algorithms owing to insufficient photons and undesirable noise. To enhance object detection in a dark environment, we propose a novel multitask auto encoding transformation (MAET) model which is able to explore the intrinsic pattern behind illumination translation. In a self-supervision manner, the MAET learns the intrinsic visual structure by encoding and decoding the realistic illumination-degrading transformation considering the physical noise model and image signal processing (ISP). Based on this representation, we achieve the object detection task by decoding the bounding box coordinates and classes. To avoid the over-entanglement of two tasks, our MAET disentangles the object and degrading features by imposing an orthogonal tangent regularity. This forms a parametric manifold along which multitask predictions can be geometrically formulated by maximizing the orthogonality between the tangents along the outputs of respective tasks. Our framework can be implemented based on the mainstream object detection architecture and directly trained end-to-end using normal target detection datasets, such as VOC and COCO. We have achieved the state-of-the-art performance using synthetic and real-world datasets. Code is available at https://github.com/cuiziteng/MAET.

Ziteng Cui, Guo-Jun Qi, Lin Gu, Shaodi You, Zenghui Zhang, Tatsuya Harada• 2022

Related benchmarks

Task	Dataset	Result
Object Detection	COCO 2017 (val)	AP33	2843
Object Detection	VOC 2007 (test)	AP@5078.8	91
Object Detection	ExDark	mAP (Mean Average Precision)77.7	58
Object Detection	ExDark (test)	mAP (Mean Average Precision)74	47
Face Detection	DARK FACE (test)	mAP0.526	24
Face Detection	UG2+ DARK FACE (test)	mAP0.558	7
Face Detection	DARK FACE (val)	mAP55.8	7
Image Classification	CODaN	Top-1 Acc56.48	4

Showing 8 of 8 rows

Other info

Code

Follow for update

@wizwand_team Discord