WaterScenes: A Multi-Task 4D Radar-Camera Fusion Dataset and Benchmark
  for Autonomous Driving on Water Surfaces

Guan, Runwei; Huang, Zile; Man, Ka Lok; Ni, Yi; Seo, Hyungjoon; Wu, Zhaodong; Yao, Shanliang; Yue, Yong; Yue, Yutao; Zhang, Zixian; Zhu, Xiaohui

WaterScenes: A Multi-Task 4D Radar-Camera Fusion Dataset and Benchmark for Autonomous Driving on Water Surfaces

Authors: Runwei Guan
Zile Huang
Ka Lok Man
Yi Ni
Hyungjoon Seo
Zhaodong Wu
Shanliang Yao
Yong Yue
Yutao Yue
Zixian Zhang
Xiaohui Zhu
Publication date: 12 July 2023
Publisher

Abstract

Autonomous driving on water surfaces plays an essential role in executing hazardous and time-consuming missions, such as maritime surveillance, survivors rescue, environmental monitoring, hydrography mapping and waste cleaning. This work presents WaterScenes, the first multi-task 4D radar-camera fusion dataset for autonomous driving on water surfaces. Equipped with a 4D radar and a monocular camera, our Unmanned Surface Vehicle (USV) proffers all-weather solutions for discerning object-related information, including color, shape, texture, range, velocity, azimuth, and elevation. Focusing on typical static and dynamic objects on water surfaces, we label the camera images and radar point clouds at pixel-level and point-level, respectively. In addition to basic perception tasks, such as object detection, instance segmentation and semantic segmentation, we also provide annotations for free-space segmentation and waterline segmentation. Leveraging the multi-task and multi-modal data, we conduct numerous experiments on the single modality of radar and camera, as well as the fused modalities. Results demonstrate that 4D radar-camera fusion can considerably enhance the robustness of perception on water surfaces, especially in adverse lighting and weather conditions. WaterScenes dataset is public on https://waterscenes.github.io

Similar works

Full text

Available Versions

arXiv.org e-Print Archive

oai:arXiv.org:2307.06505

Last time updated on 20/07/2023