Search CORE

3,378 research outputs found

PWC-Net: CNNs for Optical Flow Using Pyramid, Warping, and Cost Volume

Author: Kautz Jan
Liu Ming-Yu
Sun Deqing
Yang Xiaodong
Publication venue
Publication date: 25/06/2018
Field of study

We present a compact but effective CNN model for optical flow, called PWC-Net. PWC-Net has been designed according to simple and well-established principles: pyramidal processing, warping, and the use of a cost volume. Cast in a learnable feature pyramid, PWC-Net uses the cur- rent optical flow estimate to warp the CNN features of the second image. It then uses the warped features and features of the first image to construct a cost volume, which is processed by a CNN to estimate the optical flow. PWC-Net is 17 times smaller in size and easier to train than the recent FlowNet2 model. Moreover, it outperforms all published optical flow methods on the MPI Sintel final pass and KITTI 2015 benchmarks, running at about 35 fps on Sintel resolution (1024x436) images. Our models are available on https://github.com/NVlabs/PWC-Net.Comment: CVPR 2018 camera ready version (with github link to Caffe and PyTorch code

arXiv.org e-Print Archive

Crossref

A Deep Learning Approach to Denoise Optical Coherence Tomography Images of the Optic Nerve Head

Author: Aung Tin
Devalla Sripad Krishna
Girard Michael J. A.
Perera Shamira
Pham Tan Hung
Schmetterer Leopold
Subramanian Giridhar
Thiery Alexandre H.
Tun Tin A.
Wang Xiaofei
Publication venue
Publication date: 27/09/2018
Field of study

Purpose: To develop a deep learning approach to de-noise optical coherence tomography (OCT) B-scans of the optic nerve head (ONH). Methods: Volume scans consisting of 97 horizontal B-scans were acquired through the center of the ONH using a commercial OCT device (Spectralis) for both eyes of 20 subjects. For each eye, single-frame (without signal averaging), and multi-frame (75x signal averaging) volume scans were obtained. A custom deep learning network was then designed and trained with 2,328 "clean B-scans" (multi-frame B-scans), and their corresponding "noisy B-scans" (clean B-scans + gaussian noise) to de-noise the single-frame B-scans. The performance of the de-noising algorithm was assessed qualitatively, and quantitatively on 1,552 B-scans using the signal to noise ratio (SNR), contrast to noise ratio (CNR), and mean structural similarity index metrics (MSSIM). Results: The proposed algorithm successfully denoised unseen single-frame OCT B-scans. The denoised B-scans were qualitatively similar to their corresponding multi-frame B-scans, with enhanced visibility of the ONH tissues. The mean SNR increased from

4.02 \pm 0.68

dB (single-frame) to

8.14 \pm 1.03

dB (denoised). For all the ONH tissues, the mean CNR increased from

3.50 \pm 0.56

(single-frame) to

7.63 \pm 1.81

(denoised). The MSSIM increased from

0.13 \pm 0.02

(single frame) to

0.65 \pm 0.03

(denoised) when compared with the corresponding multi-frame B-scans. Conclusions: Our deep learning algorithm can denoise a single-frame OCT B-scan of the ONH in under 20 ms, thus offering a framework to obtain superior quality OCT B-scans with reduced scanning times and minimal patient discomfort

arXiv.org e-Print Archive

DR-NTU (Digital Repository of NTU)

ScholarBank@NUS