508 research outputs found

    Deep Residual Learning for Image Recognition

    Full text link
    Deeper neural networks are more difficult to train. We present a residual learning framework to ease the training of networks that are substantially deeper than those used previously. We explicitly reformulate the layers as learning residual functions with reference to the layer inputs, instead of learning unreferenced functions. We provide comprehensive empirical evidence showing that these residual networks are easier to optimize, and can gain accuracy from considerably increased depth. On the ImageNet dataset we evaluate residual nets with a depth of up to 152 layers---8x deeper than VGG nets but still having lower complexity. An ensemble of these residual nets achieves 3.57% error on the ImageNet test set. This result won the 1st place on the ILSVRC 2015 classification task. We also present analysis on CIFAR-10 with 100 and 1000 layers. The depth of representations is of central importance for many visual recognition tasks. Solely due to our extremely deep representations, we obtain a 28% relative improvement on the COCO object detection dataset. Deep residual nets are foundations of our submissions to ILSVRC & COCO 2015 competitions, where we also won the 1st places on the tasks of ImageNet detection, ImageNet localization, COCO detection, and COCO segmentation.Comment: Tech repor

    Spatial Pyramid Pooling in Deep Convolutional Networks for Visual Recognition

    Full text link
    Existing deep convolutional neural networks (CNNs) require a fixed-size (e.g., 224x224) input image. This requirement is "artificial" and may reduce the recognition accuracy for the images or sub-images of an arbitrary size/scale. In this work, we equip the networks with another pooling strategy, "spatial pyramid pooling", to eliminate the above requirement. The new network structure, called SPP-net, can generate a fixed-length representation regardless of image size/scale. Pyramid pooling is also robust to object deformations. With these advantages, SPP-net should in general improve all CNN-based image classification methods. On the ImageNet 2012 dataset, we demonstrate that SPP-net boosts the accuracy of a variety of CNN architectures despite their different designs. On the Pascal VOC 2007 and Caltech101 datasets, SPP-net achieves state-of-the-art classification results using a single full-image representation and no fine-tuning. The power of SPP-net is also significant in object detection. Using SPP-net, we compute the feature maps from the entire image only once, and then pool features in arbitrary regions (sub-images) to generate fixed-length representations for training the detectors. This method avoids repeatedly computing the convolutional features. In processing test images, our method is 24-102x faster than the R-CNN method, while achieving better or comparable accuracy on Pascal VOC 2007. In ImageNet Large Scale Visual Recognition Challenge (ILSVRC) 2014, our methods rank #2 in object detection and #3 in image classification among all 38 teams. This manuscript also introduces the improvement made for this competition.Comment: This manuscript is the accepted version for IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) 2015. See Changelo

    In-fiber linear polarizer based on UV-inscribed 45° tilted grating in polarization maintaining fiber

    Get PDF
    We report an in-fiber linear polarizer structured by UV-inscribing a 45° tilted fiber grating (TFG) into polarization maintaining (PM) fiber along its principal axis. The polarization extinction ratio (PER) achieved by a 48 mm long 45° TFG has reached 46 dB at 1550 nm and the overall PER is >40 dB over a 50 nm wavelength range. Such 45° TFG based polarizers have many advantages over conventional products, including low loss, low cost, simple fabrication process, and no physical modification to the fiber, thus offering high stability and capable of handling high power

    Highly sensitive temperature and strain sensors based on all-fiber 45°-TFG Lyot filter

    Get PDF
    We demonstrate highly sensitive temperature and strain sensors based on an all-fiber Lyot filter structure, which is formed by concatenating two 45°-TFGs (tilted fiber gratings) with a PM fiber cavity. The experiment results show the all-fiber 45°-TFG Lyot filter has very high sensitivity to strain and temperature. The 45°-TFG Lyot filters of two different cavity lengths (18cm and 40 cm) have been evaluated for temperature sensing by heating a section of the cavity from 10°C to 50°C. The experiment results have shown remarkably high temperature sensitivities of 0.616nm/°C for 18cm and 0.31nm/°C for 40cm long cavity filter, respectively. The 18cm long cavity filter has been subjected to strain variations up to around 550μ ε and the filter has exhibited strain sensitivities of 0.02499nm/μ ε and 0.012nm/μ ε for two straining situations, where its cavity middle section of 18cm and 9cm were stretched, respectively

    Refractometer based on fiber Bragg grating Fabry-Pérot cavity embedded with a narrow microchannel

    Get PDF
    We report on inscription of microchannels of different widths in optical fiber using femtosecond (fs) laser inscription assisted chemical etching and the narrowest channel has been created with a width down to only 1.2µm. Microchannels with 5µm and 35µm widths were fabricated together with Fabry-Pérot (FP) cavities formed by UV laser written fiber Bragg gratings (FBGs), creating high function and linear response refractometers. The device with a 5µm microchannel has exhibited a refractive index (RI) detection range up to 1.7, significantly higher than all fiber grating RI sensors. In addition, the microchannel FBG FP structures have been theoretically simulated showing excellent agreement with experimental measured characteristics
    • …
    corecore