Search CORE

30 research outputs found

A community-powered search of machine learning strategy space to find NMR property prediction models

Author: Anderson Brandon
Bai Shaojie
Bratholm Lars A.
Butts Craig P.
Choi Sunghwan
Dang Lam
Gerrard Will
Glowacki David R.
Hanchar Pavel
Howard Addison
Huard Guillaume
Kim Sanghoon
Kolter Zico
Kondor Risi
Kornbluth Mordechai
Lee Youhan
Lee Youngsoo
Mailoa Jonathan P.
Nguyen Thanh Tu
participants Kaggle
Popovic Milos
Rakocevic Goran
Reade Walter
Song Wonho
Stojanovic Luka
Thiede Erik H.
Tijanic Nebojsa
Torrubia Andres
Willmott Devin
Publication venue: 'Public Library of Science (PLoS)'
Publication date: 13/08/2020
Field of study

The rise of machine learning (ML) has created an explosion in the potential strategies for using data to make scientific predictions. For physical scientists wishing to apply ML strategies to a particular domain, it can be difficult to assess in advance what strategy to adopt within a vast space of possibilities. Here we outline the results of an online community-powered effort to swarm search the space of ML strategies and develop algorithms for predicting atomic-pairwise nuclear magnetic resonance (NMR) properties in molecules. Using an open-source dataset, we worked with Kaggle to design and host a 3-month competition which received 47,800 ML model predictions from 2,700 teams in 84 countries. Within 3 weeks, the Kaggle community produced models with comparable accuracy to our best previously published "in-house" efforts. A meta-ensemble model constructed as a linear combination of the top predictions has a prediction accuracy which exceeds that of any individual model, 7-19x better than our previous state-of-the-art. The results highlight the potential of transformer architectures for predicting quantum mechanical (QM) molecular properties

arXiv.org e-Print Archive

Explore Bristol Research

Distributed deep learning networks among institutions for medical imaging

Author: Andrew Beers
Bruce Rosen
Carson Lam
Chang
Chollet
Collobert
Daniel L Rubin
Darvin Yi
Dean
Dluhoš
Esteva
Glorot
Graham
Gulshan
Hansen
He
Hinton
James Brown
Jayashree Kalpathy-Cramer
Kaggle
Ken Chang
Krizhevsky
LeCun
Miotto
Niranjan Balachandar
Pan
Quellec
Russakovsky
Samala
Su
Theano Development Team
USF Digital Mammography
Xia
Publication venue: 'Oxford University Press (OUP)'
Publication date: 31/08/2018
Field of study

Objective Deep learning has become a promising approach for automated support for clinical diagnosis. When medical data samples are limited, collaboration among multiple institutions is necessary to achieve high algorithm performance. However, sharing patient data often has limitations due to technical, legal, or ethical concerns. In this study, we propose methods of distributing deep learning models as an attractive alternative to sharing patient data. Methods We simulate the distribution of deep learning models across 4 institutions using various training heuristics and compare the results with a deep learning model trained on centrally hosted patient data. The training heuristics investigated include ensembling single institution models, single weight transfer, and cyclical weight transfer. We evaluated these approaches for image classification in 3 independent image collections (retinal fundus photos, mammography, and ImageNet). Results We find that cyclical weight transfer resulted in a performance that was comparable to that of centrally hosted patient data. We also found that there is an improvement in the performance of cyclical weight transfer heuristic with a high frequency of weight transfer. Conclusions We show that distributing deep learning models is an effective alternative to sharing patient data. This finding has implications for any collaborative deep learning study

University of Lincoln Institutional Repository

Crossref

glaucoma image categorization

Author: glaucoma dataset kaggle
Publication venue: Zenodo
Publication date: 12/11/2023
Field of study

<p>glaucoma image categorization</p&gt

ZENODO

Jobs Dataset - Normalised Location

Author: Kaggle Charles C Newey
Publication venue
Publication date: 13/02/2017
Field of study

EdShare

Kernel Based Online Change Point Detection

Author: aminikhanghahi
basseville
inclan
kaggle
richard
rojo-álvarez
schölkopf
Publication venue: 'Institute of Electrical and Electronics Engineers (IEEE)'
Publication date: 02/09/2019
Field of study