Search CORE

20 research outputs found

Essentials of Business Analytics

Author: Bhimasankaram PochirajuSridhar Seshadri
Publication venue: 'Springer Fachmedien Wiesbaden GmbH'
Publication date: 22/04/2020
Field of study

Emerging approaches for data-driven innovation in Europe: Sandbox experiments on the governance of data and technology

Author: Dencik Lina
Granell Carlos
Jirka Simon
Kotsev Alexander
MICHELI MARINA
Minghini Marco
Mooney Peter
Oost H.
Ostermann Frank
Rieke M.
Sarretta Alessandro
Schade Sven
Van Den Broecke J.
Verhulst Stefaan
Publication venue: 'Publications Office of the European Union'
Publication date: 16/02/2022
Field of study

Europe’s digital transformation of the economy and society is one of the priorities of the current Commission and is framed by the European strategy for data. This strategy aims at creating a single market for data through the establishment of a common European data space, based in turn on domain-specific data spaces in strategic sectors such as environment, agriculture, industry, health and transportation. Acknowledging the key role that emerging technologies and innovative approaches for data sharing and use can play to make European data spaces a reality, this document presents a set of experiments that explore emerging technologies and tools for data-driven innovation, and also deepen in the socio-technical factors and forces that occur in data-driven innovation. Experimental results shed some light in terms of lessons learned and practical recommendations towards the establishment of European data spaces

Repositori Institucional de la Universitat Jaume I

Processing data close to its origin - edge computing on IoT devices to detect noise pollution

Author: Ostermann F.O.
Publication venue
Publication date: 01/01/2022
Field of study

University of Twente Research Information

Metadata-driven data integration

Author: Nadal Francesch Sergi
Publication venue: Universitat Politècnica de Catalunya
Publication date: 16/05/2019
Field of study

Cotutela: Universitat Politècnica de Catalunya i Université Libre de Bruxelles, IT4BI-DC programme for the joint Ph.D. degree in computer science.Data has an undoubtable impact on society. Storing and processing large amounts of available data is currently one of the key success factors for an organization. Nonetheless, we are recently witnessing a change represented by huge and heterogeneous amounts of data. Indeed, 90% of the data in the world has been generated in the last two years. Thus, in order to carry on these data exploitation tasks, organizations must first perform data integration combining data from multiple sources to yield a unified view over them. Yet, the integration of massive and heterogeneous amounts of data requires revisiting the traditional integration assumptions to cope with the new requirements posed by such data-intensive settings. This PhD thesis aims to provide a novel framework for data integration in the context of data-intensive ecosystems, which entails dealing with vast amounts of heterogeneous data, from multiple sources and in their original format. To this end, we advocate for an integration process consisting of sequential activities governed by a semantic layer, implemented via a shared repository of metadata. From an stewardship perspective, this activities are the deployment of a data integration architecture, followed by the population of such shared metadata. From a data consumption perspective, the activities are virtual and materialized data integration, the former an exploratory task and the latter a consolidation one. Following the proposed framework, we focus on providing contributions to each of the four activities. We begin proposing a software reference architecture for semantic-aware data-intensive systems. Such architecture serves as a blueprint to deploy a stack of systems, its core being the metadata repository. Next, we propose a graph-based metadata model as formalism for metadata management. We focus on supporting schema and data source evolution, a predominant factor on the heterogeneous sources at hand. For virtual integration, we propose query rewriting algorithms that rely on the previously proposed metadata model. We additionally consider semantic heterogeneities in the data sources, which the proposed algorithms are capable of automatically resolving. Finally, the thesis focuses on the materialized integration activity, and to this end, proposes a method to select intermediate results to materialize in data-intensive flows. Overall, the results of this thesis serve as contribution to the field of data integration in contemporary data-intensive ecosystems.Les dades tenen un impacte indubtable en la societat. La capacitat d’emmagatzemar i processar grans quantitats de dades disponibles és avui en dia un dels factors claus per l’èxit d’una organització. No obstant, avui en dia estem presenciant un canvi representat per grans volums de dades heterogenis. En efecte, el 90% de les dades mundials han sigut generades en els últims dos anys. Per tal de dur a terme aquestes tasques d’explotació de dades, les organitzacions primer han de realitzar una integració de les dades, combinantles a partir de diferents fonts amb l’objectiu de tenir-ne una vista unificada d’elles. Per això, aquest fet requereix reconsiderar les assumpcions tradicionals en integració amb l’objectiu de lidiar amb els requisits imposats per aquests sistemes de tractament massiu de dades. Aquesta tesi doctoral té com a objectiu proporcional un nou marc de treball per a la integració de dades en el context de sistemes de tractament massiu de dades, el qual implica lidiar amb una gran quantitat de dades heterogènies, provinents de múltiples fonts i en el seu format original. Per això, proposem un procés d’integració compost d’una seqüència d’activitats governades per una capa semàntica, la qual és implementada a partir d’un repositori de metadades compartides. Des d’una perspectiva d’administració, aquestes activitats són el desplegament d’una arquitectura d’integració de dades, seguit per la inserció d’aquestes metadades compartides. Des d’una perspectiva de consum de dades, les activitats són la integració virtual i materialització de les dades, la primera sent una tasca exploratòria i la segona una de consolidació. Seguint el marc de treball proposat, ens centrem en proporcionar contribucions a cada una de les quatre activitats. La tesi inicia proposant una arquitectura de referència de software per a sistemes de tractament massiu de dades amb coneixement semàntic. Aquesta arquitectura serveix com a planell per a desplegar un conjunt de sistemes, sent el repositori de metadades al seu nucli. Posteriorment, proposem un model basat en grafs per a la gestió de metadades. Concretament, ens centrem en donar suport a l’evolució d’esquemes i fonts de dades, un dels factors predominants en les fonts de dades heterogènies considerades. Per a l’integració virtual, proposem algorismes de rescriptura de consultes que usen el model de metadades previament proposat. Com a afegitó, considerem heterogeneïtat semàntica en les fonts de dades, les quals els algorismes de rescriptura poden resoldre automàticament. Finalment, la tesi es centra en l’activitat d’integració materialitzada. Per això proposa un mètode per a seleccionar els resultats intermedis a materialitzar un fluxes de tractament intensiu de dades. En general, els resultats d’aquesta tesi serveixen com a contribució al camp d’integració de dades en els ecosistemes de tractament massiu de dades contemporanisLes données ont un impact indéniable sur la société. Le stockage et le traitement de grandes quantités de données disponibles constituent actuellement l’un des facteurs clés de succès d’une entreprise. Néanmoins, nous assistons récemment à un changement représenté par des quantités de données massives et hétérogènes. En effet, 90% des données dans le monde ont été générées au cours des deux dernières années. Ainsi, pour mener à bien ces tâches d’exploitation des données, les organisations doivent d’abord réaliser une intégration des données en combinant des données provenant de sources multiples pour obtenir une vue unifiée de ces dernières. Cependant, l’intégration de quantités de données massives et hétérogènes nécessite de revoir les hypothèses d’intégration traditionnelles afin de faire face aux nouvelles exigences posées par les systèmes de gestion de données massives. Cette thèse de doctorat a pour objectif de fournir un nouveau cadre pour l’intégration de données dans le contexte d’écosystèmes à forte intensité de données, ce qui implique de traiter de grandes quantités de données hétérogènes, provenant de sources multiples et dans leur format d’origine. À cette fin, nous préconisons un processus d’intégration constitué d’activités séquentielles régies par une couche sémantique, mise en oeuvre via un dépôt partagé de métadonnées. Du point de vue de la gestion, ces activités consistent à déployer une architecture d’intégration de données, suivies de la population de métadonnées partagées. Du point de vue de la consommation de données, les activités sont l’intégration de données virtuelle et matérialisée, la première étant une tâche exploratoire et la seconde, une tâche de consolidation. Conformément au cadre proposé, nous nous attachons à fournir des contributions à chacune des quatre activités. Nous commençons par proposer une architecture logicielle de référence pour les systèmes de gestion de données massives et à connaissance sémantique. Une telle architecture consiste en un schéma directeur pour le déploiement d’une pile de systèmes, le dépôt de métadonnées étant son composant principal. Ensuite, nous proposons un modèle de métadonnées basé sur des graphes comme formalisme pour la gestion des métadonnées. Nous mettons l’accent sur la prise en charge de l’évolution des schémas et des sources de données, facteur prédominant des sources hétérogènes sous-jacentes. Pour l’intégration virtuelle, nous proposons des algorithmes de réécriture de requêtes qui s’appuient sur le modèle de métadonnées proposé précédemment. Nous considérons en outre les hétérogénéités sémantiques dans les sources de données, que les algorithmes proposés sont capables de résoudre automatiquement. Enfin, la thèse se concentre sur l’activité d’intégration matérialisée et propose à cette fin une méthode de sélection de résultats intermédiaires à matérialiser dans des flux des données massives. Dans l’ensemble, les résultats de cette thèse constituent une contribution au domaine de l’intégration des données dans les écosystèmes contemporains de gestion de données massivesPostprint (published version

UPCommons. Portal del coneixement obert de la UPC

Towards Interoperable Research Infrastructures for Environmental and Earth Sciences

Author
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 10/02/2021
Field of study

This open access book summarises the latest developments on data management in the EU H2020 ENVRIplus project, which brought together more than 20 environmental and Earth science research infrastructures into a single community. It provides readers with a systematic overview of the common challenges faced by research infrastructures and how a ‘reference model guided’ engineering approach can be used to achieve greater interoperability among such infrastructures in the environmental and earth sciences. The 20 contributions in this book are structured in 5 parts on the design, development, deployment, operation and use of research infrastructures. Part one provides an overview of the state of the art of research infrastructure and relevant e-Infrastructure technologies, part two discusses the reference model guided engineering approach, the third part presents the software and tools developed for common data management challenges, the fourth part demonstrates the software via several use cases, and the last part discusses the sustainability and future directions

Directory of Open Access Books (DOAB)

Towards Interoperable Research Infrastructures for Environmental and Earth Sciences:A Reference Model Guided Approach for Common Challenges

Author
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2020
Field of study

International Migration, Integration and Social Cohesion online publications

Towards Interoperable Research Infrastructures for Environmental and Earth Sciences:A Reference Model Guided Approach for Common Challenges

Author
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2020
Field of study

International Migration, Integration and Social Cohesion online publications

Emerging approaches for data-driven innovation in Europe

Author: Dencik Lina
Granell Carlos
Jirka Simon
Kotsev Alexander
Micheli Marina
Minghini Marco
Mooney Peter
Oost Hillen
Ostermann Frank
Rieke Matthes
Sarretta Alessandro
Schade Sven
Van Den Broecke Just
Verhulst Stefaan
Publication venue: Publications Office of the European Union
Publication date
Field of study

Europe's digital transformation of the economy and society is one of the priorities of the current Commission and is framed by the European strategy for data. This strategy aims at creating a single market for data through the establishment of a common European data space, based in turn on domain-specific data spaces in strategic sectors such as environment, agriculture, industry, health and transportation. Acknowledging the key role that emerging technologies and innovative approaches for data sharing and use can play to make European data spaces a reality, this document presents a set of experiments that explore emerging technologies and tools for data-driven innovation, and also deepen in the socio-technical factors and forces that occur in data-driven innovation. Experimental results shed some light in terms of lessons learned and practical recommendations towards the establishment of European data spaces

Goldsmiths Research Online

The use of social media big data within South African hotels and lodges

Author: Gutfreund Sebastian
Publication venue
Publication date: 01/01/2019
Field of study

Abstract: Big data is a revolutionary and disruptive technology that is used to identify behavioural patterns and track customer preferences. It has several advantages for the hospitality industry, where customer loyalty is integral for brand performance. However, big data is greatly underutilised. Therefore, a study was conducted to look at the use of social media big data within South African hotels and lodges The aim of this study was to focus on the general understanding of big data and the link it shares with social media. There was a further focus on the analytical tools that hotels and lodges make use of, as well as the benefits and challenges which social media big data elucidates for these sectors. This information provides an overall image of how South African hotels and lodges are wielding this technology, giving a future viewpoint on the progression and improvements that need to be undertaken. A comparison concerning the key similarities and differences between the lodge and hotel sector was also provided. This gave an overall picture on how South African hotels and lodges are using this technology, thus giving a future outlook on the progression and improvements that need to be taken into consideration. In order to fully grasp and appreciate big data, a literature review was provided in order to understand the relationship big data has with social media, and the impact it has within the hospitality industry, playing closer attention to hotels and lodges. The methodological approach of the study focused on the qualitative research method, where ten participants in total were interviewed - five being marketing managers in hotels and five marketing managers in lodges. The key findings of the study revealed that the South African hospitality industry is presently only at the genesis when it comes to the use of social media big data. This was revealed through the marketing manager’s generic understanding of the phenomenon. Furthermore, the data predominantly illustrated that only basic analytical tools were used, which indicates that there is a shortage of internal specialists who are capable of handling more advanced tools to further their findings. However, the benefits established were primarily related to the identification of behavioural patterns and preferences of both future and current customers, as well as the marketability of certain promotions that are placed on various platforms. In summary, the data is essentially used to enhance the guests experience through targeting their likes and dislikes. The primary challenges within both sectors of the industry emphasised areas such as education and training, the lack of advanced technology, and the security and privacy concerns pertaining to guest data. ..M.Com. (Tourism and Hospitality Management

University of Johannesburg Institutional Repository

Medical Secretaries’ Registration Work in the Data-Driven Healthcare Era

Author: Bertelsen Pernille Scholdan
Knudsen Casper
Publication venue: IOS Press
Publication date: 01/01/2023
Field of study

VBN