Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Vítor Lourenço

A Modality-level Explainable Framework for Misinformation Checking in Social Networks

Dec 08, 2022

Vítor Lourenço, Aline Paes

Abstract:The widespread of false information is a rising concern worldwide with critical social impact, inspiring the emergence of fact-checking organizations to mitigate misinformation dissemination. However, human-driven verification leads to a time-consuming task and a bottleneck to have checked trustworthy information at the same pace they emerge. Since misinformation relates not only to the content itself but also to other social features, this paper addresses automatic misinformation checking in social networks from a multimodal perspective. Moreover, as simply naming a piece of news as incorrect may not convince the citizen and, even worse, strengthen confirmation bias, the proposal is a modality-level explainable-prone misinformation classifier framework. Our framework comprises a misinformation classifier assisted by explainable methods to generate modality-oriented explainable inferences. Preliminary findings show that the misinformation classifier does benefit from multimodal information encoding and the modality-oriented explainable mechanism increases both inferences' interpretability and completeness.

* Accepted to publication at LatinX in AI workshop at the Thirty-sixth Conference on Neural Information Processing Systems, LXAI @ NeurIPS 2022

Via

Access Paper or Ask Questions

Learning Attention-based Representations from Multiple Patterns for Relation Prediction in Knowledge Graphs

Jun 07, 2022

Vítor Lourenço, Aline Paes

Figure 1 for Learning Attention-based Representations from Multiple Patterns for Relation Prediction in Knowledge Graphs

Figure 2 for Learning Attention-based Representations from Multiple Patterns for Relation Prediction in Knowledge Graphs

Figure 3 for Learning Attention-based Representations from Multiple Patterns for Relation Prediction in Knowledge Graphs

Figure 4 for Learning Attention-based Representations from Multiple Patterns for Relation Prediction in Knowledge Graphs

Abstract:Knowledge bases, and their representations in the form of knowledge graphs (KGs), are naturally incomplete. Since scientific and industrial applications have extensively adopted them, there is a high demand for solutions that complete their information. Several recent works tackle this challenge by learning embeddings for entities and relations, then employing them to predict new relations among the entities. Despite their aggrandizement, most of those methods focus only on the local neighbors of a relation to learn the embeddings. As a result, they may fail to capture the KGs' context information by neglecting long-term dependencies and the propagation of entities' semantics. In this manuscript, we propose {\AE}MP (Attention-based Embeddings from Multiple Patterns), a novel model for learning contextualized representations by: (i) acquiring entities' context information through an attention-enhanced message-passing scheme, which captures the entities' local semantics while focusing on different aspects of their neighborhood; and (ii) capturing the semantic context, by leveraging the paths and their relationships between entities. Our empirical findings draw insights into how attention mechanisms can improve entities' context representation and how combining entities and semantic path contexts improves the general representation of entities and the relation predictions. Experimental results on several large and small knowledge graph benchmarks show that {\AE}MP either outperforms or competes with state-of-the-art relation prediction methods.

* Accepted to publication at Knowledge-based Systems, 2022

Via

Access Paper or Ask Questions

Workflow Provenance in the Lifecycle of Scientific Machine Learning

Sep 30, 2020

Renan Souza, Leonardo G. Azevedo, Vítor Lourenço, Elton Soares, Raphael Thiago, Rafael Brandão, Daniel Civitarese, Emilio Vital Brazil, Marcio Moreno, Patrick Valduriez(+3 more)

Figure 1 for Workflow Provenance in the Lifecycle of Scientific Machine Learning

Figure 2 for Workflow Provenance in the Lifecycle of Scientific Machine Learning

Figure 3 for Workflow Provenance in the Lifecycle of Scientific Machine Learning

Figure 4 for Workflow Provenance in the Lifecycle of Scientific Machine Learning

Abstract:Machine Learning (ML) has already fundamentally changed several businesses. More recently, it has also been profoundly impacting the computational science and engineering domains, like geoscience, climate science, and health science. In these domains, users need to perform comprehensive data analyses combining scientific data and ML models to provide for critical requirements, such as reproducibility, model explainability, and experiment data understanding. However, scientific ML is multidisciplinary, heterogeneous, and affected by the physical constraints of the domain, making such analyses even more challenging. In this work, we leverage workflow provenance techniques to build a holistic view to support the lifecycle of scientific ML. We contribute with (i) characterization of the lifecycle and taxonomy for data analyses; (ii) design principles to build this view, with a W3C PROV compliant data representation and a reference system architecture; and (iii) lessons learned after an evaluation in an Oil & Gas case using an HPC cluster with 393 nodes and 946 GPUs. The experiments show that the principles enable queries that integrate domain semantics with ML models while keeping low overhead (<1%), high scalability, and an order of magnitude of query acceleration under certain workloads against without our representation.

* 21 pages, 10 figures, Under review in a scientific journal since June 30th, 2020. arXiv admin note: text overlap with arXiv:1910.04223

Via

Access Paper or Ask Questions

Managing Machine Learning Workflow Components

Dec 10, 2019

Marcio Moreno, Vítor Lourenço, Sandro Rama Fiorini, Polyana Costa, Rafael Brandão, Daniel Civitarese, Renato Cerqueira

Figure 1 for Managing Machine Learning Workflow Components

Figure 2 for Managing Machine Learning Workflow Components

Abstract:Machine Learning Workflows~(MLWfs) have become essential and a disruptive approach in problem-solving over several industries. However, the development process of MLWfs may be complicated, hard to achieve, time-consuming, and error-prone. To handle this problem, in this paper, we introduce \emph{machine learning workflow management}~(MLWfM) as a technique to aid the development and reuse of MLWfs and their components through three aspects: representation, execution, and creation. More precisely, we discuss our approach to structure the MLWfs' components and their metadata to aid retrieval and reuse of components in new MLWfs. Also, we consider the execution of these components within a tool. The hybrid knowledge representation, called Hyperknowledge, frames our methodology, supporting the three MLWfM's aspects. To validate our approach, we show a practical use case in the Oil \& Gas industry.

* 6 pages, 4 figures, Accepted at the 14th IEEE International Conference on SEMANTIC COMPUTING (ICSC) 2019, San Diego, California

Via

Access Paper or Ask Questions

Provenance Data in the Machine Learning Lifecycle in Computational Science and Engineering

Oct 21, 2019

Renan Souza, Leonardo Azevedo, Vítor Lourenço, Elton Soares, Raphael Thiago, Rafael Brandão, Daniel Civitarese, Emilio Vital Brazil, Marcio Moreno, Patrick Valduriez(+3 more)

Figure 1 for Provenance Data in the Machine Learning Lifecycle in Computational Science and Engineering

Figure 2 for Provenance Data in the Machine Learning Lifecycle in Computational Science and Engineering

Figure 3 for Provenance Data in the Machine Learning Lifecycle in Computational Science and Engineering

Figure 4 for Provenance Data in the Machine Learning Lifecycle in Computational Science and Engineering

Abstract:Machine Learning (ML) has become essential in several industries. In Computational Science and Engineering (CSE), the complexity of the ML lifecycle comes from the large variety of data, scientists' expertise, tools, and workflows. If data are not tracked properly during the lifecycle, it becomes unfeasible to recreate a ML model from scratch or to explain to stakeholders how it was created. The main limitation of provenance tracking solutions is that they cannot cope with provenance capture and integration of domain and ML data processed in the multiple workflows in the lifecycle while keeping the provenance capture overhead low. To handle this problem, in this paper we contribute with a detailed characterization of provenance data in the ML lifecycle in CSE; a new provenance data representation, called PROV-ML, built on top of W3C PROV and ML Schema; and extensions to a system that tracks provenance from multiple workflows to address the characteristics of ML and CSE, and to allow for provenance queries with a standard vocabulary. We show a practical use in a real case in the Oil and Gas industry, along with its evaluation using 48 GPUs in parallel.

* 10 pages, 7 figures, Accepted at Workflows in Support of Large-scale Science (WORKS) co-located with the ACM/IEEE International Conference for High Performance Computing, Networking, Storage, and Analysis (SC) 2019, Denver, Colorado

Via

Access Paper or Ask Questions