Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Xiaochen Hou

LinkSAGE: Optimizing Job Matching Using Graph Neural Networks

Feb 20, 2024

Ping Liu, Haichao Wei, Xiaochen Hou, Jianqiang Shen, Shihai He, Kay Qianqi Shen, Zhujun Chen, Fedor Borisyuk, Daniel Hewlett, Liang Wu(+4 more)

Figure 1 for LinkSAGE: Optimizing Job Matching Using Graph Neural Networks

Figure 2 for LinkSAGE: Optimizing Job Matching Using Graph Neural Networks

Figure 3 for LinkSAGE: Optimizing Job Matching Using Graph Neural Networks

Figure 4 for LinkSAGE: Optimizing Job Matching Using Graph Neural Networks

Abstract:We present LinkSAGE, an innovative framework that integrates Graph Neural Networks (GNNs) into large-scale personalized job matching systems, designed to address the complex dynamics of LinkedIns extensive professional network. Our approach capitalizes on a novel job marketplace graph, the largest and most intricate of its kind in industry, with billions of nodes and edges. This graph is not merely extensive but also richly detailed, encompassing member and job nodes along with key attributes, thus creating an expansive and interwoven network. A key innovation in LinkSAGE is its training and serving methodology, which effectively combines inductive graph learning on a heterogeneous, evolving graph with an encoder-decoder GNN model. This methodology decouples the training of the GNN model from that of existing Deep Neural Nets (DNN) models, eliminating the need for frequent GNN retraining while maintaining up-to-date graph signals in near realtime, allowing for the effective integration of GNN insights through transfer learning. The subsequent nearline inference system serves the GNN encoder within a real-world setting, significantly reducing online latency and obviating the need for costly real-time GNN infrastructure. Validated across multiple online A/B tests in diverse product scenarios, LinkSAGE demonstrates marked improvements in member engagement, relevance matching, and member retention, confirming its generalizability and practical impact.

Via

Access Paper or Ask Questions

LiGNN: Graph Neural Networks at LinkedIn

Feb 17, 2024

Fedor Borisyuk, Shihai He, Yunbo Ouyang, Morteza Ramezani, Peng Du, Xiaochen Hou, Chengming Jiang, Nitin Pasumarthy, Priya Bannur, Birjodh Tiwana(+13 more)

Figure 1 for LiGNN: Graph Neural Networks at LinkedIn

Figure 2 for LiGNN: Graph Neural Networks at LinkedIn

Figure 3 for LiGNN: Graph Neural Networks at LinkedIn

Figure 4 for LiGNN: Graph Neural Networks at LinkedIn

Abstract:In this paper, we present LiGNN, a deployed large-scale Graph Neural Networks (GNNs) Framework. We share our insight on developing and deployment of GNNs at large scale at LinkedIn. We present a set of algorithmic improvements to the quality of GNN representation learning including temporal graph architectures with long term losses, effective cold start solutions via graph densification, ID embeddings and multi-hop neighbor sampling. We explain how we built and sped up by 7x our large-scale training on LinkedIn graphs with adaptive sampling of neighbors, grouping and slicing of training data batches, specialized shared-memory queue and local gradient optimization. We summarize our deployment lessons and learnings gathered from A/B test experiments. The techniques presented in this work have contributed to an approximate relative improvements of 1% of Job application hearing back rate, 2% Ads CTR lift, 0.5% of Feed engaged daily active users, 0.2% session lift and 0.1% weekly active user lift from people recommendation. We believe that this work can provide practical solutions and insights for engineers who are interested in applying Graph neural networks at large scale.

Via

Access Paper or Ask Questions

LiRank: Industrial Large Scale Ranking Models at LinkedIn

Feb 10, 2024

Fedor Borisyuk, Mingzhou Zhou, Qingquan Song, Siyu Zhu, Birjodh Tiwana, Ganesh Parameswaran, Siddharth Dangi, Lars Hertel, Qiang Xiao, Xiaochen Hou(+24 more)

Figure 1 for LiRank: Industrial Large Scale Ranking Models at LinkedIn

Figure 2 for LiRank: Industrial Large Scale Ranking Models at LinkedIn

Figure 3 for LiRank: Industrial Large Scale Ranking Models at LinkedIn

Figure 4 for LiRank: Industrial Large Scale Ranking Models at LinkedIn

Abstract:We present LiRank, a large-scale ranking framework at LinkedIn that brings to production state-of-the-art modeling architectures and optimization methods. We unveil several modeling improvements, including Residual DCN, which adds attention and residual connections to the famous DCNv2 architecture. We share insights into combining and tuning SOTA architectures to create a unified model, including Dense Gating, Transformers and Residual DCN. We also propose novel techniques for calibration and describe how we productionalized deep learning based explore/exploit methods. To enable effective, production-grade serving of large ranking models, we detail how to train and compress models using quantization and vocabulary compression. We provide details about the deployment setup for large-scale use cases of Feed ranking, Jobs Recommendations, and Ads click-through rate (CTR) prediction. We summarize our learnings from various A/B tests by elucidating the most effective technical approaches. These ideas have contributed to relative metrics improvements across the board at LinkedIn: +0.5% member sessions in the Feed, +1.76% qualified job applications for Jobs search and recommendations, and +4.3% for Ads CTR. We hope this work can provide practical insights and solutions for practitioners interested in leveraging large-scale deep ranking systems.

Via

Access Paper or Ask Questions

Semantic Categorization of Social Knowledge for Commonsense Question Answering

Sep 11, 2021

Gengyu Wang, Xiaochen Hou, Diyi Yang, Kathleen McKeown, Jing Huang

Figure 1 for Semantic Categorization of Social Knowledge for Commonsense Question Answering

Figure 2 for Semantic Categorization of Social Knowledge for Commonsense Question Answering

Figure 3 for Semantic Categorization of Social Knowledge for Commonsense Question Answering

Figure 4 for Semantic Categorization of Social Knowledge for Commonsense Question Answering

Abstract:Large pre-trained language models (PLMs) have led to great success on various commonsense question answering (QA) tasks in an end-to-end fashion. However, little attention has been paid to what commonsense knowledge is needed to deeply characterize these QA tasks. In this work, we proposed to categorize the semantics needed for these tasks using the SocialIQA as an example. Building upon our labeled social knowledge categories dataset on top of SocialIQA, we further train neural QA models to incorporate such social knowledge categories and relation information from a knowledge base. Unlike previous work, we observe our models with semantic categorizations of social knowledge can achieve comparable performance with a relatively simple model and smaller size compared to other complex approaches.

* Accepted by SustaiNLP 2021 on EMNLP 2021

Via

Access Paper or Ask Questions

Graph Ensemble Learning over Multiple Dependency Trees for Aspect-level Sentiment Classification

Mar 12, 2021

Xiaochen Hou, Peng Qi, Guangtao Wang, Rex Ying, Jing Huang, Xiaodong He, Bowen Zhou

Figure 1 for Graph Ensemble Learning over Multiple Dependency Trees for Aspect-level Sentiment Classification

Figure 2 for Graph Ensemble Learning over Multiple Dependency Trees for Aspect-level Sentiment Classification

Figure 3 for Graph Ensemble Learning over Multiple Dependency Trees for Aspect-level Sentiment Classification

Figure 4 for Graph Ensemble Learning over Multiple Dependency Trees for Aspect-level Sentiment Classification

Abstract:Recent work on aspect-level sentiment classification has demonstrated the efficacy of incorporating syntactic structures such as dependency trees with graph neural networks(GNN), but these approaches are usually vulnerable to parsing errors. To better leverage syntactic information in the face of unavoidable errors, we propose a simple yet effective graph ensemble technique, GraphMerge, to make use of the predictions from differ-ent parsers. Instead of assigning one set of model parameters to each dependency tree, we first combine the dependency relations from different parses before applying GNNs over the resulting graph. This allows GNN mod-els to be robust to parse errors at no additional computational cost, and helps avoid overparameterization and overfitting from GNN layer stacking by introducing more connectivity into the ensemble graph. Our experiments on the SemEval 2014 Task 4 and ACL 14 Twitter datasets show that our GraphMerge model not only outperforms models with single dependency tree, but also beats other ensemble mod-els without adding model parameters.

* Accepted by NAACL 2021

Via

Access Paper or Ask Questions

Selective Attention Based Graph Convolutional Networks for Aspect-Level Sentiment Classification

Oct 24, 2019

Xiaochen Hou, Jing Huang, Guangtao Wang, Kevin Huang, Xiaodong He, Bowen Zhou

Figure 1 for Selective Attention Based Graph Convolutional Networks for Aspect-Level Sentiment Classification

Figure 2 for Selective Attention Based Graph Convolutional Networks for Aspect-Level Sentiment Classification

Figure 3 for Selective Attention Based Graph Convolutional Networks for Aspect-Level Sentiment Classification

Figure 4 for Selective Attention Based Graph Convolutional Networks for Aspect-Level Sentiment Classification

Abstract:Aspect-level sentiment classification aims to identify the sentiment polarity towards a specific aspect term in a sentence. Most current approaches mainly consider the semantic information by utilizing attention mechanisms to capture the interactions between the context and the aspect term. In this paper, we propose to employ graph convolutional networks (GCNs) on the dependency tree to learn syntax-aware representations of aspect terms. GCNs often show the best performance with two layers, and deeper GCNs do not bring additional gain due to over-smoothing problem. However, in some cases, important context words cannot be reached within two hops on the dependency tree. Therefore we design a selective attention based GCN block (SA-GCN) to find the most important context words, and directly aggregate these information into the aspect-term representation. We conduct experiments on the SemEval 2014 Task 4 datasets. Our experimental results show that our model outperforms the current state-of-the-art.

Via

Access Paper or Ask Questions