已入深夜,您辛苦了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!祝你早点完成任务,早点休息,好梦!

Embedding Approaches for Relational Data

作者
Yu Wu
标识
DOI:10.17638/03016866
摘要

​Embedding methods for searching latent representations of the data are very important tools for unsupervised and supervised machine learning as well as information visualisation. Over the years, such methods have continually progressed towards the ability to capture and analyse the structure and latent characteristics of larger and more complex data. In this thesis, we examine the problem of developing efficient and reliable embedding methods for revealing, understanding, and exploiting the different aspects of the relational data. We split our work into three pieces, where each deals with a different relational data structure. In the first part, we are handling with the weighted bipartite relational structure. Based on the relational measurements between two groups of heterogeneous objects, our goal is to generate low dimensional representations of these two different types of objects in a unified common space. We propose a novel method that models the embedding of each object type symmetrically to the other type, subject to flexible scale constraints and weighting parameters. The embedding generation relies on an efficient optimisation despatched using matrix decomposition. And we have also proposed a simple way of measuring the conformity between the original object relations and the ones re-estimated from the embeddings, in order to achieve model selection by identifying the optimal model parameters with a simple search procedure. We show that our proposed method achieves consistently better or on-par results on multiple synthetic datasets and real world ones from the text mining domain when compared with existing embedding generation approaches. In the second part of this thesis, we focus on the multi-relational data, where objects are interlinked by various relation types. Embedding approaches are very popular in this field, they typically encode objects and relation types with hidden representations and use the operations between them to compute the positive scalars corresponding to the linkages' likelihood score. In this work, we aim at further improving the existing embedding techniques by taking into account the multiple facets of the different patterns and behaviours of each relation type. To the best of our knowledge, this is the first latent representation model which considers relational representations to be dependent on the objects they relate in this field. The multi-modality of the relation type over different objects is effectively formulated as a projection matrix over the space spanned by the object vectors. Two large benchmark knowledge bases are used to evaluate the performance with respect to the link prediction task. And a new test data partition scheme is proposed to offer a better understanding of the behaviour of a link prediction model. In the last part of this thesis, a much more complex relational structure is considered. In particular, we aim at developing novel embedding methods for jointly modelling the linkage structure and objects' attributes. Traditionally, link prediction task is carried out on either the linkage structure or the objects' attributes, which does not aware of their semantic connections and is insufficient for handling the complex link prediction task. Thus, our goal in this work is to build a reliable model that can fuse both sources of information to improve the link prediction problem. The key idea of our approach is to encode both the linkage validities and the nodes neighbourhood information into embedding-based conditional probabilities. Another important aspect of our proposed algorithm is that we utilise a margin-based contrastive training process for encoding the linkage structure, which relies on a more appropriate assumption and dramatically reduces the number of training links. In the experiments, our proposed method indeed improves the link prediction performance on three citation/hyperlink datasets, when compared with those methods relying on only the nodes' attributes or the linkage structure, and it also achieves much better performances compared with the state-of-arts.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
越幸运完成签到 ,获得积分10
1秒前
休斯顿发布了新的文献求助10
3秒前
从容的柠檬完成签到 ,获得积分10
3秒前
4秒前
4秒前
纯真哈密瓜完成签到,获得积分10
5秒前
5秒前
5秒前
霸气一斩完成签到,获得积分10
6秒前
6秒前
小巧的傲晴完成签到,获得积分10
7秒前
7秒前
8秒前
狂野的含烟完成签到 ,获得积分10
9秒前
zhaoshuo发布了新的文献求助10
10秒前
zhaoshuo发布了新的文献求助10
10秒前
zhaoshuo发布了新的文献求助10
10秒前
zhaoshuo发布了新的文献求助10
10秒前
zhaoshuo发布了新的文献求助10
10秒前
zhaoshuo发布了新的文献求助10
10秒前
YYL完成签到 ,获得积分10
12秒前
zhaoshuo发布了新的文献求助10
14秒前
15秒前
cscgood完成签到,获得积分10
15秒前
cxw陈祥薇完成签到 ,获得积分10
15秒前
细腻访云完成签到,获得积分20
15秒前
19秒前
休斯顿发布了新的文献求助30
19秒前
雪儿发布了新的文献求助10
22秒前
Ding发布了新的文献求助10
23秒前
tjnksy完成签到,获得积分0
25秒前
传统的松鼠完成签到 ,获得积分10
26秒前
阿容发布了新的文献求助10
29秒前
缓慢曼易完成签到,获得积分20
30秒前
32秒前
32秒前
baihehuakai完成签到 ,获得积分10
32秒前
陶醉的美女完成签到,获得积分10
33秒前
我是老大的应助被Ding采纳,获得10
33秒前
35秒前
高分求助中
(应助此贴封号)通过应助OA文献获取积分 10000
The Student's Guide to Social Neuroscience 800
Rosenblum, Global Change Biology 800
Computational Chemical Reaction Engineering: Modeling, Simulation, and Design with MATLAB 600
Organizational Behavior 510
Management and the Arts 510
Production Logging: Theoretical and Interpretive Elements 400
热门求助领域 (近24小时)
化学 材料科学 医学 生物 计算机科学 工程类 纳米技术 内科学 物理 有机化学 化学工程 生物化学 复合材料 光电子学 细胞生物学 心理学 量子力学 催化作用 物理化学 电极
热门帖子
关注 科研通微信公众号,转发送积分 7812607
求助须知:如何正确求助?哪些是违规求助? 9343667
关于积分的说明 20519093
捐赠科研通 7405336
什么是DOI,文献DOI怎么找? 3330174
关于科研通互助平台的介绍 2476762
邀请新用户注册赠送积分活动 2349723