WELMSD – word embedding and language model based sarcasm detection

讽刺 计算机科学 自然语言处理 人工智能 情绪分析 语言模型 分类器(UML) 文字嵌入 独创性 词(群论) 特征工程 嵌入 语言学 深度学习 心理学 讽刺 社会心理学 哲学 创造力
作者
Pradeep Kumar,Gaurav Sarin
出处
期刊:Online Information Review [Emerald Publishing Limited]
卷期号:46 (7): 1242-1256 被引量:14
标识
DOI:10.1108/oir-03-2021-0184
摘要

Purpose Sarcasm is a sentiment in which human beings convey messages with the opposite meanings to hurt someone emotionally or condemn something in a witty manner. The difference between the text's literal and its intended meaning makes it tough to identify. Mostly, researchers and practitioners only consider explicit information for text classification; however, considering implicit with explicit information will enhance the classifier's accuracy. Several sarcasm detection studies focus on syntactic, lexical or pragmatic features that are uttered using words, emoticons and exclamation marks. Discrete models, which are utilized by many existing works, require manual features that are costly to uncover. Design/methodology/approach In this research, word embeddings used for feature extraction are combined with context-aware language models to provide automatic feature engineering capabilities as well superior classification performance as compared to baseline models. Performance of the proposed models has been shown on three benchmark datasets over different evaluation metrics namely misclassification rate, receiver operating characteristic (ROC) curve and area under curve (AUC). Findings Experimental results suggest that FastText word embedding technique with BERT language model gives higher accuracy and helps to identify the sarcastic textual element correctly. Originality/value Sarcasm detection is a sub-task of sentiment analysis. To help in appropriate data-driven decision-making, the sentiment of the text that gets reversed due to sarcasm needs to be detected properly. In online social environments, it is critical for businesses and individuals to detect the correct sentiment polarity. This will aid in the right selling and buying of products and/or services, leading to higher sales and better market share for businesses, and meeting the quality requirements of customers.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
刚刚
2052669099发布了新的文献求助10
刚刚
刚刚
鲁文杰完成签到,获得积分10
1秒前
克里斯蒂娜完成签到,获得积分10
1秒前
1秒前
希望天下0贩的0应助xm采纳,获得10
2秒前
喔喔佳佳发布了新的文献求助10
3秒前
zzz发布了新的文献求助10
3秒前
科研浦东发布了新的文献求助10
4秒前
4秒前
充电宝应助mookie采纳,获得10
4秒前
5秒前
情怀应助李哈哈采纳,获得10
7秒前
烟花应助mei采纳,获得10
7秒前
9秒前
小太阳发布了新的文献求助10
9秒前
10秒前
希望天下0贩的0应助MAR采纳,获得10
10秒前
佘拜拜完成签到,获得积分10
10秒前
Ruolin发布了新的文献求助10
10秒前
超级的鹅完成签到,获得积分10
11秒前
11秒前
12秒前
张晓艳完成签到,获得积分20
12秒前
光亮千易完成签到,获得积分10
13秒前
Lucas应助莫小烦采纳,获得10
14秒前
整齐的幻香完成签到,获得积分20
15秒前
15秒前
Owen应助xm采纳,获得10
15秒前
学术小垃圾完成签到,获得积分10
16秒前
16秒前
mookie发布了新的文献求助10
16秒前
科研通AI6.2应助伍若舟采纳,获得10
17秒前
18秒前
asaso发布了新的文献求助10
18秒前
Naruto完成签到,获得积分10
19秒前
cblll完成签到 ,获得积分20
20秒前
20秒前
20秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Principles of town planning: translating concepts to applications 1000
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
核安全综合知识2024版 500
Photothermal Science and Techniques 500
The Effective Clinical Neurologist 3ed 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7714322
求助须知:如何正确求助?哪些是违规求助? 9269771
关于积分的说明 20078363
捐赠科研通 7290670
什么是DOI,文献DOI怎么找? 3298173
关于科研通互助平台的介绍 2452391
邀请新用户注册赠送积分活动 2305488