BearGen: LLM-guided signal generation framework for bearing fault diagnosis

可解释性 计算机科学 可靠性(半导体) 稀缺 信号(编程语言) 断层(地质) 数据挖掘 机器学习 数据建模 人工智能 可靠性工程 故障检测与隔离 信号处理 方位(导航) 服务器 决策树 生成语法 数据驱动 实时计算 数据安全 状态监测 数据类型 滤波器(信号处理) Web应用程序 生成模型 深度学习
作者
Jaeyoung Lee,Hyuna Jeon,Uiin Kim,Misuk Kim
出处
期刊:Advanced Engineering Informatics [Elsevier BV]
卷期号:71: 104400-104400
标识
DOI:10.1016/j.aei.2026.104400
摘要

Signal data are essential for condition monitoring, fault diagnosis, and decision-making across industrial domains, and research leveraging signal data has been actively pursued in areas such as healthcare and manufacturing. However, acquiring such data is costly and difficult due to factors such as the risk of equipment damage, the need for expert labeling, and the scarcity of fault data. Moreover, collected data often contain sensitive operational information, making sharing difficult, and enterprises are restricted from using high-performance models hosted on external servers due to security concerns. To address these challenges, we propose BearGen , a novel framework that combines the strong generative capabilities of Large Language Models (LLMs) with the precise data distribution learning of diffusion models to synthesize high-quality signal data in on-premise environments. BearGen first employs an LLM to generate descriptions of existing signals and then conditions a description-guided diffusion model on these descriptions to generate high-quality synthetic signals. We evaluated BearGen on eight publicly available bearing fault diagnosis datasets, and the results showed superior performance compared to existing approaches. In addition, we experimentally validated the reliability and usefulness of the generated signal descriptions. Further experiments under conditions simulating real industrial environments — such as limited data availability and severe data imbalance — verified the practical applicability of the framework. By operating in on-premise environments, BearGen resolves data security concerns while alleviating data scarcity and imbalance. Furthermore, by providing natural language descriptions, it enhances interpretability and offers significant potential for decision support in real-world industrial applications.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
刚刚
彪行天下完成签到,获得积分10
刚刚
落后的滑板完成签到,获得积分10
1秒前
踏实的不愁完成签到,获得积分20
2秒前
勤奋兔子完成签到,获得积分10
2秒前
ly发布了新的文献求助30
2秒前
情怀应助苗条的荧荧采纳,获得10
3秒前
zhou发布了新的文献求助30
3秒前
4秒前
茶茶发布了新的文献求助10
5秒前
6秒前
6秒前
隐形的依霜完成签到,获得积分10
6秒前
乐乐应助gis采纳,获得10
7秒前
科研通AI6.4应助热情馒头采纳,获得10
7秒前
秋名山喵喵完成签到,获得积分10
7秒前
zhgj发布了新的文献求助100
8秒前
淡定如芳完成签到,获得积分20
8秒前
dd发布了新的文献求助10
8秒前
孙梦涵完成签到,获得积分10
9秒前
幽一完成签到,获得积分10
9秒前
10秒前
夏天发布了新的文献求助10
11秒前
Xyan完成签到 ,获得积分10
11秒前
冷静的豪完成签到 ,获得积分10
11秒前
Maocan完成签到,获得积分10
12秒前
淡然迎波发布了新的文献求助10
12秒前
12秒前
12秒前
13秒前
14秒前
14秒前
14秒前
14秒前
15秒前
lcy完成签到 ,获得积分20
15秒前
15秒前
meng完成签到,获得积分10
16秒前
捻念发布了新的文献求助10
16秒前
虚拟的夏菡完成签到,获得积分10
16秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Principles of town planning: translating concepts to applications 1000
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
The Effective Clinical Neurologist 3ed 500
The Great Hymn to Šamaš 500
Moody's Ratings Rising AI spending narrows the gap, but US hyperscalers retain edge over Chinese peers 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7695688
求助须知:如何正确求助?哪些是违规求助? 9256163
关于积分的说明 20001072
捐赠科研通 7270154
什么是DOI,文献DOI怎么找? 3292558
关于科研通互助平台的介绍 2448209
邀请新用户注册赠送积分活动 2298217