A MIL-based framework via contrastive instance learning and multimodal learning for long-term ECG classification

Haozhan Han; Cheng Lian; Bingrong Xu; Zhigang Zeng; Adi Alhudhaif; Kemal Polat

doi:10.1016/j.asoc.2024.112372

A MIL-based framework via contrastive instance learning and multimodal learning for long-term ECG classification

Haozhan Han
, Cheng Lian
, Bingrong Xu
, Zhigang Zeng
, Adi Alhudhaif
, Kemal Polat

Computer Sciences

Research output: Contribution to journal › Article › peer-review

1 Scopus citations

Abstract

Recently, deep learning-based models are widely employed for electrocardiogram (ECG) classification. However, classifying long-term ECGs, which contain vast amounts of data, is challenging. Due to the limitation of memory with respect to the original data size, preprocessing techniques such as resizing or cropping are often applied, leading to information loss. Therefore, introducing multi-instance learning (MIL) to address long-term ECG classification problems is crucial. However, a major drawback of employing MIL is the destruction of sample integrity, which consequently hinders the interaction among instances. To tackle this challenge, we proposed a multimodal MIL neural network named CIMIL, which consists of three key components: an instance interactor, a feature fusion method based on attention mechanisms, and a multimodal contrastive instance loss. First, we designed an instance interactor to improve the interaction and keep continuity among instances. Second, we proposed a novel feature fusion method based on attention mechanisms to effectively aggregate multimodal instance features for final classification, which selects key instances within each class, not only enhances the performance of our model but also reduces the number of parameters. Third, a multimodal contrastive instance loss is proposed to enhance the model's ability to distinguish positive and negative multimodal instances. Finally, we evaluated CIMIL on both intrapatient and interpatient patterns of two commonly used ECG datasets. The experimental results show that the proposed CIMIL outperforms existing state-of-the-art methods on long-term ECG tasks.

Original language	English
Article number	112372
Journal	Applied Soft Computing
Volume	167
DOIs	https://doi.org/10.1016/j.asoc.2024.112372
State	Published - Dec 2024

Keywords

Attention mechanism
Long-term ECG
Multi-instance learning
Multimodal learning

Access to Document

10.1016/j.asoc.2024.112372

Cite this

@article{f17d239447254227bd0b1e15fdde4fa9,

title = "A MIL-based framework via contrastive instance learning and multimodal learning for long-term ECG classification",

abstract = "Recently, deep learning-based models are widely employed for electrocardiogram (ECG) classification. However, classifying long-term ECGs, which contain vast amounts of data, is challenging. Due to the limitation of memory with respect to the original data size, preprocessing techniques such as resizing or cropping are often applied, leading to information loss. Therefore, introducing multi-instance learning (MIL) to address long-term ECG classification problems is crucial. However, a major drawback of employing MIL is the destruction of sample integrity, which consequently hinders the interaction among instances. To tackle this challenge, we proposed a multimodal MIL neural network named CIMIL, which consists of three key components: an instance interactor, a feature fusion method based on attention mechanisms, and a multimodal contrastive instance loss. First, we designed an instance interactor to improve the interaction and keep continuity among instances. Second, we proposed a novel feature fusion method based on attention mechanisms to effectively aggregate multimodal instance features for final classification, which selects key instances within each class, not only enhances the performance of our model but also reduces the number of parameters. Third, a multimodal contrastive instance loss is proposed to enhance the model's ability to distinguish positive and negative multimodal instances. Finally, we evaluated CIMIL on both intrapatient and interpatient patterns of two commonly used ECG datasets. The experimental results show that the proposed CIMIL outperforms existing state-of-the-art methods on long-term ECG tasks.",

keywords = "Attention mechanism, Long-term ECG, Multi-instance learning, Multimodal learning",

author = "Haozhan Han and Cheng Lian and Bingrong Xu and Zhigang Zeng and Adi Alhudhaif and Kemal Polat",

note = "Publisher Copyright: {\textcopyright} 2024 Elsevier B.V.",

year = "2024",

month = dec,

doi = "10.1016/j.asoc.2024.112372",

language = "English",

volume = "167",

journal = "Applied Soft Computing",

issn = "1568-4946",

publisher = "Elsevier B.V.",

}

TY - JOUR

T1 - A MIL-based framework via contrastive instance learning and multimodal learning for long-term ECG classification

AU - Han, Haozhan

AU - Lian, Cheng

AU - Xu, Bingrong

AU - Zeng, Zhigang

AU - Alhudhaif, Adi

AU - Polat, Kemal

PY - 2024/12

Y1 - 2024/12

N2 - Recently, deep learning-based models are widely employed for electrocardiogram (ECG) classification. However, classifying long-term ECGs, which contain vast amounts of data, is challenging. Due to the limitation of memory with respect to the original data size, preprocessing techniques such as resizing or cropping are often applied, leading to information loss. Therefore, introducing multi-instance learning (MIL) to address long-term ECG classification problems is crucial. However, a major drawback of employing MIL is the destruction of sample integrity, which consequently hinders the interaction among instances. To tackle this challenge, we proposed a multimodal MIL neural network named CIMIL, which consists of three key components: an instance interactor, a feature fusion method based on attention mechanisms, and a multimodal contrastive instance loss. First, we designed an instance interactor to improve the interaction and keep continuity among instances. Second, we proposed a novel feature fusion method based on attention mechanisms to effectively aggregate multimodal instance features for final classification, which selects key instances within each class, not only enhances the performance of our model but also reduces the number of parameters. Third, a multimodal contrastive instance loss is proposed to enhance the model's ability to distinguish positive and negative multimodal instances. Finally, we evaluated CIMIL on both intrapatient and interpatient patterns of two commonly used ECG datasets. The experimental results show that the proposed CIMIL outperforms existing state-of-the-art methods on long-term ECG tasks.

AB - Recently, deep learning-based models are widely employed for electrocardiogram (ECG) classification. However, classifying long-term ECGs, which contain vast amounts of data, is challenging. Due to the limitation of memory with respect to the original data size, preprocessing techniques such as resizing or cropping are often applied, leading to information loss. Therefore, introducing multi-instance learning (MIL) to address long-term ECG classification problems is crucial. However, a major drawback of employing MIL is the destruction of sample integrity, which consequently hinders the interaction among instances. To tackle this challenge, we proposed a multimodal MIL neural network named CIMIL, which consists of three key components: an instance interactor, a feature fusion method based on attention mechanisms, and a multimodal contrastive instance loss. First, we designed an instance interactor to improve the interaction and keep continuity among instances. Second, we proposed a novel feature fusion method based on attention mechanisms to effectively aggregate multimodal instance features for final classification, which selects key instances within each class, not only enhances the performance of our model but also reduces the number of parameters. Third, a multimodal contrastive instance loss is proposed to enhance the model's ability to distinguish positive and negative multimodal instances. Finally, we evaluated CIMIL on both intrapatient and interpatient patterns of two commonly used ECG datasets. The experimental results show that the proposed CIMIL outperforms existing state-of-the-art methods on long-term ECG tasks.

KW - Attention mechanism

KW - Long-term ECG

KW - Multi-instance learning

KW - Multimodal learning

UR - https://www.scopus.com/pages/publications/85207903508

U2 - 10.1016/j.asoc.2024.112372

DO - 10.1016/j.asoc.2024.112372

M3 - Article

AN - SCOPUS:85207903508

SN - 1568-4946

VL - 167

JO - Applied Soft Computing

JF - Applied Soft Computing

M1 - 112372

ER -

A MIL-based framework via contrastive instance learning and multimodal learning for long-term ECG classification

Abstract

Keywords

Access to Document

Other files and links

Fingerprint

Cite this