Constarium
← Search

Data · dataset · 2026

Explainable AI for malware opcode sequence analysis and explainability-motivated saliency-map based spurious correlation robustness in images

Listed in ZivaHub and Deakin Research Online and DMU Figshare — shown once because both records carry DOI 10.17034/32805341.v1

Explainability offers a powerful lens for understanding and improving the robustness of Machine Learning (ML) models.

Description

This work demonstrates how eXplainable AI (XAI) techniques can be used not only to interpret model behaviour, but also to develop robust training algorithms that encourage the learning of semantically meaningful features.<br><br>The first contribution is Hierarchical-LIME (H-LIME), a novel XAI method tailored to malicious Android opcode sequence analysis.

Unlike standard LIME, which treats individual opcodes as flat, independent features, H-LIME leverages the hierarchical structure of programs, such as classes and methods, to produce sparser and more descriptively accurate explanations, improving the localisation of malicious code segments.<br><br>The second contribution, UnLearning from Experience (ULE), integrates explainability into the training process of image classifiers by leveraging saliency maps to guide model behaviour.

Read the rest (4 more)

ULE trains two models concurrently: a student model and a teacher model. The student is trained using standard Empirical Risk Minimisation and learns to rely on spurious correlations present in the data. Meanwhile, the teacher is trained to actively avoid these spurious correlations by minimising alignment with the student’s saliency maps.

This encourages the teacher to focus on more semantically meaningful features, resulting in a model that learns a more robust feature representation. Importantly, ULE achieves this without requiring x access to group labels or prior information about spurious features, making it widely applicable in real-world scenarios.<br><br>The final contribution, Weighted UnLearning from Experience (wULE), builds upon ULE by introducing a targeted sample re-weighting strategy that distinguishes between samples likely and unlikely to contain spurious correlations.

Instead of treating all samples uniformly, wULE leverages the simplicity bias principle to estimate the set of training samples suspected to contain spurious correlations. The loss function is adjusted on a per-sample basis: samples in this set receive stronger penalties for gradient alignment to encourage unlearning, while cleaner samples are weighted more heavily in the classification objective to reinforce reliable feature learning.

This adaptive strategy leads to improved worst-group performance and further enhances interpretability through more focused and meaningful saliency maps. Together, these contributions position explainability, not just as a diagnostic tool, but a core strategy for training resilient and robust machine learning systems.

Links

Where it is published

Catalogue records · 1

Topics

Inferred from text
Image 75%
Provenance · 3 source records, 14 field assertions
SourceKeyLast seenRaw
ZivaHuboai:figshare.com:article/328053417 d agoJSON v1
Deakin Research Onlineoai:figshare.com:article/328053417 d agoJSON v1
DMU Figshareoai:figshare.com:article/328053417 d agoJSON v1
FieldAssertionExtractorEvidence
concepts[field].anzsrc:field:461103mapping · dro deakin edu auvocabulary-mapper@1.0.0keywords['deep learning']
concepts[field].anzsrc:field:461103mapping · zivahub uct ac zavocabulary-mapper@1.0.0keywords['deep learning']
concepts[field].anzsrc:field:461103mapping · figshare dmu ac ukvocabulary-mapper@1.0.0keywords['deep learning']
concepts[field].local:field:computer-science-aimapping · dro deakin edu auconnector:dro_deakin_edu_au@1.0.0
concepts[field].local:field:computer-science-aimapping · figshare dmu ac ukconnector:figshare_dmu_ac_uk@1.0.0
concepts[field].local:field:computer-science-aimapping · zivahub uct ac zaconnector:zivahub_uct_ac_za@1.0.0
concepts[field].local:field:earth-environmentalmapping · zivahub uct ac zaconnector:zivahub_uct_ac_za@1.0.0
concepts[field].local:field:earth-environmentalmapping · dro deakin edu auconnector:dro_deakin_edu_au@1.0.0
concepts[field].local:field:earth-environmentalmapping · figshare dmu ac ukconnector:figshare_dmu_ac_uk@1.0.0
concepts[modality].local:modality:imageenrichment · zivahub uct ac zakeyword-concept-rules@1.0.0title+description (75%)
descriptionsource · zivahub uct ac zaconnector:zivahub_uct_ac_za@1.0.0/metadata/dc/description
license_textsource · zivahub uct ac zaconnector:zivahub_uct_ac_za@1.0.0
publication_datesource · zivahub uct ac zaconnector:zivahub_uct_ac_za@1.0.0
titlesource · zivahub uct ac zaconnector:zivahub_uct_ac_za@1.0.0/metadata/dc/title