Data · dataset · 2026
Explainable AI for malware opcode sequence analysis and explainability-motivated saliency-map based spurious correlation robustness in images
Listed in ZivaHub and Deakin Research Online and DMU Figshare — shown once because both records carry DOI 10.17034/32805341.v1
Explainability offers a powerful lens for understanding and improving the robustness of Machine Learning (ML) models.
Description
This work demonstrates how eXplainable AI (XAI) techniques can be used not only to interpret model behaviour, but also to develop robust training algorithms that encourage the learning of semantically meaningful features.<br><br>The first contribution is Hierarchical-LIME (H-LIME), a novel XAI method tailored to malicious Android opcode sequence analysis.
Unlike standard LIME, which treats individual opcodes as flat, independent features, H-LIME leverages the hierarchical structure of programs, such as classes and methods, to produce sparser and more descriptively accurate explanations, improving the localisation of malicious code segments.<br><br>The second contribution, UnLearning from Experience (ULE), integrates explainability into the training process of image classifiers by leveraging saliency maps to guide model behaviour.
Read the rest (4 more)
ULE trains two models concurrently: a student model and a teacher model. The student is trained using standard Empirical Risk Minimisation and learns to rely on spurious correlations present in the data. Meanwhile, the teacher is trained to actively avoid these spurious correlations by minimising alignment with the student’s saliency maps.
This encourages the teacher to focus on more semantically meaningful features, resulting in a model that learns a more robust feature representation. Importantly, ULE achieves this without requiring x access to group labels or prior information about spurious features, making it widely applicable in real-world scenarios.<br><br>The final contribution, Weighted UnLearning from Experience (wULE), builds upon ULE by introducing a targeted sample re-weighting strategy that distinguishes between samples likely and unlikely to contain spurious correlations.
Instead of treating all samples uniformly, wULE leverages the simplicity bias principle to estimate the set of training samples suspected to contain spurious correlations. The loss function is adjusted on a per-sample basis: samples in this set receive stronger penalties for gradient alignment to encourage unlearning, while cleaner samples are weighted more heavily in the classification objective to reinforce reliable feature learning.
This adaptive strategy leads to improved worst-group performance and further enhances interpretability through more focused and meaningful saliency maps. Together, these contributions position explainability, not just as a diagnostic tool, but a core strategy for training resilient and robust machine learning systems.
Links
Where it is published
- DOI doi.org/10.17034/32805341.v1 ↗
DOI / persistent id · from zivahub uct ac za
Catalogue records · 1
- OAI-PMH record api.figshare.com/v2/oai?verb=GetRecord&metadataPrefix=oai_dc&identifier=oai%3Af… ↗
metadata API · from zivahub uct ac za
Topics
- From keywords
- Computer Science & AI · Computer Science & AI · Computer Science & AI · Deep learning · Deep learning · Deep learning · Earth & Environmental Science · Earth & Environmental Science · Earth & Environmental Science
- Inferred from text
- Image 75%
Provenance · 3 source records, 14 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| ZivaHub | oai:figshare.com:article/32805341 | 7 d ago | JSON v1 |
| Deakin Research Online | oai:figshare.com:article/32805341 | 7 d ago | JSON v1 |
| DMU Figshare | oai:figshare.com:article/32805341 | 7 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| concepts[field].anzsrc:field:461103 | mapping · dro deakin edu au | vocabulary-mapper@1.0.0 | keywords['deep learning'] |
| concepts[field].anzsrc:field:461103 | mapping · zivahub uct ac za | vocabulary-mapper@1.0.0 | keywords['deep learning'] |
| concepts[field].anzsrc:field:461103 | mapping · figshare dmu ac uk | vocabulary-mapper@1.0.0 | keywords['deep learning'] |
| concepts[field].local:field:computer-science-ai | mapping · dro deakin edu au | connector:dro_deakin_edu_au@1.0.0 | |
| concepts[field].local:field:computer-science-ai | mapping · figshare dmu ac uk | connector:figshare_dmu_ac_uk@1.0.0 | |
| concepts[field].local:field:computer-science-ai | mapping · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | |
| concepts[field].local:field:earth-environmental | mapping · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | |
| concepts[field].local:field:earth-environmental | mapping · dro deakin edu au | connector:dro_deakin_edu_au@1.0.0 | |
| concepts[field].local:field:earth-environmental | mapping · figshare dmu ac uk | connector:figshare_dmu_ac_uk@1.0.0 | |
| concepts[modality].local:modality:image | enrichment · zivahub uct ac za | keyword-concept-rules@1.0.0 | title+description (75%) |
| description | source · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | /metadata/dc/description |
| license_text | source · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | |
| publication_date | source · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | |
| title | source · zivahub uct ac za | connector:zivahub_uct_ac_za@1.0.0 | /metadata/dc/title |