liu.seSök publikationer i DiVA
Ändra sökning
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • oxford
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Handling Novel and Out-Of-Distribution Data in Deep Learning: OOD Detection and Shortcut Mitigation
Linköpings universitet, Institutionen för datavetenskap, Statistik och maskininlärning. Linköpings universitet, Tekniska fakulteten.ORCID-id: 0000-0001-7411-2177
2025 (Engelska)Doktorsavhandling, sammanläggning (Övrigt vetenskapligt)
Abstract [en]

Advancements in machine learning, and particularly deep learning, have revolutionized the real-world applications of artificial intelligence in recent years. A main property of deep neural models is their ability to learn a task based on a set of examples, that is, the training data. Although the state-of-the-art performance of such models is promising in many tasks, this of-ten holds only as long as the inputs to the model are “sufficiently similar” to the training data. Mathematically, a ubiquitous assumption in machine learning studies is that the test data used for evaluating a model are sampled from the same probability distribution as the training data. It is challenging to approach any problem where this assumption is violated, as it requires handling Out-Of-Distribution (OOD) data, i.e., data points that are systematically different from the training (in-distribution) data. In particular, one might be interested in detecting OOD inputs at test time given an unlabeled training set, which is the main problem explored in this thesis. This type of OOD detection (a.k.a. novelty/anomaly detection) has various applications in discovering unusual events and phenomena as well as improving safety in AI systems. Another challenging problem in deep learning is that a model might rely on certain trivial relations (spurious correlations) existing in training data to solve a task. Such “shortcuts” can bring a high performance on in-distribution data, but they may collapse on more realistic OOD data. It is therefore vital to mitigate the shortcut learning effects in deep models, which is the second topic studied in this thesis.

A part of the present thesis is concerned with leveraging pretrained deep models for OOD detection on images, without modifying their standard training algorithms. A method is proposed to use invertible (flow-based) generative models based on null hypothesis testing ideas, leading to an OOD detection method that is fast and more reliable than the traditional likelihood-based method. Diffusion (score-based) models are another type of modern generative models used for OOD detection in this thesis, in combination with pretrained deep encoders. Another contribution of the thesis is in leveraging the power of large self-supervised models in fully unsupervised fine-grained OOD detection. It is shown that the simple k-nearest neighbor distance in the representation space of such models results in a reasonable performance but can be boosted substantially through the proposed adjustments, without any model fine-tuning. The local geometry of representations and background (irrelevant) features are considered to this end.

OOD detection with time series data is another problem studied in this thesis. Specifically, a method is proposed based on Contrastive Predictive Coding (CPC) self-supervised learning, and applied to detect novel categories in human activity data. It is demonstrated, both empirically and through theoretical motivation, that modifying the CPC to use a radial basis function instead of the conventional log-bilinear function is a requirement for reliable and efficient OOD detection. This extension is combined with quantization of representation vectors to achieve better performance.

This thesis also addresses the problem of learning deep representations (transfer learning) in a situation where a shortcut exists in data. In this problem, a deep model is trained on a shortcut-biased image dataset to solve a self-supervised or supervised classification task. The representations learned by this model are used to train a smaller model on a related but different downstream task, and the adverse effect of the shortcut is verified empirically there. Moreover, a method is proposed to enhance the representation learning in this scenario, based on an auxiliary model trained in an adversarial manner along with the upstream classifier.

Ort, förlag, år, upplaga, sidor
Linköping: Linköping University Electronic Press, 2025. , s. 76
Serie
Linköping Studies in Science and Technology. Dissertations, ISSN 0345-7524 ; 2445
Nationell ämneskategori
Artificiell intelligens
Identifikatorer
URN: urn:nbn:se:liu:diva-212978DOI: 10.3384/9789181180749ISBN: 9789181180732 (tryckt)ISBN: 9789181180749 (digital)OAI: oai:DiVA.org:liu-212978DiVA, id: diva2:1951761
Disputation
2025-05-16, Ada Lovelace, B-building, Campus Valla, Linköping, 09:30 (Engelska)
Opponent
Handledare
Tillgänglig från: 2025-04-14 Skapad: 2025-04-14 Senast uppdaterad: 2025-09-09Bibliografiskt granskad
Delarbeten
1. Likelihood-free Out-of-Distribution Detection with Invertible Generative Models
Öppna denna publikation i ny flik eller fönster >>Likelihood-free Out-of-Distribution Detection with Invertible Generative Models
2021 (Engelska)Ingår i: Proceedings of the Thirtieth International Joint Conference on Artificial Intelligence (IJCAI 2021), International Joint Conferences on Artifical Intelligence (IJCAI) , 2021, s. 2119-2125Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

Likelihood of generative models has been used traditionally as a score to detect atypical (Out-of-Distribution, OOD) inputs. However, several recent studies have found this approach to be highly unreliable, even with invertible generative models, where computing the likelihood is feasible. In this paper, we present a different framework for generative model--based OOD detection that employs the model in constructing a new representation space, instead of using it directly in computing typicality scores, where it is emphasized that the score function should be interpretable as the similarity between the input and training data in the new space. In practice, with a focus on invertible models, we propose to extract low-dimensional features (statistics) based on the model encoder and complexity of input images, and then use a One-Class SVM to score the data. Contrary to recently proposed OOD detection methods for generative models, our method does not require computing likelihood values. Consequently, it is much faster when using invertible models with iteratively approximated likelihood (e.g. iResNet), while it still has a performance competitive with other related methods

Ort, förlag, år, upplaga, sidor
International Joint Conferences on Artifical Intelligence (IJCAI), 2021
Serie
Proceedings of the International Joint Conference on Artificial Intelligence, ISSN 1045-0823
Nyckelord
Deep Learning, Anomaly/Outlier Detection, Uncertainty Representations
Nationell ämneskategori
Datavetenskap (datalogi)
Identifikatorer
urn:nbn:se:liu:diva-188936 (URN)10.24963/ijcai.2021/292 (DOI)001202335502027 ()2-s2.0-85125461759 (Scopus ID)9780999241196 (ISBN)
Konferens
International Joint Conference on Artificial Intelligence (IJCAI), 19-26 August, 2021
Tillgänglig från: 2022-10-03 Skapad: 2022-10-03 Senast uppdaterad: 2025-04-14Bibliografiskt granskad
2. Unsupervised Novelty Detection in Pretrained Representation Space with Locally Adapted Likelihood Ratio
Öppna denna publikation i ny flik eller fönster >>Unsupervised Novelty Detection in Pretrained Representation Space with Locally Adapted Likelihood Ratio
2024 (Engelska)Ingår i: International Conference on Artificial Intelligence and Statistics 2024, Proceedings of Machine Learning Research, 2024, Vol. 238Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

Detecting novelties given unlabeled examples of normal data is a challenging task in machine learning, particularly when the novel and normal categories are semantically close. Large deep models pretrained on massive datasets can provide a rich representation space in which the simple k-nearest neighbor distance works as a novelty measure. However, as we show in this paper, the basic k-NN method might be insufficient in this context due to ignoring the 'local geometry' of the distribution over representations as well as the impact of irrelevant 'background features'. To address this, we propose a fully unsupervised novelty detection approach that integrates the flexibility of k-NN with a locally adapted scaling of dimensions based on the 'neighbors of nearest neighbor' and computing a 'likelihood ratio' in pretrained (self-supervised) representation spaces. Our experiments with image data show the advantage of this method when off-the-shelf vision transformers (e.g., pretrained by DINO) are used as the feature extractor without any fine-tuning.

Serie
Proceedings of Machine Learning Research, ISSN 2640-3498
Nationell ämneskategori
Datavetenskap (datalogi) Datorgrafik och datorseende Signalbehandling
Identifikatorer
urn:nbn:se:liu:diva-203391 (URN)001221034002024 ()
Konferens
27th International Conference on Artificial Intelligence and Statistics (AISTATS), Valencia, SPAIN, MAY 02-04, 2024
Tillgänglig från: 2024-05-08 Skapad: 2024-05-08 Senast uppdaterad: 2025-04-14
3. Improved Contrastive Predictive Coding for Time Series Out-Of-Distribution Detection Applied to Human Activity Data
Öppna denna publikation i ny flik eller fönster >>Improved Contrastive Predictive Coding for Time Series Out-Of-Distribution Detection Applied to Human Activity Data
2025 (Engelska)Ingår i: Pattern Recognition Letters, ISSN 0167-8655, E-ISSN 1872-7344, Vol. 197, s. 132-138Artikel i tidskrift (Refereegranskat) Published
Abstract [en]

Contrastive Predictive Coding (CPC) is a well-established self-supervised learning method that naturally fits time series data. This method has been recently leveraged to detect anomalous inputs, viewed as the task of classifying positive pairs of context-feature representations versus negative ones in order to employ classifier uncertainty measures. In this paper, by taking a different perspective, we propose a CPC-based Out-Of-Distribution (OOD) detection method for time series data that does not require any negative samples at test time and is theoretically related to a probabilistic type of uncertainty estimation in the latent representation space. Our method extends the standard CPC by using a radial (distance-based) score function both in the training loss and as the OOD measure, in addition to quantizing the context (replacing it by cluster prototypes) during inference. The proposed method is applied to detecting OOD human activities with smartphone sensors data and shows promising performance on two primary datasets without using activity labels in training.

Ort, förlag, år, upplaga, sidor
ELSEVIER, 2025
Nyckelord
Out-Of-Distribution/anomaly/novelty; detection; Self-supervised machine learning; Contrastive learning; Human activity recognition; Time series analysis; Deep learning
Nationell ämneskategori
Datavetenskap (datalogi)
Identifikatorer
urn:nbn:se:liu:diva-217248 (URN)10.1016/j.patrec.2025.07.011 (DOI)001545804100001 ()2-s2.0-105012301703 (Scopus ID)
Anmärkning

Funding Agencies|Swedish Research Council [2024-05011]; Wallenberg AI, Autonomous Systems and Software Program (WASP) - Knut and Alice Wallenberg Foundation; Excellence Center at Linkoeping-Lund in Information Technology (ELLIIT)

Tillgänglig från: 2025-09-03 Skapad: 2025-09-03 Senast uppdaterad: 2025-12-18
4. Enhancing Representation Learning with Deep Classifiers in Presence of Shortcut
Öppna denna publikation i ny flik eller fönster >>Enhancing Representation Learning with Deep Classifiers in Presence of Shortcut
2023 (Engelska)Ingår i: Proceedings of IEEE ICASSP 2023, 2023Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

A deep neural classifier trained on an upstream task can be leveraged to boost the performance of another classifier in a related downstream task through the representations learned in hidden layers. However, presence of shortcuts (easy-to-learn features) in the upstream task can considerably impair the versatility of intermediate representations and, in turn, the downstream performance. In this paper, we propose a method to improve the representations learned by deep neural image classifiers in spite of a shortcut in upstream data. In our method, the upstream classification objective is augmented with a type of adversarial training where an auxiliary network, so called lens, fools the classifier by exploiting the shortcut in reconstructing images. Empirical comparisons in self-supervised and transfer learning problems with three shortcut-biased datasets suggest the advantages of our method in terms of downstream performance and/or training time.

Nyckelord
Deep Representation Learning, Shortcut Learning, Transfer Learning, Adversarial Methods, Computer Vision
Nationell ämneskategori
Datavetenskap (datalogi) Datorgrafik och datorseende
Identifikatorer
urn:nbn:se:liu:diva-198763 (URN)10.1109/ICASSP49357.2023.10096346 (DOI)001595432200428 ()2-s2.0-86000380743 (Scopus ID)
Konferens
2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)
Anmärkning

Funding Agencies|Swedish Research Council via the project Handling Uncertainty in Machine Learning Systems [2020-04122]; Wallenberg AI, Autonomous Systems and Software Program (WASP) - Knut and Alice Wallenberg Foundation; Excellence Center at Linkoping-Lund in Information Technology (ELLIIT)

Tillgänglig från: 2023-10-26 Skapad: 2023-10-26 Senast uppdaterad: 2026-02-05

Open Access i DiVA

fulltext(9856 kB)1056 nedladdningar
Filinformation
Filnamn FULLTEXT01.pdfFilstorlek 9856 kBChecksumma SHA-512
0a05daea7a47d7054745f306578a078f2eaa489683915657e2954c5f9d480cdb63813070f78caaec3b373fb836170fed3b027ded4b7cec1e1f0f25c4b99933de
Typ fulltextMimetyp application/pdf
Beställ online >>

Övriga länkar

Förlagets fulltext

Person

Ahmadian, Amirhossein

Sök vidare i DiVA

Av författaren/redaktören
Ahmadian, Amirhossein
Av organisationen
Statistik och maskininlärningTekniska fakulteten
Artificiell intelligens

Sök vidare utanför DiVA

GoogleGoogle Scholar
Totalt: 1062 nedladdningar
Antalet nedladdningar är summan av nedladdningar för alla fulltexter. Det kan inkludera t.ex tidigare versioner som nu inte längre är tillgängliga.

doi
isbn
urn-nbn

Altmetricpoäng

doi
isbn
urn-nbn
Totalt: 3365 träffar
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • oxford
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf