liu.seSearch for publications in DiVA
Change search
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • harvard1
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • oxford
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf
A Data Mining Approach to Analyze Non-compliance with a Guideline for the Treatment of Breast Cancer
Linköping University, Department of Biomedical Engineering, Medical Informatics. Linköping University, The Institute of Technology.
Linköping University, Department of Biomedical Engineering, Medical Informatics. Linköping University, The Institute of Technology.
Linköping University, Department of Biomedical Engineering, Medical Informatics. Linköping University, The Institute of Technology.
Linköping University, Department of Biomedical Engineering, Medical Informatics. Linköping University, The Institute of Technology.
2007 (English)In: Studies in Health Technology and Informatics, ISSN 0926-9630, Vol. 129, 591-597 p.Article in journal (Refereed) Published
Abstract [en]

Postmastectomy radiotherapy (PMRT) is prescribed in order to reduce the local recurrence of breast cancer and improve overall survival. A guideline supports the trade-off between benefits and adverse effects of PMRT. However, this guideline is not always followed in practice. This study tries to find a method for revealing patterns of non-compliance between the actual treatment and the PMRT guideline.

Data from breast cancer patients admitted to Linköping University Hospital between 1990 and 2000 were analyzed in this study. Cases that were not treated in accordance with the guideline were selected and analyzed by decision tree induction (DTI). Thereafter, four resulting rules, as representations for groups of patients, were compared to the guideline.

Finding patterns of non-compliance with guidelines by means of rules can be an appropriate alternative to manual methods, i.e. a case-by-case comparison when studying very large datasets. The resulting rules can be used in a knowledge base of a guideline-based decision support system to alert when inconsistencies with the guidelines may appear.

Place, publisher, year, edition, pages
2007. Vol. 129, 591-597 p.
National Category
Biomedical Laboratory Science/Technology
Identifiers
URN: urn:nbn:se:liu:diva-12709OAI: oai:DiVA.org:liu-12709DiVA: diva2:16893
Available from: 2007-10-30 Created: 2007-10-30 Last updated: 2009-03-16
In thesis
1. Applications of Knowledge Discovery in Quality Registries - Predicting Recurrence of Breast Cancer and Analyzing Non-compliance with a Clinical Guideline
Open this publication in new window or tab >>Applications of Knowledge Discovery in Quality Registries - Predicting Recurrence of Breast Cancer and Analyzing Non-compliance with a Clinical Guideline
2007 (English)Doctoral thesis, comprehensive summary (Other academic)
Abstract [en]

In medicine, data are produced from different sources and continuously stored in data depositories. Examples of these growing databases are quality registries. In Sweden, there are many cancer registries where data on cancer patients are gathered and recorded and are used mainly for reporting survival analyses to high level health authorities.

In this thesis, a breast cancer quality registry operating in South-East of Sweden is used as the data source for newer analytical techniques, i.e. data mining as a part of knowledge discovery in databases (KDD) methodology. Analyses are done to sift through these data in order to find interesting information and hidden knowledge. KDD consists of multiple steps, starting with gathering data from different sources and preparing them in data pre-processing stages prior to data mining.

Data were cleaned from outliers and noise and missing values were handled. Then a proper subset of the data was chosen by canonical correlation analysis (CCA) in a dimensionality reduction step. This technique was chosen because there were multiple outcomes, and variables had complex relationship to one another.

After data were prepared, they were analyzed with a data mining method. Decision tree induction as a simple and efficient method was used to mine the data. To show the benefits of proper data pre-processing, results from data mining with pre-processing of the data were compared with results from data mining without data pre-processing. The comparison showed that data pre-processing results in a more compact model with a better performance in predicting the recurrence of cancer.

An important part of knowledge discovery in medicine is to increase the involvement of medical experts in the process. This starts with enquiry about current problems in their field, which leads to finding areas where computer support can be helpful. The experts can suggest potentially important variables and should then approve and validate new patterns or knowledge as predictive or descriptive models. If it can be shown that the performance of a model is comparable to domain experts, it is more probable that the model will be used to support physicians in their daily decision-making. In this thesis, we validated the model by comparing predictions done by data mining and those made by domain experts without finding any significant difference between them.

Breast cancer patients who are treated with mastectomy are recommended to receive radiotherapy. This treatment is called postmastectomy radiotherapy (PMRT) and there is a guideline for prescribing it. A history of this treatment is stored in breast cancer registries. We analyzed these datasets using rules from a clinical guideline and identified cases that had not been treated according to the PMRT guideline. Data mining revealed some patterns of non-compliance with the PMRT guideline. Further analysis with data mining revealed some reasons for guideline non-compliance. These patterns were then compared with reasons acquired from manual inspection of patient records. The comparisons showed that patterns resulting from data mining were limited to the stored variables in the registry. A prerequisite for better results is availability of comprehensive datasets.

Medicine can take advantage of KDD methodology in different ways. The main advantage is being able to reuse information and explore hidden knowledge that can be obtained using advanced analysis techniques. The results depend on good collaboration between medical informaticians and domain experts and the availability of high quality data.

Place, publisher, year, edition, pages
Institutionen för medicinsk teknik, 2007. 58 p.
Series
Linköping University Medical Dissertations, ISSN 0345-0082 ; 1018
Keyword
Breast cancer, Clinical guidelines, Canonical correlation analysis, Data Mining, Data pre-processing, Decision tree induction, Knowledge Discovery in Databases
National Category
Biomedical Laboratory Science/Technology
Identifiers
urn:nbn:se:liu:diva-10142 (URN)978-91-85895-81-6 (ISBN)
Public defence
2007-11-22, Elsa Brändström, Campus US, Linköpings universitet, Linköping, 09:00 (English)
Opponent
Supervisors
Available from: 2007-10-30 Created: 2007-10-30 Last updated: 2009-05-12

Open Access in DiVA

No full text

Other links

Link to articleLink to Ph.D. Thesis

Authority records BETA

Razavi, Amir RezaGill, HansÅhlfeldt, HansShahsavar, Nosrat

Search in DiVA

By author/editor
Razavi, Amir RezaGill, HansÅhlfeldt, HansShahsavar, Nosrat
By organisation
Medical InformaticsThe Institute of Technology
In the same journal
Studies in Health Technology and Informatics
Biomedical Laboratory Science/Technology

Search outside of DiVA

GoogleGoogle Scholar

urn-nbn

Altmetric score

urn-nbn
Total: 579 hits
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • harvard1
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • oxford
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf