In April 2016 Manchester eScholar was replaced by the University of Manchester’s new Research Information Management System, Pure. In the autumn the University’s research outputs will be available to search and browse via a new Research Portal. Until then the University’s full publication record can be accessed via a temporary portal and the old eScholar content is available to search and browse via this archive.

Related resources

Search for item elsewhere

University researcher(s)

Academic department(s)

Quantifying the Stability of Feature Selection

Nogueira, Sarah

[Thesis]. Manchester, UK: The University of Manchester; 2018.

Access to files

FULL-TEXT.PDF (pdf)

Abstract

Feature Selection is central to modern data science, from exploratory data analysis to predictive model-building. The "stability"of a feature selection algorithm refers to the robustness of its feature preferences, with respect to data sampling and to its stochastic nature. An algorithm is "unstable" if a small change in data leads to large changes in the chosen feature subset. Whilst the idea is simple, quantifying this has proven more challenging---we note numerous proposals in the literature, each with different motivation and justification. We present a rigorous statistical and axiomatic treatment for this issue. In particular, with this work we consolidate the literature and provide (1) a deeper understanding of existing work based on a small set of properties, and (2) a clearly justified statistical approach with several novel benefits. This approach serves to identify a stability measure obeying all desirable properties, and (for the first time in the literature) allowing confidence intervals and hypothesis tests on the stability of an approach, enabling rigorous comparison of feature selection algorithms.

Keyword(s)

Feature Selection; Stability; Variable Selection

Bibliographic metadata

Type of resource:

text

Content type:

Administered thesis

Form of thesis:

Traditional

Type of submission:

Doctoral level ETD - final

Thesis title:

Quantifying the Stability of Feature Selection

Degree type:

Doctor of Philosophy

Degree programme:

PhD Computer Science (CDT)

Publication date:

2018-02-02T15:36:35

Institution:

The University of Manchester

Location:

Manchester, UK

Total pages:

126

Abstract:

Keyword(s):

Thesis main supervisor(s):

BROWN, GAVIN G

Thesis co-supervisor(s):

SHAPIRO, JONATHAN JL

Degree grantor:

The University of Manchester

Language:

Institutional metadata

University researcher(s):

Nogueira, Sarah

Academic department(s):

Record metadata

Manchester eScholar ID:

uk-ac-man-scw:313287

Created by:

Nogueira, Sarah

Created:

2nd February, 2018, 15:36:35

Last modified by:

Nogueira, Sarah

Last modified:

2nd March, 2018, 10:30:41