Classification in high dimensional feature spaces

Van Dyk, Hendrik Oostewald

Classification in high dimensional feature spaces

Files

vandyk_hendriko.pdf (721.12 KB)

Date

2009

Authors

Van Dyk, Hendrik Oostewald

Supervisors

Barnard, E.

Publisher

North-West University

Abstract

In this dissertation we developed theoretical models to analyse Gaussian and multinomial distributions. The analysis is focused on classification in high dimensional feature spaces and provides a basis for dealing with issues such as data sparsity and feature selection (for Gaussian and multinomial distributions, two frequently used models for high dimensional applications). A Naïve Bayesian philosophy is followed to deal with issues associated with the curse of dimensionality. The core treatment on Gaussian and multinomial models consists of finding analytical expressions for classification error performances. Exact analytical expressions were found for calculating error rates of binary class systems with Gaussian features of arbitrary dimensionality and using any type of quadratic decision boundary (except for degenerate paraboloidal boundaries). Similarly, computationally inexpensive (and approximate) analytical error rate expressions were derived for classifiers with multinomial models. Additional issues with regards to the curse of dimensionality that are specific to multinomial models (feature sparsity) were dealt with and tested on a text-based language identification problem for all eleven official languages of South Africa.

Description

Thesis (M.Ing. (Computer Engineering))--North-West University, Potchefstroom Campus, 2009.

Keywords

Naïve Bayesian, Maximum likelihood, Curse of dimensionality, Gaussian distribution, Multinomial distribution, Feature selection, Data sparsity, Chi-square variates, Hyperboloidal decision boundaries

URI

http://hdl.handle.net/10394/4091

Collections

Engineering

Full item page

Classification in high dimensional feature spaces

Files

Date

Authors

Researcher ID

Supervisors

Journal Title

Journal ISSN

Volume Title

Publisher

Record Identifier

Abstract

Sustainable Development Goals

Description

Keywords

Citation

URI

Collections

Endorsement

Review

Supplemented By

Referenced By