Global ETD Search

Return to search

Classification in high dimensional feature spaces / by H.O. van Dyk

In this dissertation we developed theoretical models to analyse Gaussian and multinomial distributions. The analysis is focused on classification in high dimensional feature spaces and provides a basis for dealing with issues such as data sparsity and feature selection (for Gaussian and multinomial distributions, two frequently used models for high dimensional applications). A Naïve Bayesian philosophy is followed to deal with issues associated with the curse of dimensionality. The core treatment on Gaussian and multinomial models consists of finding analytical expressions for classification error performances. Exact analytical expressions were found for calculating error rates of binary class systems with Gaussian features of arbitrary dimensionality and using any type of quadratic decision boundary (except for degenerate paraboloidal boundaries).
Similarly, computationally inexpensive (and approximate) analytical error rate expressions were derived for classifiers with multinomial models. Additional issues with regards to the curse of dimensionality that are specific to multinomial models (feature sparsity) were dealt with and tested on a text-based language identification problem for all eleven official languages of South Africa. / Thesis (M.Ing. (Computer Engineering))--North-West University, Potchefstroom Campus, 2009.

http://hdl.handle.net/10394/4091

Naïve Bayesian

Maximum likelihood

Curse of dimensionality

Gaussian distribution

Multinomial distribution

Feature selection

Data sparsity

Chi-square variates

Hyperboloidal decision boundaries

Identifer	oai:union.ndltd.org:NWUBOLOKA1/oai:dspace.nwu.ac.za:10394/4091
Date	January 2009
Creators	Van Dyk, Hendrik Oostewald
Publisher	North-West University
Source Sets	North-West University
Detected Language	English
Type	Thesis

Page generated in 0.0013 seconds

Classification in high dimensional feature spaces / by H.O. van Dyk

Description

Links & Downloads

Tags

Additional Fields