Structure from Randomness in Halfspace Learning with the Zero-One Loss

Ata Kaban; Robert J. Durrant

Structure from Randomness in Halfspace Learning with the Zero-One Loss

Ata Kaban, Robert J. Durrant

Computer Science

Research output: Contribution to journal › Article › peer-review

199 Downloads (Pure)

Abstract

We prove risk bounds for halfspace learning when the data dimensionality is allowed to be larger than the sample size, using a notion of compressibility by random projection. In particular, we give upper bounds for the empirical risk minimizer learned efficiently from randomly projected data, as well as uniform upper bounds in the full high-dimensional space. Our main findings are the following: i) In both settings, the obtained bounds are able to discover and take advantage of benign geometric structure, which turns out to depend on the cosine similarities between the classifier and points of the input space, and provide a new interpretation of margin distribution type arguments. ii) Furthermore our bounds allow us to draw new connections between several existing successful classification algorithms, and we also demonstrate that our theory is predictive of empirically observed performance in numerical simulations and experiments. iii) Taken together, these results suggest that the study of compressive learning can improve our understanding of which benign structural traits - if they are possessed by the data generator - make it easier to
learn an effective classifier from a sample.

Original language	English
Number of pages	32
Journal	Journal of Artificial Intelligence Research
Publication status	Accepted/In press - 15 Sept 2020

Access to Document

Structure from RandomnessAccepted author manuscript, 463 KB

Cite this

@article{d091fb88b8454962aa8978f870be26ba,

title = "Structure from Randomness in Halfspace Learning with the Zero-One Loss",

abstract = "We prove risk bounds for halfspace learning when the data dimensionality is allowed to be larger than the sample size, using a notion of compressibility by random projection. In particular, we give upper bounds for the empirical risk minimizer learned efficiently from randomly projected data, as well as uniform upper bounds in the full high-dimensional space. Our main findings are the following: i) In both settings, the obtained bounds are able to discover and take advantage of benign geometric structure, which turns out to depend on the cosine similarities between the classifier and points of the input space, and provide a new interpretation of margin distribution type arguments. ii) Furthermore our bounds allow us to draw new connections between several existing successful classification algorithms, and we also demonstrate that our theory is predictive of empirically observed performance in numerical simulations and experiments. iii) Taken together, these results suggest that the study of compressive learning can improve our understanding of which benign structural traits - if they are possessed by the data generator - make it easier tolearn an effective classifier from a sample.",

author = "Ata Kaban and Durrant, {Robert J.}",

year = "2020",

month = sep,

day = "15",

language = "English",

journal = "Journal of Artificial Intelligence Research",

issn = "1076-9757",

publisher = "Morgan Kaufmann Publishers, Inc.",

}

TY - JOUR

T1 - Structure from Randomness in Halfspace Learning with the Zero-One Loss

AU - Kaban, Ata

AU - Durrant, Robert J.

PY - 2020/9/15

Y1 - 2020/9/15

N2 - We prove risk bounds for halfspace learning when the data dimensionality is allowed to be larger than the sample size, using a notion of compressibility by random projection. In particular, we give upper bounds for the empirical risk minimizer learned efficiently from randomly projected data, as well as uniform upper bounds in the full high-dimensional space. Our main findings are the following: i) In both settings, the obtained bounds are able to discover and take advantage of benign geometric structure, which turns out to depend on the cosine similarities between the classifier and points of the input space, and provide a new interpretation of margin distribution type arguments. ii) Furthermore our bounds allow us to draw new connections between several existing successful classification algorithms, and we also demonstrate that our theory is predictive of empirically observed performance in numerical simulations and experiments. iii) Taken together, these results suggest that the study of compressive learning can improve our understanding of which benign structural traits - if they are possessed by the data generator - make it easier tolearn an effective classifier from a sample.

AB - We prove risk bounds for halfspace learning when the data dimensionality is allowed to be larger than the sample size, using a notion of compressibility by random projection. In particular, we give upper bounds for the empirical risk minimizer learned efficiently from randomly projected data, as well as uniform upper bounds in the full high-dimensional space. Our main findings are the following: i) In both settings, the obtained bounds are able to discover and take advantage of benign geometric structure, which turns out to depend on the cosine similarities between the classifier and points of the input space, and provide a new interpretation of margin distribution type arguments. ii) Furthermore our bounds allow us to draw new connections between several existing successful classification algorithms, and we also demonstrate that our theory is predictive of empirically observed performance in numerical simulations and experiments. iii) Taken together, these results suggest that the study of compressive learning can improve our understanding of which benign structural traits - if they are possessed by the data generator - make it easier tolearn an effective classifier from a sample.

M3 - Article

SN - 1076-9757

JO - Journal of Artificial Intelligence Research

JF - Journal of Artificial Intelligence Research

ER -

Structure from Randomness in Halfspace Learning with the Zero-One Loss

Abstract

Access to Document

Fingerprint

Cite this