Data Science Decoded

Data Science #18 - The k-nearest neighbors algorithm (1951)


Listen Later

In the 18th episode we go over the original k-nearest neighbors algorithm;

Fix, Evelyn; Hodges, Joseph L. (1951). Discriminatory Analysis. Nonparametric Discrimination: Consistency Properties USAF School of Aviation Medicine, Randolph Field, Texas
They introduces a nonparametric method for classifying a new observation 𝑧 z as belonging to one of two distributions, 𝐹 F or 𝐺 G, without assuming specific parametric forms.
Using 𝑘 k-nearest neighbor density estimates, the paper implements a likelihood ratio test for classification and rigorously proves the method's consistency.


The work is a precursor to the modern 𝑘 k-Nearest Neighbors (KNN) algorithm and established nonparametric approaches as viable alternatives to parametric methods.

Its focus on consistency and data-driven learning influenced many modern machine learning techniques, including kernel density estimation and decision trees.


This paper's impact on data science is significant, introducing concepts like neighborhood-based learning and flexible discrimination.


These ideas underpin algorithms widely used today in healthcare, finance, and artificial intelligence, where robust and interpretable models are critical.

...more
View all episodesView all episodes
Download on the App Store

Data Science DecodedBy Mike E

  • 3
  • 3
  • 3
  • 3
  • 3

3

3 ratings


More shows like Data Science Decoded

View all
Science Friday by Science Friday and WNYC Studios

Science Friday

6,074 Listeners

More or Less: Behind the Stats by BBC Radio 4

More or Less: Behind the Stats

897 Listeners

The Quanta Podcast by Quanta Magazine

The Quanta Podcast

483 Listeners

Hidden Brain by Hidden Brain, Shankar Vedantam

Hidden Brain

43,460 Listeners

Space Nuts: Astronomy Insights & Cosmic Discoveries by Professor Fred Watson and Andrew Dunkley

Space Nuts: Astronomy Insights & Cosmic Discoveries

223 Listeners

Something You Should Know by Mike Carruthers | OmniCast Media

Something You Should Know

4,185 Listeners

Super Data Science: ML & AI Podcast with Jon Krohn by Jon Krohn

Super Data Science: ML & AI Podcast with Jon Krohn

295 Listeners

The Daily by The New York Times

The Daily

110,976 Listeners

Practical AI by Practical AI LLC

Practical AI

189 Listeners

The Origins Podcast with Lawrence Krauss by Lawrence M. Krauss

The Origins Podcast with Lawrence Krauss

488 Listeners

The Supermassive Podcast by The Royal Astronomical Society

The Supermassive Podcast

284 Listeners

Machine Learning Street Talk (MLST) by Machine Learning Street Talk (MLST)

Machine Learning Street Talk (MLST)

88 Listeners

The Ancients by History Hit

The Ancients

2,953 Listeners

The Rest Is Politics by Goalhanger

The Rest Is Politics

3,110 Listeners

The Bull - Il tuo podcast di finanza personale by Riccardo Spada

The Bull - Il tuo podcast di finanza personale

21 Listeners