May 23, 2025

Data Science #28 - The Bloom filter algorithm

Listen Later

39 minutes

In the 28th episode, we go over Burton Bloom's Bloom filter from 1970, a groundbreaking data structure that enables fast, space-efficient set membership checks by allowing a small, controllable rate of false positives.Unlike traditional methods that store full data, Bloom filters use a compact bit array and multiple hash functions, trading exactness for speed and memory savings.

This idea transformed modern data science and big data systems, powering tools like Apache Spark, Cassandra, and Kafka, where fast filtering and memory efficiency are critical for performance at scale.

...more

View all episodes

View all episodes

Download on the App Store

Download on the App Store

Get it on Google Play

Data Science Decoded

By Mike E

3

33 ratings

May 23, 2025

Data Science #28 - The Bloom filter algorithm

Listen Later

39 minutes

In the 28th episode, we go over Burton Bloom's Bloom filter from 1970, a groundbreaking data structure that enables fast, space-efficient set membership checks by allowing a small, controllable rate of false positives.Unlike traditional methods that store full data, Bloom filters use a compact bit array and multiple hash functions, trading exactness for speed and memory savings.

This idea transformed modern data science and big data systems, powering tools like Apache Spark, Cassandra, and Kafka, where fast filtering and memory efficiency are critical for performance at scale.

...more

More shows like Data Science Decoded

Science Friday by Science Friday and WNYC Studios

Science Friday

6,133 Listeners

More or Less: Behind the Stats by BBC Radio 4

More or Less: Behind the Stats

901 Listeners

The Quanta Podcast by Quanta Magazine

The Quanta Podcast

501 Listeners

Hidden Brain by Hidden Brain, Shankar Vedantam

Hidden Brain

43,483 Listeners

Space Nuts: Astronomy Insights & Cosmic Discoveries by Professor Fred Watson and Andrew Dunkley

Space Nuts: Astronomy Insights & Cosmic Discoveries

223 Listeners

Something You Should Know by Mike Carruthers | OmniCast Media

Something You Should Know

4,171 Listeners

Super Data Science: ML & AI Podcast with Jon Krohn by Jon Krohn

Super Data Science: ML & AI Podcast with Jon Krohn

298 Listeners

The Daily by The New York Times

The Daily

111,917 Listeners

Practical AI by Practical AI LLC

Practical AI

192 Listeners

The Origins Podcast with Lawrence Krauss by Lawrence M. Krauss

The Origins Podcast with Lawrence Krauss

488 Listeners

The Supermassive Podcast by The Royal Astronomical Society

The Supermassive Podcast

287 Listeners

Machine Learning Street Talk (MLST) by Machine Learning Street Talk (MLST)

Machine Learning Street Talk (MLST)

88 Listeners

The Ancients by History Hit

The Ancients

3,049 Listeners

The Rest Is Politics by Goalhanger

The Rest Is Politics

3,289 Listeners

The Bull - Il tuo podcast di finanza personale by Riccardo Spada – Corax

The Bull - Il tuo podcast di finanza personale

17 Listeners