
Sign up to save your podcasts
Or


Mark Grover LinkedIn Profile and Github Profile
"Hadoop Application Architectures"
"Drill to Detail Ep. 7 'Apache Spark and Hadoop Application Architectures'
Lyft Engineering Blog
"Software Engineer to Product Manager" blog by Gwen Shapira
"Introduction to the Oracle Data Integrator Topology" from the Oracle Data Integrator docs site
Apache Airflow and Amazon Kinesis homepages
"Experimentation in a Ridesharing Marketplace" by Nicholas Chamandy, Head of Data Science at Lyft
"How Uber Eats Works with Restaurants"
"Deliveroo has built a bunch of tiny kitchens to feed more hungry Londoners" - Wired.co.uk
Mark Rittman is joined by Special Guest Fangjin Yang to talk about the history of Druid, a high-performance, column-oriented, distributed data store originally developed by the team at Metamarkets to provide fast ad-hoc access to large amounts of event-level marketing data, and his work at Imply to commercialise Druid and build a suite of supporting query and data management tools.
Druid project homepage
Druid - A Real-Time Analytical Data Store (pdf)
Druid - Learning about the Druid Architecture
Imply.io homepage
Druid, Imply and Looker 5 bring OLAP Analysis to BigQuery’s Data Warehouse
Mark Rittman is joined in this 50th Episode Special by our original guest on the first episode of Drill to Detail, Stewart Bryson, to talk about developing agile BI applications using FiveTran, SnowflakeDB and Looker and his recent work developing a BI solution for Google Play Marketing using Google Data Studio and Google Cloud Platform. We're also joined later in the show by Alex Gorbachev from Pythian, our mystery guest who Stewart then interviews flawlessly armed only with a set of questions given to him as the guest was unveiled ... though be sure to listen past the final closing music for the bonus out-takes.
#115 Google Play Marketing with Dom Elliott and Stewart Bryson
The Next-Generation Jump Program
The Data Sharehouse is Here
From Data Warehouse to Data Sharehouse
Alex Gorbachev profile on Pythian.com
Mark Rittman is joined by Will Davis from Trifacta to talk about the public beta of Google Cloud Dataprep, Trifacta's data wrangling platform and topics including metadata management, data quality and data management for big data and cloud data sources.
Google Cloud Dataprep on Google Cloud Platform
"Google Cloud Dataprep: Spreadsheet-Style Data Wrangling Powered by Google Cloud Dataflow"
"A New Cloud-Based Data Prep Solution from Google & Trifacta"
Trifacta website
"A Breakthrough Approach to Exploring and Preparing Data"
Trifacta platform architecture
"Garbage In, Garbage Out: Why Data Quality Matters"
"How to Put an Effective Metadata Strategy in Place"
- Oracle Designer page on Oracle.com
- Bitmap Index page on Wikipedia
- Mondrian project page on Github
- Mondrian OLAP Server page on Wikipedia
- MultiDimensional eXpressions (MDX) page on Wikipedia
- Julian Hyde blog
- Apache Calcite project homepage
- Apache Calcite Introduction and Overview deck
- Streaming SQL presentation at Apex Big Data World 2017, Mountain View, California
From the publisher's feed
Mark Rittman is joined each episode by a special guest from the world of business intelligence, analytics and big data.