Data Engineering Podcast

ArangoDB: Fast, Scalable, and Multi-Model Data Storage with Jan Steeman and Jan Stücke - Episode 34


Listen Later

Summary

Using a multi-model database in your applications can greatly reduce the amount of infrastructure and complexity required. ArangoDB is a storage engine that supports documents, dey/value, and graph data formats, as well as being fast and scalable. In this episode Jan Steeman and Jan Stücke explain where Arango fits in the crowded database market, how it works under the hood, and how you can start working with it today.

Preamble
  • Hello and welcome to the Data Engineering Podcast, the show about modern data management
  • When you’re ready to build your next pipeline you’ll need somewhere to deploy it, so check out Linode. With private networking, shared block storage, node balancers, and a 40Gbit network, all controlled by a brand new API you’ve got everything you need to run a bullet-proof data platform. Go to dataengineeringpodcast.com/linode to get a $20 credit and launch a new server in under a minute.
  • Go to dataengineeringpodcast.com to subscribe to the show, sign up for the newsletter, read the show notes, and get in touch.
  • Your host is Tobias Macey and today I’m interviewing Jan Stücke and Jan Steeman about ArangoDB, a multi-model distributed database for graph, document, and key/value storage.
  • Interview
    • Introduction
    • How did you get involved in the area of data management?
    • Can you give a high level description of what ArangoDB is and the motivation for creating it?
      • What is the story behind the name?

      • How is ArangoDB constructed?

        • How does the underlying engine store the data to allow for the different ways of viewing it?

        • What are some of the benefits of multi-model data storage?

          • When does it become problematic?

          • For users who are accustomed to a relational engine, how do they need to adjust their approach to data modeling when working with Arango?

          • How does it compare to OrientDB?

          • What are the options for scaling a running system?

            • What are the limitations in terms of network architecture or data volumes?

            • One of the unique aspects of ArangoDB is the Foxx framework for embedding microservices in the data layer. What benefits does that provide over a three tier architecture?

              • What mechanisms do you have in place to prevent data breaches from security vulnerabilities in the Foxx code?
              • What are some of the most interesting or surprising uses of this functionality that you have seen?

              • What are some of the most challenging technical and business aspects of building and promoting ArangoDB?

              • What do you have planned for the future of ArangoDB?

              • Contact Info
                • Jan Steemann
                  • jsteemann on GitHub
                  • @steemann on Twitter

                  • Parting Question
                    • From your perspective, what is the biggest gap in the tooling or technology for data management today?
                    • Links

                      • ArangoDB
                      • Köln
                      • Multi-model Database
                      • Graph Algorithms
                      • Apache 2
                      • C++
                      • ArangoDB Foxx
                      • Raft Protocol
                      • Target Partners
                      • RocksDB
                      • AQL (ArangoDB Query Language)
                      • OrientDB
                      • PostGreSQL
                      • OrientDB Studio
                      • Google Spanner
                      • 3-Tier Architecture
                      • Thomson-Reuters
                      • Arango Search
                      • Dell EMC
                      • Google S2 Index
                      • ArangoDB Geographic Functionality
                      • JSON Schema
                      • The intro and outro music is from The Hug by The Freak Fandango Orchestra / CC BY-SA

                        Support Data Engineering Podcast

                        ...more
                        View all episodesView all episodes
                        Download on the App Store

                        Data Engineering PodcastBy Tobias Macey

                        • 4.5
                        • 4.5
                        • 4.5
                        • 4.5
                        • 4.5

                        4.5

                        142 ratings


                        More shows like Data Engineering Podcast

                        View all
                        The Changelog: Software Development, Open Source by Changelog Media

                        The Changelog: Software Development, Open Source

                        289 Listeners

                        Software Engineering Daily by Software Engineering Daily

                        Software Engineering Daily

                        624 Listeners

                        Talk Python To Me by Michael Kennedy

                        Talk Python To Me

                        583 Listeners

                        Super Data Science: ML & AI Podcast with Jon Krohn by Jon Krohn

                        Super Data Science: ML & AI Podcast with Jon Krohn

                        302 Listeners

                        NVIDIA AI Podcast by NVIDIA

                        NVIDIA AI Podcast

                        343 Listeners

                        Practical AI by Practical AI LLC

                        Practical AI

                        204 Listeners

                        AWS Podcast by Amazon Web Services

                        AWS Podcast

                        205 Listeners

                        Last Week in AI by Skynet Today

                        Last Week in AI

                        305 Listeners

                        Dwarkesh Podcast by Dwarkesh Patel

                        Dwarkesh Podcast

                        523 Listeners

                        The Data Engineering Show by The Firebolt Data Bros

                        The Data Engineering Show

                        8 Listeners

                        No Priors: Artificial Intelligence | Technology | Startups by Conviction

                        No Priors: Artificial Intelligence | Technology | Startups

                        129 Listeners

                        Latent Space: The AI Engineer Podcast by swyx + Alessio

                        Latent Space: The AI Engineer Podcast

                        92 Listeners

                        This Day in AI Podcast by Michael Sharkey, Chris Sharkey

                        This Day in AI Podcast

                        227 Listeners

                        The AI Daily Brief: Artificial Intelligence News and Analysis by Nathaniel Whittemore

                        The AI Daily Brief: Artificial Intelligence News and Analysis

                        633 Listeners

                        AI + a16z by a16z

                        AI + a16z

                        36 Listeners