"The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

RAG & Beyond: Semantic Storage and Retrieval


Listen Later

Anton Troynikov, cofounder of Chroma, joins Nathan Labenz to discuss the importance of keeping the retrieval-augmented generation (RAG) loop in house, what it means for Chroma to be in “wartime” mode right now, and how much of the data going into Chroma has never been in a database before. If you need an ERP platform, check out our sponsor NetSuite: http://netsuite.com/cognitive.


SPONSORS: NetSuite | Omneky

NetSuite has 25 years of providing financial software for all your business needs. More than 36,000 businesses have already upgraded to NetSuite by Oracle, gaining visibility and control over their financials, inventory, HR, eCommerce, and more. If you're looking for an ERP platform ✅ head to NetSuite: http://netsuite.com/cognitive and download your own customized KPI checklist.


Omneky is an omnichannel creative generation platform that lets you launch hundreds of thousands of ad iterations that actually work customized across all platforms, with a click of a button. Omneky combines generative AI and real-time advertising data. Mention "Cog Rev" for 10% off.


LINKS:

Part 1 with Anton: https://youtu.be/ogy37CdIljg


X/SOCIAL:

@labenz (Nathan)

@atroyn (Anton)

@eriktorenberg (Erik)

@CogRev_Podcast


TIMESTAMPS:

(00:00:00) - Introduction by Nathan, setting up the conversation with Anton

(00:02:16) - Anton articulates Chroma's mission to build a horizontally scalable system

(00:03:06) - Rise in popularity of retrieval-augmented generation (RAG)

(00:06:03) - Chroma's focus on delivering a horizontally scalable cloud service for vector search and storage

(00:08:07) - Nathan describes his experience building a RAG application for a client profiling use case

(00:10:27) - Anton advises measuring retrieval quality and maximizing relevant information returned

(00:15:05) - Sponsors: Netsuite | Omneky

(00:17:02) - Popular use of open source vs. proprietary embedding models like Anthropic's Ada

(00:19:30) - The importance of keeping the RAG loop in house and not relying solely on external APIs

(00:27:41) - The huge amount of unstructured data that can now be processed by AI

(00:30:40) - Providing a unified interface to structured and unstructured data

(00:33:15) - Much of the data going into Chroma has never been in a database before

(00:38:47) - Categories of organizations adapting to AI: legacy, AI-native, and AI-first

(00:40:55) - Where Chroma is seeing most of its growth right now

(00:46:20) - Interpretability work like Anthropic's circuit evaluation

(00:52:23) - Anton believes new tooling can make latent spaces accessible without AI expertise

(01:06:08) - Scaling constraints between search indexes vs. application databases

(01:09:10) - Potential for time as a dimension in embedding spaces

(01:13:46) - Likelihood of missing results due to representational issues vs. approximate nearest neighbor

(01:15:22) - Automatically handling small data sets without needing elaborate indexing

(01:17:20) - Anton's perspective on whether OpenAI will build its own database

(01:19:43) - Partnering with OpenAI and other labs to increase use of their models

(01:21:19) - Anton's experiments probing GPT's reasoning abilities with Game of Life

(01:25:41) - Closing thoughts on the conversation


This show is produced by Turpentine: a network of podcasts, newsletters, and more, covering technology, business, and culture — all from the perspective of industry insiders and experts. We’re launching new shows every week, and we’re looking for industry-leading sponsors — if you think that might be you and your company, email us at [email protected].

...more
View all episodesView all episodes
Download on the App Store

"The Cognitive Revolution" | AI Builders, Researchers, and Live Player AnalysisBy Erik Torenberg, Nathan Labenz

  • 4.6
  • 4.6
  • 4.6
  • 4.6
  • 4.6

4.6

81 ratings


More shows like "The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

View all
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence) by Sam Charrington

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

429 Listeners

Practical AI by Practical AI LLC

Practical AI

196 Listeners

Last Week in AI by Skynet Today

Last Week in AI

274 Listeners

Machine Learning Street Talk (MLST) by Machine Learning Street Talk (MLST)

Machine Learning Street Talk (MLST)

90 Listeners

Dwarkesh Podcast by Dwarkesh Patel

Dwarkesh Podcast

326 Listeners

"Moment of Zen" by Erik Torenberg, Dan Romero, Antonio Garcia Martinez

"Moment of Zen"

89 Listeners

No Priors: Artificial Intelligence | Technology | Startups by Conviction

No Priors: Artificial Intelligence | Technology | Startups

103 Listeners

This Day in AI Podcast by Michael Sharkey, Chris Sharkey

This Day in AI Podcast

193 Listeners

Latent Space: The AI Engineer Podcast by swyx + Alessio

Latent Space: The AI Engineer Podcast

64 Listeners

"Upstream" with Erik Torenberg by Erik Torenberg

"Upstream" with Erik Torenberg

65 Listeners

The AI Daily Brief (Formerly The AI Breakdown): Artificial Intelligence News and Analysis by Nathaniel Whittemore

The AI Daily Brief (Formerly The AI Breakdown): Artificial Intelligence News and Analysis

421 Listeners

"In The Arena" by Turpentine

"In The Arena"

16 Listeners

"The Hill & Valley" by Jacob Helberg, Delian Asparouhov, Christian Garrett

"The Hill & Valley"

11 Listeners

"Econ 102" with Noah Smith and Erik Torenberg by Turpentine

"Econ 102" with Noah Smith and Erik Torenberg

138 Listeners

"Turpentine VC" | Venture Capital and Investing by Erik Torenberg

"Turpentine VC" | Venture Capital and Investing

20 Listeners

"Tech Finance" with Sasha Orloff: B2B Fintech | AI | Finance Tech by Puzzle, Turpentine

"Tech Finance" with Sasha Orloff: B2B Fintech | AI | Finance Tech

43 Listeners

"1 to 1000" | Scaling Startups with CEOs by Turpentine

"1 to 1000" | Scaling Startups with CEOs

2 Listeners

"The Riff" with Byrne Hobart and Erik Torenberg by Byrne Hobart, Erik Torenberg

"The Riff" with Byrne Hobart and Erik Torenberg

21 Listeners

"Live Players" with Samo Burja and Erik Torenberg by Turpentine

"Live Players" with Samo Burja and Erik Torenberg

39 Listeners

AI and I by Dan Shipper

AI and I

30 Listeners

History 102 with WhatifAltHist's Rudyard Lynch and Austin Padgett by Turpentine

History 102 with WhatifAltHist's Rudyard Lynch and Austin Padgett

99 Listeners

Emergent Behavior by Turpentine

Emergent Behavior

7 Listeners

"WhatifAlthist" | World History, Philosophy, Culture by Rudyard Lynch

"WhatifAlthist" | World History, Philosophy, Culture

59 Listeners

"Autopilot" with Will Summerlin by Will Summerlin | Turpentine

"Autopilot" with Will Summerlin

4 Listeners

"Company Breakdowns" by Turpentine

"Company Breakdowns"

4 Listeners

Training Data by Sequoia Capital

Training Data

31 Listeners

Complex Systems with Patrick McKenzie (patio11) by Patrick McKenzie

Complex Systems with Patrick McKenzie (patio11)

113 Listeners

"Second Opinion" with Christina Farr, Ash Zenooz MD & Luba Greenwood JD by Christina Farr, Luba Greenwood, Ash Zenooz

"Second Opinion" with Christina Farr, Ash Zenooz MD & Luba Greenwood JD

15 Listeners

"1 to 100" | Hypergrowth Startups Worth Joining by Turpentine, Why You Should Join

"1 to 100" | Hypergrowth Startups Worth Joining

0 Listeners

"This Won't Last" with Keith Rabois, Kevin Ryan, Logan Bartlett, and Zach Weinberg by Turpentine, Keith Rabois, Logan Bartlett, Zach Weinberg, Kevin Ryan

"This Won't Last" with Keith Rabois, Kevin Ryan, Logan Bartlett, and Zach Weinberg

14 Listeners

"Modern Relationships" by Erik Torenberg, Turpentine

"Modern Relationships"

7 Listeners