OpenAI model escaped its sandbox, but is that an almighty model or just bad engineering? We unpack why engineering failures will keep happening and what reliable AI systems need to do differently.
We also cover why the battle between model providers is all going to come down to who can do inference most profitably, which, as of right now, gives Google a big advantage. We also debate whether we can ever build the primitives needed for AI to actually help with large chunks of knowledge work. And, if so, how.
Vibes & Benchmarks is a weekly AI news podcast hosted by Ali Rohde (Outset Capital) and Josh Albrecht (Imbue + Outset Capital). Each week, they pick apart the latest news, from startups to SpaceX to OpenAI. They argue through what is actually new, what is overfit to Twitter, and what might still matter 6 months from now.
Spotify: https://open.spotify.com/show/033s9Dt...
Apple Podcasts: https://podcasts.apple.com/us/podcast...
Substack: https://outsetcapital.substack.com