Interconnects

Interconnects

By Nathan LambertScienceTechnology
Download on the App Store

Interconnects episodes

  • OpenAI's Model (behavior) Spec, RLHF transparency, and personalization questions

    Now we will have some grounding for when weird ChatGPT behaviors are intended or side-effects -- shrinking the Overton window of RLHF bugs.
    This is AI generated audio with Python and 11Labs.
    Source code: https://github.com/natolambert/interconnects-tools
    Original post: https://www.interconnects.ai/p/openai-rlhf-model-spec

    00:00 OpenAI's Model (behavior) Spec, RLHF transparency, and personalization questions
    02:56 Reviewing the Model Spec
    08:26 Where RLHF can fail OpenAI
    12:23 From Model Spec's to personalization

    Fig 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/model-spec/img_027.png
    Fig 2: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/model-spec/img_029.png
    Fig 3: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/model-spec/img_033.png
    Fig 4: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/model-spec/img_034.png
    Fig 5: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/model-spec/img_041.webp
    Fig 6: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/model-spec/img_046.webp



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    15 min
  • RLHF: A thin line between useful and lobotomized

    Many, many signs of life for preference fine-tuning beyond spoofing chat evaluation tools.
    This is AI generated audio with Python and 11Labs.
    Source code: https://github.com/natolambert/interconnects-tools
    Original post: https://www.interconnects.ai/p/how-rlhf-works-2

    00:00 How RLHF works, part 2: A thin line between useful and lobotomized
    04:27 The chattiness paradox
    08:09 The mechanism for making models chattier
    10:42 Next steps for RLHF research

    Fig 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/rlhf/img_012.webp
    Fig 2: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/rlhf/img_018.png
    Fig 3: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/rlhf/img_025.png



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    14 min
  • Phi 3 and Arctic: Outlier LMs are hints

    Models that seem totally out of scope from recent open LLMs give us a sneak peek of where the industry will be in 6 to 18 months.
    This is AI generated audio with Python and 11Labs.
    Source code: https://github.com/natolambert/interconnects-tools
    Original post: https://www.interconnects.ai/p/phi-3-and-arctic-llms

    0:00 Phi 3 and Arctic: Outlier LMs are hints
    1:01 Arctic & open mixture of expert trends
    6:10 Phi 3, synthetic data, and small models

    Fig 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/phi3/img_004.png
    Fig 2: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/phi3/img_008.png
    Fig 3: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/phi3/img_018.png



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    10 min
  • AGI is what you want it to be

    Certain definitions of AGI are backing people into a pseudo-religious corner.
    This is AI generated audio with Python and 11Labs.
    Source code: https://github.com/natolambert/interconnects-tools
    Original post: https://www.interconnects.ai/p/agi-is-what-you-want-it-to-be

    00:00 AGI is what you want it to be
    04:01 RL still rules the AGI discourse
    05:43 Modern AGI tests
    07:37 Agency and shifting goalposts

    Fig 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/agi/img_018.png
    Fig 2: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/agi/img_020.png



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    11 min
  • Llama 3: Scaling open LLMs to AGI

    Meta shows that scaling won't be a limit for open LLM players in the near future.
    This is AI generated audio with Python and 11Labs.
    Source code: https://github.com/natolambert/interconnects-tools
    Original post: https://www.interconnects.ai/p/llama-3-and-scaling-open-llms

    00:00 Llama 3; scaling open LLMs to AGI
    01:44 Pretraining, data, and basic evals
    06:06 Alignment and human evaluations
    10:08 Chatting with Meta AI and Llama 3 70B Instruct
    11:55 Same Llama license (mostly)
    12:52 The healthy open LLM ecosystem

    Fig 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_011.jpeg
    Fig 2: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_013.png
    Fig 3: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_015.png
    Fig 4: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_020.png
    Fig 5: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_036.png
    Fig 6: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_040.png
    Fig 7: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_046.jpeg
    Fig 8: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_061.png
    Fig 9: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_063.webp
    Fig 10: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_066.png
    Fig 11: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/llama3/img_068.jpeg



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    16 min
  • Stop "reinventing" everything to "solve" alignment

    Integrating some non computing science into reinforcement learning from human feedback can give us the models we want.
    This is AI generated audio with Python and 11Labs.
    Source code: https://github.com/natolambert/interconnects-tools
    Original post: https://www.interconnects.ai/p/reinventing-llm-alignment

    0:00 Stop "reinventing" everything to "solve" AI alignment
    2:19 Social Choice for AI Alignment: Dealing with Diverse Human Feedback
    7:03 OLMo 1.7 7B: A truly open model with actually good benchmarks


    Fig 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/reinvention/img_013.png
    Fig 2: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/reinvention/img_015.png
    Fig 3: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/reinvention/img_018.png
    Fig 4: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/reinvention/img_024.png
    Fig 5: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/reinvention/img_027.png



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    8 min
  • The end of the "best open LLM"

    Modeling the compute versus performance tradeoff of many open LLMs.
    This is AI generated audio with Python and 11Labs.
    Source code: https://github.com/natolambert/interconnects-tools
    Original post: https://www.interconnects.ai/p/compute-efficient-open-llms

    0:00 The end of the "best open LLM"
    3:05 Compute efficient open LLMs

    Fig 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/scaling/img_004.jpeg
    Fig 2: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/scaling/img_009.png
    Fig 3: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/scaling/img_014.png
    Fig 4: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/scaling/img_016.png
    Fig 5: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/scaling/img_018.png
    Fig 6: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/scaling/img_020.png
    Fig 7: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/scaling/img_022.png
    Fig 8: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/scaling/img_024.png
    Fig 9: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/scaling/img_028.png



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    7 min
  • Why we disagree on what open-source AI should be

    Last minute title change from: The tech industry can't agree on what open-source AI means. That's the process.
    How to read what multiple people mean by the word openness and see through the PR speak.
    This is AI generated audio with Python and 11Labs.
    Source code: https://github.com/natolambert/interconnects-tools
    Original post: https://www.interconnects.ai/p/flavors-of-open-source-ai

    0:00 The tech industry can't agree on what open-source AI means. That's the process.
    2:45 1. Effective Accelerationists, Techno-Optimists, capitalists, etc.
    3:39 2. Scientists, promoting understanding and transparency
    5:16 3. Inclusion, public interest, and fighting concentration of power
    6:19 4. Freedom advocates
    7:25 Dissecting "openness"

    Fig 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/openness/img_004.png



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    9 min
  • DBRX: The new best open LLM and Databricks' ML strategy

    Databricks' new model is surpassing the performance of Mixtral and Llama 2 while still being in a size category that's reasonably accessible.
    This is AI generated audio with Python and 11Labs.
    Source code: https://github.com/natolambert/interconnects-tools
    https://www.interconnects.ai/p/databricks-dbrx-open-llm

    00:00 DBRX: The new best open model and Databricks' ML strategy
    03:36 The DBRX narrative
    07:33 Databricks' open LLM (and AI) strategy
    09:42 Playing with DBRX Instruct
    14:54 Digging for details

    Fig 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/dbrx/img_007.png
    Fig 2: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/dbrx/img_012.png
    Fig 3: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/dbrx/img_023.png
    Fig 4: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/dbrx/img_045.png
    Fig 5: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/dbrx/img_047.png
    Fig 6: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/dbrx/img_059.png
    Fig 7: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/dbrx/img_066.jpeg
    Fig 8: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/dbrx/img_068.png



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    17 min
  • Evaluations: Trust, performance, and price (bonus, announcing RewardBench)

    Evaluation is not only getting harder with modern LLMs, it's getting harder because it means something different.
    This is AI generated audio with Python and 11Labs. Music generated by Meta's MusicGen.
    Source code: https://github.com/natolambert/interconnects-tools
    Original post: https://www.interconnects.ai/p/evaluations-trust-performance-and-price

    00:00 Evaluations: Trust, performance, and price (bonus, announcing RewardBench)
    03:14 The rising price of evaluation
    05:40 Announcing RewardBench: The First reward model evaluation tool
    08:37 Updates to RLHF evaluation tools

    YouTube code intro: https://youtu.be/CAaHAfCqrBA

    Figure 1: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/evals/img_026.png
    Figure 2: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/evals/img_030.png
    Figure 3: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/evals/img_034.png
    Figure 4: https://huggingface.co/datasets/natolambert/interconnects-figures/resolve/main/evals/img_040.png



    This is a public episode. If you'd like to discuss this with other subscribers or get access to bonus episodes, visit www.interconnects.ai/subscribe
    13 min

About Interconnects

From the publisher's feed

Audio essays about the latest developments in AI and interviews with leading scientists in the field. Breaking the hype, understanding what's under the hood, and telling stories.

More shows like Interconnects

The Twenty Minute VC (20VC): Venture Capital | Startup Funding | The Pitch by Harry Stebbings

The Twenty Minute VC (20VC): Venture Capital | Startup Funding | The Pitch

542 Listeners

The a16z Show by Andreessen Horowitz

The a16z Show

1,089 Listeners

ChinaTalk by Jordan Schneider

ChinaTalk

288 Listeners

Practical AI by Daniel Whitenack and Chris Benson

Practical AI

203 Listeners

Google DeepMind: The Podcast by Hannah Fry

Google DeepMind: The Podcast

205 Listeners

Last Week in AI by Skynet Today

Last Week in AI

315 Listeners

Machine Learning Street Talk (MLST) by Machine Learning Street Talk (MLST)

Machine Learning Street Talk (MLST)

98 Listeners

Dwarkesh Podcast by Dwarkesh Patel

Dwarkesh Podcast

565 Listeners

No Priors: Artificial Intelligence | Technology | Startups by Conviction

No Priors: Artificial Intelligence | Technology | Startups

141 Listeners

Latent Space: The AI Engineer Podcast by Latent.Space

Latent Space: The AI Engineer Podcast

102 Listeners

This Day in AI Podcast by Michael Sharkey, Chris Sharkey

This Day in AI Podcast

222 Listeners

"Econ 102" with Noah Smith and Erik Torenberg by Turpentine

"Econ 102" with Noah Smith and Erik Torenberg

145 Listeners

BG2Pod with Brad Gerstner and Bill Gurley by BG2Pod

BG2Pod with Brad Gerstner and Bill Gurley

455 Listeners

AI + a16z by a16z

AI + a16z

30 Listeners

Training Data by Sequoia Capital

Training Data

39 Listeners