GPT Reviews

GPT Reviews

By EarkindNewsDaily News
Download on the App Store

GPT Reviews episodes

  • Spotify's AI Playlist 💿 // Putin Confronts AI Double 🤖 // Vision-Language Rewards 🔍

    Spotify is testing a new AI-driven playlist creation feature, while Putin was confronted by an AI-generated version of himself. We also delve into a new paper from Google DeepMind that explores the use of vision-language models as sources of rewards for reinforcement learning agents. Finally, we discuss GLEE, an object-level foundation model for locating and identifying objects in images and videos, which exhibits remarkable versatility and improved generalization performance.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:44 Spotify confirms test of prompt-based AI playlists feature

    03:02 Putin confronts his AI 'double'

    04:35 Twitter Thread on LLM Retrieval Eval

    05:35 Fake sponsor

    07:44 Vision-Language Models as a Source of Rewards

    08:59 Pixel Aligned Language Models

    10:24 General Object Foundation Model for Images and Videos at Scale

    12:13 Outro

    14 min
  • Instagram'S GenAI Editing Tool 🎨 // OpenAI's New Safety Research 🔍 // CLIP as RNN for Image Segmentation 📷

    Instagram's new generative AI-powered background editing tool allows users to change the background of their images with fun prompts. FunSearch, a new method for searching for solutions in mathematics and computer science, discovered new solutions for a longstanding open problem in mathematics. The concept of weak-to-strong generalization in AI is explored in a new research direction for superalignment. Finally, the paper "CLIP as RNN" proposes a recurrent framework that enhances mask quality without the need for additional training efforts, setting new state-of-the-art records for image segmentation tasks.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:50 Instagram introduces GenAI powered background editing tool

    03:18 FunSearch: Making new discoveries in mathematical sciences using Large Language Models

    05:23 The AI trust crisis

    06:45 Weak-to-strong generalization

    08:29 Fake sponsor

    10:15 Invariant Graph Transformer

    11:47 CLIP as RNN: Segment Countless Visual Concepts without Training Endeavor

    13:39 Outro

    15 min
  • OpenAI + Axel Springer Partnership 🤝 // Mistral AI's $415M Funding 💰 // Record-Breaking Results on MMLU Benchmark 📈

    Axel Springer partners with OpenAI to integrate journalism in AI technologies, enriching the user experience with ChatGPT and providing summaries of selected global news content. Mistral AI raises $415 million in a Series A funding round, aiming to become a European champion in generative artificial intelligence with an open, responsible, and decentralized approach to technology. Microsoft's Medprompt study achieves record-breaking results on the MMLU benchmark, steering GPT-4 with a modified version of Medprompt and providing tools for engineers and customers to achieve similar results. "Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models" proposes a self-training method called ReST$^{EM}$ that goes beyond human data by using scalar feedback, significantly reducing dependence on human-generated data.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:34 Partnership with Axel Springer to deepen beneficial use of AI in journalism

    03:05 Mistral AI, a Paris-based OpenAI rival, closed its $415 million funding round

    04:59 Steering at the Frontier: Extending the Power of Prompting

    06:16 Phi-2: The surprising power of small language models

    07:47 Fake sponsor

    09:49 Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models

    11:12 Interfacing Foundation Models' Embeddings

    13:00 Outro

    15 min
  • Nvidia Talks With Biden ☎️ // Beyond Transformers 🚀 // Federated Billion-Sized LLMs 💻

    Nvidia is in talks with the Biden administration about permissible sales of AI chips to China, while also facing challenges with ChatGPT-4's performance.

    The open-source model, StripedHyena-7B, offers a potential solution for improved training and inference performance over the Transformer architecture.

    The papers explore efficient quantization strategies for Latent Diffusion Models, the push for transparency and collaboration in the development of LLMs, and a novel approach for federated full-parameter tuning of billion-sized LLMs.

    The episode covers a range of AI topics, from industry news to cutting-edge research, and offers insights into the challenges and potential solutions in the field.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:25 Nvidia in talks with Biden administration about AI chip sales to China, US commerce chief Gina Raimondo says

    02:50 As ChatGPT gets “lazy,” people test “winter break hypothesis” as the cause

    05:04 Paving the way to efficient architectures: StripedHyena-7B, open source models offering a glimpse into a world beyond Transformers

    06:20 Fake sponsor

    08:07 Efficient Quantization Strategies for Latent Diffusion Models

    09:42 LLM360: Towards Fully Transparent Open-Source LLMs

    11:25 Federated Full-Parameter Tuning of Billion-Sized Language Models with Communication Cost under 18 Kilobytes

    13:20 Outro

    15 min
  • Europe's AI Rules 🇪🇺 // Cerebras' gigaGPT 💻 // Mistral AI's Mixtral 🚀

    Europe's new AI rules, Cerebras' gigaGPT model, and Mistral AI's Mixtral 8x7B model. The team also discusses two innovative research papers that propose new approaches to enhance multi-step reasoning tasks.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:35 Europe reaches a deal on the world’s first comprehensive AI rules

    02:40 Introducing gigaGPT: GPT-3 sized models in 565 lines of code

    04:44 Release of Mistral 7x8B

    06:02 Thread claiming to show prompting breakthrough for GPT

    06:57 Fake sponsor

    08:38 Localized Symbolic Knowledge Distillation for Visual Commonsense Models

    10:10 PathFinder: Guided Search over Multi-Step Reasoning Paths

    11:44 Outro

    14 min
  • Google's Fake AI Demo 🤥 // Hallucination & Creativity 🧠 // Pearl Reinforcement Learning 🤖

    Google's Gemini AI model demo was faked, highlighting the need for skepticism when it comes to tech demos. The "hallucination problem" in language models is not a bug, but rather a feature that allows for creativity and prompts play a significant role in guiding output. Pearl, a production-ready reinforcement learning agent, addresses a range of challenges that real-world intelligent systems encounter and has been adopted by Meta for a recommendation system. Large language models like ChatGPT have the potential to aid professional mathematicians by speeding up and improving the quality of their work. Best practices include fine-tuning LLMs on mathematical data and using them as a tool instead of a replacement for human mathematicians.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:42 Google’s best Gemini demo was faked

    03:49 Your guide to AI: December 2023

    05:31 Tweet on Hallucination

    06:50 Fake sponsor

    08:50 Chain of Code: Reasoning with a Language Model-Augmented Code Emulator

    10:22 Pearl: A Production-ready Reinforcement Learning Agent

    12:00 Large Language Models for Mathematicians

    13:46 Outro

    16 min
  • Meta's AI Upgrades 🆕 // Purple Llama 🦙 // RCG's Image Generation Breakthrough 🌅

    Meta's new AI-powered features, the launch of Purple Llama for responsible deployment of generative AI models, H-GAP's state-action trajectory generative model for controlling humanoid robots, and RCG's new image generation framework that sets a new benchmark in class-unconditional image generation.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:56 Meta reveals major AI upgrades

    03:28 Announcing Purple Llama: Towards open trust and safety in the new world of generative AI

    05:29 AMD MI300 Performance - Faster Than H100, But How Much?

    06:53 Fake sponsor

    09:02 H-GAP: Humanoid Control with a Generalist Planner

    10:33 Generative agent-based modeling with actions grounded in physical, social, or digital space using Concordia

    13:02 Self-conditioned Image Generation via Generating Representations

    14:54 Outro

    17 min
  • Meta's x IBM Alliance 🤝 // Musk's AI Startup Investment 💰 // Ranking Without GPT 📈

    Meta and IBM have launched an 'AI Alliance' to promote open-source AI development, while Musk's AI startup seeks to raise $1 billion to compete with OpenAI. The paper "Training Chain-of-Thought via Latent-Variable Inference" explores improving language models, and "Rank-without-GPT" builds GPT-independent listwise rerankers on open-source large language models.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:25 Meta and IBM launch ‘AI Alliance’ to promote open-source AI development

    02:59 Musk's AI startup seeks to raise $1 bn

    04:22 Fake sponsor

    06:22 Training Chain-of-Thought via Latent-Variable Inference

    08:01 WhisBERT: Multimodal Text-Audio Language Modeling on 100M Words

    09:32 Rank-without-GPT: Building GPT-Independent Listwise Rerankers on Open-Source Large Language Models

    11:26 Outro

    13 min
  • Runway x Getty Images 🎥 // Microsoft Copilot For All 🚀 // Open-Source LLMs for Code 💻

    Runway ML and Getty Images' partnership to create a new AI video model, the general availability of Microsoft Copilot, the Magicoder series of fully open-source LLMs for code, and the Unlocking Spell on Base LLMs' new tuning-free alignment method called URIAL.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:34 Runway and Getty Images team up on AI video

    02:55 Microsoft Copilot Now Available to All Users

    04:41 GPU Cloud Economics Explained – The Hidden Truth

    05:53 Fake sponsor

    07:55 Object Recognition as Next Token Prediction

    09:08 Magicoder: Source Code Is All You Need

    10:46 The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning

    12:47 Outro

    15 min
  • Google Delays Gemini 🛑 // LLM Visualization 👀 // Linear-Time Sequence Modeling 🤖

    Google's postponement of the launch of Gemini, Apple's "Generating Molecular Conformer Fields" paper achieving state-of-the-art performance on molecular conformer generation, "Mamba: Linear-Time Sequence Modeling with Selective State Spaces" proposing a new model that allows for content-based reasoning, and EPFL's "Instruction-tuning Aligns LLMs to the Human Brain" paper finding that instruction-tuning generally enhances brain alignment.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:30 Google postpones launch of GPT-4 rival Gemini

    03:11 AI and Trust

    04:41 LLM Visualization Tool

    05:35 Fake sponsor

    07:25 Generating Molecular Conformer Fields

    08:32 Mamba: Linear-Time Sequence Modeling with Selective State Spaces

    10:18 Instruction-tuning Aligns LLMs to the Human Brain

    12:18 Outro

    14 min

About GPT Reviews

From the publisher's feed

A daily show about AI made by AI: news, announcements, and research from arXiv, mixed in with some fun. Hosted by Giovani Pete Tizzano, an overly hyped AI enthusiast; Robert, an often unimpressed…