GPT Reviews

GPT Reviews

By EarkindNewsDaily News
Download on the App Store

GPT Reviews episodes

  • Jony Ive & OpenAI's AI Device 📱 // Mistral 7B Language Model 🤯 // Tom Hanks vs. AI 👥

    Jony Ive and OpenAI are teaming up to create the "iPhone of artificial intelligence" with over $1 billion in funding. Mistral 7B is the most powerful language model for its size to date, outperforming Llama 2 13B on all benchmarks. Tom Hanks warns fans about a fake dental plan ad that uses his image created with AI. Belinda, the AI research expert, discusses advancements in long-context scaling models, multi-agent motion forecasting, and vision transformers.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:36 Details emerge on Jony Ive and OpenAI’s plan to build the ‘iPhone of artificial intelligence’

    02:59 Mistral 7B is Released

    04:23 Tom Hanks says AI version of him used in dental plan ad without his consent

    05:49 Fake sponsor

    07:35 Effective Long-Context Scaling of Foundation Models

    09:14 MotionLM: Multi-Agent Motion Forecasting as Language Modeling

    10:44 Vision Transformers Need Registers

    12:39 Outro

    15 min
  • ChatGPT Browsing 🌐 // Meta's AI on Social Media 🤖 // VQ-VAE Made Simple 💡

    ChatGPT can now browse the internet to provide users with current information, but concerns about accuracy and reliability remain. Meta has introduced social profiles for its AIs, allowing users to interact with them directly on Instagram, Messenger, and WhatsApp. The paper "Finite Scalar Quantization: VQ-VAE Made Simple" proposes a simpler alternative to vector quantization in VAEs, which could be a game-changer for the field. "Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation" proposes a hybrid model that combines pixel-based and latent-based VDMs for more efficient and accurate text-to-video generation.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:46 ChatGPT can now access up to date information

    03:44 Introducing Social Profiles for Meta’s AIs

    05:27 WebGPU Technical Report

    06:29 Fake sponsor

    09:02 Finite Scalar Quantization: VQ-VAE Made Simple

    10:28 Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack

    12:09 Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation

    14:04 Outro

    16 min
  • Meta's AI Everywhere 🤖 // Factually Augmented RLHF 🤔 // Vision-language models 🌅

    OpenAI CEO Sam Altman caused a stir by posting a comment that AGI has been achieved internally, only to edit the post later to dispel the news.  Meta is bringing its AI assistant to WhatsApp, Messenger, and Instagram, and releasing dozens of AI characters based on celebrities like MrBeast and Charli D’Amelio.  "Aligning Large Multimodal Models with Factually Augmented RLHF" proposes a new approach to address the issue of misalignment between modalities in large multimodal models, achieving a remarkable 94% performance level on the LLaVA-Bench dataset. "InternLM-XComposer" is a vision-language large model that enables advanced image-text comprehension and composition, seamlessly integrating images into generated articles and consistently achieving state-of-the-art results on various benchmarks for vision-language foundational models.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:32 Sam Altman says (trolls) saying AGI has been achieved

    03:04 Meta is putting AI chatbots everywhere

    05:08 26% of the top 100 websites are now blocking GPTBot

    06:04 Fake sponsor

    08:19 Aligning Large Multimodal Models with Factually Augmented RLHF

    10:07 VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

    11:59 InternLM-XComposer: A Vision-Language Large Model for Advanced Text-image Comprehension and Composition

    13:58 Outro

    16 min
  • AI Girlfriends 🤖 // AI for Competitive Intelligence 💼 // Large Language Models & Causality 📊

    The rise of AI girlfriends and their impact on America's future population is a thought-provoking topic discussed in this episode. The use of AI for competitive intelligence and its advantages for companies is explored in-depth, with a focus on a new company called Prelaunch.com. The papers discussed in this episode shed light on important topics such as understanding large language models, training instabilities, and causality for machine learning. The insights provided by the experts on the show offer valuable perspectives on the implications of AI for various industries and society as a whole.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:45 AI girlfriends are ruining an entire generation of men

    03:06 AI Driven Competitive Intelligence Accelerates Research Productivity

    05:03 Causality for Machine Learning

    06:11 Fake sponsor

    08:02 Studying Large Language Model Generalization with Influence Functions

    09:26 Embers of Autoregression: Understanding Large Language Models Through the Problem They are Trained to Solve

    11:26 Small-scale proxies for large-scale Transformer training instabilities

    13:20 Outro

    15 min
  • ChatGPT Can Now See, Hear, and Speak 🗣️ // Amazon Invests $4B in AI Chatbot Rival 🤑 // Spotify Clones Podcasters' Voices 🎙️

    OpenAI's ChatGPT-4 can now see, hear, and speak, making it more intuitive for daily use. Amazon is investing up to $4 billion in OpenAI rival Anthropic, which is developing a new AI chatbot called Claude 2. Spotify is partnering with OpenAI to clone podcasters' voices and translate their shows into other languages. Lastly, Text2Reward is a new framework that automates the generation of dense reward functions for reinforcement learning.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:32 ChatGPT can now see, hear, and speak: Announcing GPT-4 multimodal

    03:16 Amazon to invest up to $4bn in OpenAI rival Anthropic

    04:51 Spotify is going to clone podcasters’ voices — and translate them to other languages

    06:14 Fake sponsor

    08:11 ChatGPT v Bard v Bing v Claude 2 v Aria v human-expert. How good are AI chatbots at scientific writing? (ver. 23Q3)

    09:49 DePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning

    11:35 Text2Reward: Automated Dense Reward Function Generation for Reinforcement Learning

    13:21 Outro

    15 min
  • Meta's Sassy Robot 🤖 // Google's Bard Extensions 📧 // MetaMath for Math Reasoning 🔢

    Meta's plan to release a "sassy robot" for younger users and concerns about potential implications of these chatbots. Google's new Bard Extensions feature and its performance in retrieving emails and drafting responses. Research on language models with Claude's long context window and the need for models that can effectively use information from the middle of long contexts. The development of MetaMath, a fine-tuned language model that specializes in mathematical reasoning, and its superior performance on mathematical reasoning benchmarks.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:33 Meta’s AI chatbot plan includes a ‘sassy robot’ for younger users

    03:17 I tried Google's new Bard Extensions feature which integrates with apps like Gmail. The AI assistant isn't perfect, but it has one clear strength.

    04:59 Prompt engineering for Claude's long context window

    06:42 Fake sponsor

    08:30 Lost in the Middle: How Language Models Use Long Contexts

    09:52 Evaluating Large Language Models for Document-grounded Response Generation in Information-Seeking Dialogues

    11:21 MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models

    13:22 Outro

    15 min
  • OpenAI Wins Lawsuit 🧑‍⚖️ // Human vs. Model Carbon Footprint 🌿 // Chain-of-Verification ⛓️

    OpenAI's privacy lawsuit has been dismissed, but the company is still facing legal controversies. YouTube is introducing new AI-powered tools for creators, including AI-generated backgrounds and personalized music recommendations. We also discuss the potential impact of open source AI on language and image models, and a new study that shows AI systems have a much lower carbon footprint than humans when it comes to tasks like writing and illustrating.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:25 A lawsuit alleging privacy violations by OpenAI was dismissed

    03:01 YouTube is going all in on AI

    04:51 Why Open Source AI Will Win

    06:30 Fake sponsor

    08:42 Chain-of-Verification Reduces Hallucination in Large Language Models

    10:02 Kosmos-2.5: A Multimodal Literate Model

    11:31 The Carbon Emissions of Writing and Illustrating Are Lower for AI than for Humans

    13:37 Outro

    15 min
  • DALL-E 3 meets ChatGPT 🖼️ // Google's BART Extensions for Workplace 📈 // Neuralink's Clinical Trial 🧠

    OpenAI's latest release of DALL-E 3 integrates with ChatGPT, making it more accessible for people who struggle with prompts. Google's BART Extensions for the Workplace uses a reinforced learning model to pull relevant information from all their Google tools and services, potentially setting it apart from its competitors. Neuralink has opened recruitment for their first-in-human clinical trial for their brain-computer interface, which aims to help those with paralysis control external devices with their thoughts. The research papers discussed in this episode explore the compression capabilities of large language models, a new text generation method called Contrastive Decoding, and Compositional Foundation Models for Hierarchical Planning, which leverages multiple expert foundation models to solve long-horizon tasks.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:38 OpenAI releases third version of DALL-E

    03:04 Google Announces BART Extensions for the Workplace

    04:31 Neuralink’s First-in-Human Clinical Trial is Open for Recruitment

    06:02 Fake sponsor

    08:06 Language Modeling Is Compression

    09:46 Contrastive Decoding Improves Reasoning in Large Language Models

    11:24 Compositional Foundation Models for Hierarchical Planning

    13:12 Outro

    15 min
  • Microsoft's Data Scandal 💻 // UK's AI Principles for Transparency 🇬🇧 // Efficient Language Models 🚀

    Microsoft's AI research team accidentally exposed 38 terabytes of private data while publishing open-source training data on GitHub, posing a significant security risk. The UK's new AI principles focus on accountability and transparency, seeking views from leading AI developers and governments to ensure the development and use of foundation models evolves in a way that promotes competition and protects consumers. Two new papers explore ways to improve the efficiency and quality of large language models, including a new inference scheme called self-speculative decoding and the ability to prune pretraining data while still retaining performance. A third paper introduces a new type of prompt called the "Chain of Density" or CoD, which generates increasingly dense summaries without increasing their length, resulting in more abstractive and human-preferred summaries.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:45 38TB of data accidentally exposed by Microsoft AI researchers

    03:05 UK focuses on transparency and access with new AI principles

    04:40 Jason Wei Tweet on the role of task-specific LLMs

    06:11 Fake sponsor

    07:43 Draft & Verify: Lossless Large Language Model Acceleration via Self-Speculative Decoding

    09:20 When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

    10:55 From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting

    12:37 Outro

    14 min
  • Amazon's Kindle Limits 📖 // Interactive AI Future 🤖 // Scaling Sparsely-Connected Models 🔍

    Amazon is limiting new Kindle books due to the rapid evolution of generative AI, which has flooded the market with low-quality content. DeepMind's co-founder believes that interactive AI is the future, which can carry out tasks by calling on other software and people to get things done. "Compositional Foundation Models for Hierarchical Planning" proposes a solution for effective decision-making in novel environments with long-horizon goals. "Scaling Laws for Sparsely-Connected Foundation Models" explores the impact of parameter sparsity on the scaling behavior of transformers trained on massive datasets, which can lead to more efficient and scalable models in the future.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:42 Citing “rapid evolution of generative AI,” Amazon limits new Kindle books

    02:56 DeepMind’s cofounder: Generative AI is just a phase. What’s next is interactive AI.

    05:10 Mitigating LLM Hallucinations: a multifaceted approach

    06:27 Fake sponsor

    08:52 Compositional Foundation Models for Hierarchical Planning

    10:39 Replacing softmax with ReLU in Vision Transformers

    11:56 Scaling Laws for Sparsely-Connected Foundation Models

    13:37 Outro

    16 min

About GPT Reviews

From the publisher's feed

A daily show about AI made by AI: news, announcements, and research from arXiv, mixed in with some fun. Hosted by Giovani Pete Tizzano, an overly hyped AI enthusiast; Robert, an often unimpressed…