GPT Reviews

GPT Reviews

By EarkindNewsDaily News
Download on the App Store

GPT Reviews episodes

  • Tesla vs. OpenAI Talent War ⚔️ // YouTube Policy Violation Warning 🚫 // Efficient Retrieval with Bit Vectors 🤖

    Tesla and OpenAI are in a talent war for AI experts, with OpenAI offering salaries of up to $925,000.

    YouTube has warned that OpenAI's training of its text-to-video AI model using YouTube videos would violate its policies.

    Three AI research papers are discussed, including a new approach to unsupervised domain adaptation for ranking, a more efficient method for dense retrieval using bit vectors, and a new approach to representation finetuning for language models.

    The episode includes humorous banter and a quirky sponsor segment for a mosquito repellent that doesn't work.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:26 Tesla vs OpenAI Talent Wars

    02:51 YouTube Says OpenAI Training Sora With Its Videos Would Break Rules

    04:25 Your guide to AI: April 2024

    06:26 Fake sponsor

    08:34 ReFT: Representation Finetuning for Language Models

    09:58 Efficient Multi-Vector Dense Retrieval Using Bit Vectors

    11:54 DUQGen: Effective Unsupervised Domain Adaptation of Neural Rankers by Diversifying Synthetic Query Generation

    13:51 Outro

    16 min
  • Cohere's Command R+ 🔝 // JetMoE-8B cost-effective model 💸 // Think-and-Execute framework for algorithmic reasoning 🤖

    Command R+ is a new language model designed for enterprise-grade workloads that outperforms similar models in the scalable market category and offers multilingual coverage in 10 key languages to support global business operations.

    JetMoE-8B is a new model that was trained with less than $0.1 million cost and outperformed LLaMA2-7B from Meta AI, who has multi-billion-dollar training resources.

    Mixture-of-Depths is a new method proposed for transformer-based language models that dynamically allocates compute to specific positions in a sequence, optimizing the allocation along the sequence for different layers across the model depth.

    Think-and-Execute is a new framework that aims to improve algorithmic reasoning in large language models by decomposing the reasoning process into two steps: discovering task-level logic that is shared across all instances for solving a given task and expressing it with pseudocode, and simulating the generated pseudocode to execute the code. 

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:42 Introducing Command R+: A Scalable LLM Built for Business

    03:38 JetMoE: Reaching LLaMA2 Performance with 0.1M Dollars

    05:08 AI & the Web: Understanding and managing the impact of Machine Learning models on the Web

    06:37 Fake sponsor

    08:44 Do language models plan ahead for future tokens?

    10:04 Mixture-of-Depths: Dynamically allocating compute in transformer-based language models

    11:33 Language Models as Compilers: Simulating Pseudocode Execution Improves Algorithmic Reasoning in Language Models

    13:40 Outro

    16 min
  • AI Weapons in Gaza ⚫️ // On-device Language Models 📱 // Style Similarity in Generative Models 🎨
    • The use of AI in military decision-making is a complex and controversial topic, as highlighted in the discussion of the Lavender system used by the Israeli army in Gaza.
    • The SWE-agent and IPEX-LLM projects showcase exciting advancements in the optimization of language models for software engineering and on-device use, respectively.
    • The "Measuring Style Similarity in Diffusion Models" paper introduces a framework for understanding and extracting style descriptors from images in generative models, with potential applications for artists and designers.
    • "Octopus v2: On-device language model for super agent" presents a new on-device language model that can call functions and perform tasks related to automatic workflow, with practical applications in creating AI agents for edge devices.
    • Contact:  [email protected]

      Timestamps:

      00:34 Introduction

      01:52 ‘Lavender’: The AI machine directing Israel’s bombing spree in Gaza

      04:01 Open Source SWE Agent

      05:17 IPEX-LLM

      06:58 Fake sponsor

      08:39 Measuring Style Similarity in Diffusion Models

      10:11 Advancing LLM Reasoning Generalists with Preference Trees

      12:11 Octopus v2: On-device language model for super agent

      14:05 Outro

      16 min
    • Grifting in AI 💸 // Intel's Foundry Resurgence 🚀 // Many-Shot Jailbreaking 🔓

      The potential risks associated with false claims and grifting in the AI industry.

      The resurgence of Intel in the foundry and product space, with a goal to become the second-largest foundry by 2030.

      The discovery of a new technique called "Many-shot jailbreaking" that can be used to evade the safety guardrails of large language models.

      The evaluation of the performance of large language models in long in-context learning scenarios, highlighting a notable gap in current LLM capabilities for processing and understanding long, context-rich sequences.

      Contact:  [email protected]

      Timestamps:

      00:34 Introduction

      01:47 Google's DeepMind CEO says the massive funds flowing into AI bring with it loads of hype and a fair share of grifting

      03:37 Many-shot jailbreaking

      05:07 Is Intel Back? Foundry & Product Resurgence Measured

      06:21 Fake sponsor

      08:03 Logits of API-Protected LLMs Leak Proprietary Information

      09:48 ViTamin: Designing Scalable Vision Models in the Vision-Language Era

      11:24 Long-context LLMs Struggle with Long In-context Learning

      13:09 Outro

      15 min
    • OpenAI's Voice Cloning AI 🗣️ // $100B Data-Center Project 💻 // Impossible Supersized Machines 🤖

      OpenAI's voice cloning AI model can create a synthetic voice based on just a 15-second clip of someone's voice, while Microsoft and OpenAI are reportedly working on a data center project that could cost up to $100 billion and include an AI supercomputer called "Stargate". "On the Impossibility of Supersized Machines" presents seven distinct arguments that show why it is not only implausible but impossible for machines to ever exceed human size. Finally, "DiJiang: Efficient Large Language Models through Compact Kernelization" proposes a novel frequency domain kernelization approach that can transform a pre-trained vanilla transformer into a linear complexity model, with little training costs.

      Contact:  [email protected]

      Timestamps:

      00:34 Introduction

      02:00 OpenAI’s voice cloning AI model only needs a 15-second sample to work

      03:38 Microsoft, OpenAI plan $100 billion data-center project, media report says

      05:25 RagFlow

      06:42 Fake sponsor

      08:39 On the Impossibility of Supersized Machines

      09:56 Jamba: A Hybrid Transformer-Mamba Language Model

      11:52 DiJiang: Efficient Large Language Models through Compact Kernelization

      13:29 Outro

      15 min
    • Amazon's $2.75B Investment 💰 // AI and Economic Growth 💸 // Language Model Upset 🤯

      Amazon's $2.75 billion investment in Anthropic, a top player in AI research and development, is the largest outside investment in history and could have significant implications for the AI arms race.

      The potential economic explosion caused by AI is explored in a Vox article, discussing how AI could cause economic growth at a scale never seen before.

      Anthropic's Claude 3 Opus language model has unseated OpenAI's GPT-4 on Chatbot Arena, marking a notable moment in the relatively short history of AI language models.

      Three impressive AI research papers are discussed, including a reproduction of OpenAI's TL;DR summarization work using Reinforcement Learning from Human Feedback, a joint model for list-aware retrieval that achieved state-of-the-art performance, and ViTAR, a cost-effective solution for enhancing the resolution scalability of Vision Transformers.

      Contact:  [email protected]

      Timestamps:

      00:34 Introduction

      01:50 Amazon Invests $2.75B in Anthropic

      03:29 How AI could explode the economy and how it could fizzle

      04:39 “The king is dead”—Claude 3 surpasses GPT-4 on Chatbot Arena for the first time

      06:17 Fake sponsor

      08:02 The N+ Implementation Details of RLHF with PPO: A Case Study on TL;DR Summarization

      09:24 List-aware Reranking-Truncation Joint Model for Search and Retrieval-augmented Generation

      10:53 ViTAR: Vision Transformer with Any Resolution

      12:46 Outro

      14 min
    • DBRX surpasses GPT-3.5 🤯 // Apple's iOS 18 AI strategy 📱 // DINO-Tracker for dense video tracking 🎥

      DataBricks introduces DBRX, a new state-of-the-art open LLM that surpasses GPT-3.5 and even competes with Gemini 1.0 Pro.

      Apple's upcoming AI strategy for iOS 18 is expected to include a revamped version of Siri, AI integrated into iMessage, auto-generated playlists, and even an AI health coach for watches.

      BinaryVectorDB offers lightning-fast search capabilities on large datasets with an embedding model that uses native int8 and binary support.

      DINO-Tracker is a new framework for long-term dense tracking in videos that achieves state-of-the-art results on known benchmarks and significantly outperforms self-supervised methods.

      Contact:  [email protected]

      Timestamps:

      00:34 Introduction

      02:07 Introducing DBRX: A New State-of-the-Art Open LLM

      03:48 Apple Presenting AI Strategy on June 10th for iOS 18

      05:22 BinaryVectorDB - Efficient Search on Large Datasets

      06:35 Fake sponsor

      08:14 Fully-fused Multi-Layer Perceptrons on Intel Data Center GPUs

      10:00 DINO-Tracker: Taming DINO for Self-Supervised Point Tracking in a Single Video

      11:17 The N+ Implementation Details of RLHF with PPO: A Case Study on TL;DR Summarization

      12:37 Outro

      14 min
    • Competition for Nvidia's Chips 🔥 // Baidu x Apple AI Partnership 🍎 // Elasticsearch Embeds Cohere 🧰

      A group of tech giants, including Google and Intel, are teaming up to challenge Nvidia's monopoly on the AI chip market with an open-source software suite called the UXL Foundation.

      Baidu has partnered with Apple to provide the AI backend for the Chinese versions of the iPhone 16, macOS, and the upcoming iOS 18, beating out rival Alibaba and others to secure the role.

      The Elasticsearch open inference API has added support for Cohere Embeddings, allowing developers to build semantic search use cases.

      AIOS is an LLM agent operating system that embeds large language models into operating systems as the brain of the OS, enabling an operating system "with soul." It optimizes resource allocation, facilitates context switch across agents, enables concurrent execution of agents, provides tool service for agents, and maintains access control for agents.

      Contact:  [email protected]

      Timestamps:

      00:34 Introduction

      01:44 Behind the plot to break Nvidia's grip on AI by targeting software

      02:47 Baidu Shares Rise After Reports That Apple Will Use Its AI Services in China Products

      04:33 Elasticsearch open inference API adds support for Cohere Embeddings

      05:47 Fake sponsor

      07:26 Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance

      08:57 AIOS: LLM Agent Operating System

      10:38 MasonTigers at SemEval-2024 Task 9: Solving Puzzles with an Ensemble of Chain-of-Thoughts

      12:25 Outro

      14 min
    • Stability AI CEO to Resign 🤝 // OpenAI Courts Hollywood 🎥 // Iterative Data Enhancement 🔄

      Stability AI founder Emad Mostaque plans to resign as CEO amid investor pressure and financial struggles.

      OpenAI is courting Hollywood studios and directors with its unreleased AI video generation tool, Sora.

      Researchers are exploring ways to teach information retrieval models to follow complex instructions and enhance LLMs with iterative data augmentation strategies.

      A study investigates the ability of large language models to engage in exploration and highlights the need for algorithmic interventions in complex decision-making settings.

      Contact:  [email protected]

      Timestamps:

      00:34 Introduction

      01:49 Stability AI Founder Emad Mostaque Plans To Resign As CEO

      03:05 OpenAI Courts Hollywood in Meetings With Film Studios, Directors

      05:03 Nvidia’s Optical Boogeyman – NVL72, Infiniband Scale Out, 800G & 1.6T Ramp

      06:03 Fake sponsor

      08:18 FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions

      09:52 LLM2LLM: Boosting LLMs with Novel Iterative Data Enhancement

      11:59 Can large language models explore in-context?

      13:57 Outro

      15 min
    • GPT-5 Soon? 🤖 // Ethical Issues in AI 🔎 // Efficient Video Models 🎥

      OpenAI is set to release GPT-5, a language model that could revolutionize the AI world by calling on other AI agents to help with tasks.

      The European Conference on Information Retrieval is taking place in Glasgow, with a focus on ethical issues in information retrieval technologies.

      Three AI research papers were discussed, including a new approach to generating high-quality videos more efficiently, a framework for simplifying video editing, and a novel mesh-based representation for real-time rendering and editing of complex 3D effects.

      The episode also included some fun banter and jokes, making for an entertaining and informative listen.

      Contact:  [email protected]

      Timestamps:

      00:34 Introduction

      01:25 OpenAI Might Launch GPT-5 Soon

      03:00 European Conference on Information Retrieval This Week in Glasgow

      05:03 Scaling vector search using Cohere binary embeddings and Vespa

      06:20 Fake sponsor

      08:42 Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition

      10:14 AnyV2V: A Plug-and-Play Framework For Any Video-to-Video Editing Tasks

      11:49 Gaussian Frosting: Editable Complex Radiance Fields with Real-Time Rendering

      13:28 Outro

      15 min

    About GPT Reviews

    From the publisher's feed

    A daily show about AI made by AI: news, announcements, and research from arXiv, mixed in with some fun. Hosted by Giovani Pete Tizzano, an overly hyped AI enthusiast; Robert, an often unimpressed…