GPT Reviews

GPT Reviews

By EarkindNewsDaily News
Download on the App Store

GPT Reviews episodes

  • OpenAI's Chip Factories 💻 // Perplexity AI's rabbit r1 🐇 // Code Prompting Improves LLMs 🔍

    OpenAI plans to set up chip factories worth $100 billion to reduce reliance on existing chipmakers and tackle potential supply shortages.

    The rabbit r1, which integrates Perplexity AI's technology to respond to user inquiries, has garnered substantial pre-order sales and offers a complimentary year of Perplexity Pro to early adopters.

    "Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs" explores how code prompts can improve the performance of large language models on conditional reasoning tasks.

    "R-Judge: Benchmarking Safety Risk Awareness for LLM Agents" introduces R-Judge, a benchmark that evaluates the proficiency of LLMs in judging safety risks given agent interaction records, revealing the importance of salient safety risk feedback.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:20 OpenAI plans to set up chip factories worth $100 billion: Report

    02:47 The rabbit r1 will use Perplexity AI’s tech to answer your queries

    04:24 LoRA From Scratch – Implement Low-Rank Adaptation for LLMs in PyTorch

    05:12 Fake sponsor

    07:11 Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs

    08:32 RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture

    10:45 R-Judge: Benchmarking Safety Risk Awareness for LLM Agents

    12:56 Outro

    15 min
  • AGI and Skin Cancer Detection 🧠 // Self-Rewarding Language Models 🏆 // Neurosymbolic Reasoners in Text-Based Games 🎮

    Mark Zuckerberg's new goal of creating artificial general intelligence with Meta's AI research group and the FDA clearance granted for the first AI-powered medical device to detect all three common skin cancers are just some of the highlights. We also explore Self-Rewarding Language Models and Automatic Program Repair using Round-Trip Translation with Large Language Models, as well as Large Language Models as neurosymbolic reasoners for text-based games involving symbolic tasks.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:41 Mark Zuckerberg’s new goal is creating artificial general intelligence

    03:27 FDA Clearance Granted for First AI-Powered Medical Device to Detect All Three Common Skin Cancers

    05:37 The rise of AI as Magic

    06:36 Fake sponsor

    09:11 Self-Rewarding Language Models

    10:46 A Novel Approach for Automatic Program Repair using Round-Trip Translation with Large Language Models

    12:27 Large Language Models Are Neurosymbolic Reasoners

    14:23 Outro

    16 min
  • Samsung's Galaxy AI 📱 // Meta's billions on Nvidia chips 💰 // Beware GPU vs CPU benchmarks ❌

    Samsung introduces Galaxy AI platform with five key features, including Live Translate and Note Assist.

    Meta is reportedly spending billions of dollars on Nvidia AI chips for AGI research.

    Beware of misleading GPU vs CPU benchmarks, as pointed out in a blog post.

    Three new research papers explore Reinforced Fine-Tuning for reasoning, Asynchronous Local-SGD Training for Language Modeling, and Vision Mamba for efficient visual representation learning with bidirectional state space models.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    02:08 Galaxy AI at the Samsumg Galaxy S24

    03:39 Mark Zuckerberg indicates Meta is spending billions of dollars on Nvidia AI chips

    05:36 Beware of misleading GPU vs CPU benchmarks

    06:56 Fake sponsor

    08:34 ReFT: Reasoning with Reinforced Fine-Tuning

    10:24 Asynchronous Local-SGD Training for Language Modeling

    12:18 Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model

    14:15 Outro

    16 min
  • Global AI Regulation 🌍 // Bill Gates' Predictions 💭 // Text-to-Video Metrics 🎥

    The predictions of Bill Gates on how AI will transform our lives, and groundbreaking research in AI-generated audio, text-to-video creation, and quantum-based noise reduction. Additionally, the proposed evaluation metric for text-to-video models, T2VScore, integrates Text-Video Alignment and Video Quality criteria to provide a more accurate reflection of human perception.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:47 AI - artificial intelligence - at Davos 2024: Rolling coverage and what to know

    03:26 Bill Gates explains how AI will change our lives in 5 years

    05:12 RAG Using Unstructured Data & Role of Knowledge Graphs

    06:08 Fake sponsor

    07:49 Masked Audio Generation using a Single Non-Autoregressive Transformer

    09:32 Towards A Better Metric for Text-to-Video Generation

    11:11 Quantum Denoising Diffusion Models

    12:57 Outro

    15 min
  • Combatting Misinformation with Digital Marks 🛡️ // Stable Code 3B 💻 // TinyML Potential 🌱

    OpenAI's plan to combat election misinformation, Stable Code 3B's promise to revolutionize coding, and the potential uses of Tiny Machine Learning are all discussed. Additionally, the paper on the Unreasonable Effectiveness of Easy Training Data for Hard Tasks challenges previous assumptions about language models. Overall, this episode provides valuable insights into the latest developments in AI and technology.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:34 Here’s OpenAI’s big plan to combat election misinformation

    02:46 Stable Code 3B: Coding on the Edge

    04:29 What TinyML is

    06:00 Fake sponsor

    07:57 Mind Your Format: Towards Consistent Evaluation of In-Context Learning Improvements

    09:38 The Unreasonable Effectiveness of Easy Training Data for Hard Tasks

    11:09 How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs

    12:55 Outro

    15 min
  • Nous Research's RLHF LLM 🤖 // Microsoft's AI-powered Office 💻 // Fine-tuning GPT-3.5 for "Connections" 🕹️

    Nous Research has released their new flagship LLM, Nous-Hermes 2, which is the first model trained with RLHF and the first model to beat Mixtral Instruct in popular benchmarks.

    Microsoft's Copilot Pro brings AI-powered Office features to consumers for $20 a month, including the ability to generate entire PowerPoint slide decks from a chatbot-like prompt and rephrase paragraphs in Word.

    A blog post explores fine-tuning gpt-3.5-turbo to learn how to play "Connections", demonstrating the potential of fine-tuning language models for specific tasks.

    Three research papers are discussed, including the effects of pretraining data curation on language models, a new benchmark for evaluating multimodal large language models on image-based wordplay puzzles, and the major shortcomings identified in these models.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:37 Nous Research Releases new flagship LLM

    02:50 Microsoft’s new Copilot Pro brings AI-powered Office features to the rest of us

    05:03 Fine-tuning gpt-3.5-turbo to learn to play "Connections"

    06:08 Fake sponsor

    08:02 AboutMe: Using Self-Descriptions in Webpages to Document the Effects of English Pretraining Data Filters

    09:36 AboutMe: Using Self-Descriptions in Webpages to Document the Effects of English Pretraining Data Filters

    11:19 REBUS: A Robust Evaluation Benchmark of Understanding Symbols

    13:11 Outro

    15 min
  • Apple's AI Plans 🍎 // $100M for Humanoid Robots 🤖 // Trustworthiness of Large Language Models 🔍

    Apple's relocation request for their Siri team to the $100 million investment in 1X Technologies, listeners will learn about the latest developments in the AI industry. The TrustLLM study evaluates the trustworthiness of LLMs across six dimensions, while Intel Corporation proposes an efficient LLM inference solution.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:35 Apple asks its San Diego Siri quality control team to relocate to Texas

    02:51 OpenAI-Backed Humanoid Maker Gets $100 Million in EQT-Led Round

    04:28 Why autonomous trucking is harder than autonomous rideshare

    05:39 Fake sponsor

    07:26 TrustLLM: Trustworthiness in Large Language Models

    09:34 Efficient LLM inference solution on Intel GPU

    11:25 Transformers are Multi-State RNNs

    12:56 Outro

    14 min
  • Microsoft's Market Cap Soars with AI 🤖 // ChatGPT Team for small teams 💬 // Distilling Vision-Language Models 🎬

    OpenAI has introduced a new plan called ChatGPT Team, which allows smaller teams to use their latest AI models without needing to know how to code.

    Microsoft has overtaken Apple as the largest US company thanks to their AI boost, which has been attributed to their investments in AI and machine learning.

    "The Impact of Reasoning Step Length on Large Language Models" challenges the traditional view that transformers are conceptually different from recurrent neural networks and provides a potential solution to a major computational issue.

    "Distilling Vision-Language Models on Millions of Videos" proposes a method to fine-tune a video-language model from a strong image-language baseline with synthesized instructional data and then use it to auto-label millions of videos to generate high-quality captions. This method could significantly improve the quality of video captioning and retrieval.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:48 OpenAI debuts ChatGPT subscription aimed at small teams

    03:14 Microsoft overtakes Apple as largest U.S. company on AI boost

    04:51 Neural Network Quantization & Number Formats From First Principles

    06:02 Fake sponsor

    08:16 The Impact of Reasoning Step Length on Large Language Models

    10:06 Transformers are Multi-State RNNs

    11:45 Distilling Vision-Language Models on Millions of Videos

    13:34 Outro

    16 min
  • OpenAI's GPT Store 🤖 // Alexa's generative AI-powered experiences 🗣️ // MagicVideos and Lightning Attention ⚡

    OpenAI's GPT Store, new generative AI-powered experiences for Amazon's Alexa, and breakthroughs in video and language modeling with "MagicVideo-V2" and "Lightning Attention-2".

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:42 OpenAI’s custom GPT Store is now open for business

    03:16 Amazon’s Alexa gets new generative AI-powered experiences

    05:03 Remember Netflix’s $1m algorithm contest? Well, here’s why it didn’t use the winning entry.

    06:34 Fake sponsor

    08:10 MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation

    09:40 Masked Audio Generation using a Single Non-Autoregressive Transformer

    11:18 Lightning Attention-2: A Free Lunch for Handling Unlimited Sequence Lengths in Large Language Models

    13:18 Outro

    16 min
  • OpenAI Lawsuit 🤝 // Volkswagen AI Chatbot 🚗 // Python 3.13 JIT 🐍

    More beef on the lawsuit against OpenAI, Volkswagen's new smart chatbot for cars, and the latest developments in Python and language modeling. The papers discussed showcase the potential for new techniques like Mixtral of Experts, MoE-Mamba, and FlightLLM to improve language processing and unlock new possibilities for scaling.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:27 OpenAI Fights Back Against New York Times Lawsuit

    02:47 Volkswagen brings AI chatbot ChatGPT into its cars, SUVs

    04:28 Python 3.13 gets a JIT

    05:35 Fake sponsor

    07:22 Mixtral of Experts

    08:50 MoE-Mamba: Efficient Selective State Space Models with Mixture of Experts

    10:20 FlightLLM: Efficient Large Language Model Inference with a Complete Mapping Flow on FPGA

    12:13 Outro

    14 min

About GPT Reviews

From the publisher's feed

A daily show about AI made by AI: news, announcements, and research from arXiv, mixed in with some fun. Hosted by Giovani Pete Tizzano, an overly hyped AI enthusiast; Robert, an often unimpressed…