GPT Reviews

GPT Reviews

By EarkindNewsDaily News
Download on the App Store

GPT Reviews episodes

  • AI & Antibiotics 💊 // Audio Open Source 🎶 // AI Ethics & Robotics 🤖

    AI used to predict potential new antibiotics in groundbreaking study.

    Stable Audio Open: an open source model that allows users to create short audio samples and sound effects from text prompts.

    The ethical responsibilities of AI researchers when it comes to warning about the dangers of advanced artificial intelligence.

    Cutting-edge research on AI and robotics, including large-scale simulations, in-context learning, and skill composition in modular arithmetic tasks.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:21 AI used to predict potential new antibiotics in groundbreaking study

    02:40 Introducing Stable Audio Open - An Open Source Model for Audio Samples and Sound Design

    04:22 A Right to Warn about Advanced Artificial Intelligence

    05:23 Fake sponsor

    07:10 RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots

    08:56 Learning to grok: Emergence of in-context learning and skill composition in modular arithmetic tasks

    10:36 Guiding a Diffusion Model with a Bad Version of Itself

    12:24 Outro

    14 min
  • Amazon AI Detects Damaged Goods 📦 // Musk Prioritizes xAI 🚘 // Uncertainty in LLMs 🤔

    Amazon's new AI system to detect damaged or incorrect items before they ship.

    Elon Musk's controversial decision to prioritize X and xAI over Tesla for AI chips.

    "To Believe or Not to Believe Your LLM" paper on uncertainty quantification in Large Language Models.

    "Guiding a Diffusion Model with a Bad Version of Itself" paper on improving image generation with diffusion models.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:47 Learn how Amazon uses AI to spot damaged products before they’re shipped to customers

    03:17 Elon Musk ordered Nvidia to ship thousands of AI chips reserved for Tesla to X and xAI

    05:08 FineWeb: decanting the web for the finest text data at scale

    06:16 Fake sponsor

    08:05 To Believe or Not to Believe Your LLM

    09:33 Guiding a Diffusion Model with a Bad Version of Itself

    11:06 Learning to grok: Emergence of in-context learning and skill composition in modular arithmetic tasks

    12:54 Outro

    15 min
  • Microsoft's Latest $3.2B AI Investment 🇸🇪 // Grokfast Algorithm 💪 // Zipper Decoder Architecture 🎧

    Microsoft is investing $3.2 billion in Sweden for cloud and AI infrastructure, deploying 20,000 advanced graphics processing units and training 250,000 Swedes with AI skills over three years.

    "Grokfast" is a new algorithm that accelerates generalization under the grokking phenomenon in machine learning by amplifying the slow-varying component of gradients, improving performance on tasks like image classification.

    "Zipper" is a multi-tower decoder architecture that uses cross-attention to flexibly compose multimodal generative models from independently pre-trained unimodal decoders, showcasing superior performance in tasks like speech-to-text generation.

    "MetRag" is a new framework for retrieval augmented generation that combines similarity and utility-oriented models, using an LLM as a task adaptive summarizer to generate knowledge-augmented text and outperforming existing models on knowledge-intensive tasks like finance and medicine.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:49 Microsoft to invest $3.2 bln in Swedish cloud, AI

    03:42 State Space Duality (Mamba-2) Part I - The Model

    04:47 Sam Altman, Lately

    06:08 Fake sponsor

    08:39 Grokfast: Accelerated Grokking by Amplifying Slow Gradients

    10:11 Zipper: A Multi-Tower Decoder Architecture for Fusing Modalities

    11:38 Similarity is Not All You Need: Endowing Retrieval Augmented Generation with Multi Layered Thoughts

    13:52 Outro

    15 min
  • Nvidia's AI Factories 🏭 // AI Gadget for Recycling 🌍 // Intellectual Obesity Crisis 📚

    Nvidia unveils plans to accelerate the advance of artificial intelligence, partnering with companies and countries to build AI factories and releasing Nvidia ACE generative AI.

    Finnish startup Binit develops an AI gadget that tracks household waste to encourage recycling, with potential benefits in improving recycling efficiency.

    "The Intellectual Obesity Crisis" article discusses how we've become addicted to useless information, just like we evolved to crave sugar because it was a scarce source of energy.

    Three AI research papers are discussed, including a method to compress second-order optimizer states to lower bitwidths, the first-ever full-spectrum, multi-modal evaluation benchmark of MLLMs in video analysis, and a theoretical connection between Transformers and state-space models leading to a faster and more efficient alternative to existing models.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:21 AI hardware firm Nvidia unveils next-gen products at Taiwan tech expo

    02:48 Binit is bringing AI to trash

    04:39 The Intellectual Obesity Crisis

    06:22 Fake sponsor

    08:18 4-bit Shampoo for Memory-Efficient Network Training

    10:01 Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    11:51 Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

    13:30 Outro

    15 min
  • Google's Apology 🤖 // Nvidia's Top-Ranked Embedding Model 🥇 // Matryoshka Query Transformer 🌟

    Google's AI Overviews are improving to provide accurate and helpful information.

    Nvidia's new embedding model, NV-Embed-v1, ranks number one on the Massive Text Embedding Benchmark.

    Matryoshka Query Transformer (MQT) offers flexibility to Large Vision-Language Models (LVLMs) by encoding an image into a variable number of visual tokens during inference.

    Contextual Position Encoding (CoPE) improves the position encoding method in Large Language Models (LLMs) and solves tasks where popular position embeddings fail. 

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:35 AI Overviews: About last week

    03:58 Nvidia Releases Embedding Model NV-Embed-v1

    04:53 Multi-camera YOLOv5 on Zynq UltraScale+ with Hailo-8 AI Acceleration

    06:31 Fake sponsor

    08:28 Matryoshka Query Transformer for Large Vision-Language Models

    10:24 Similarity is Not All You Need: Endowing Retrieval Augmented Generation with Multi Layered Thoughts

    11:51 Contextual Position Encoding: Learning to Count What's Important

    13:30 Outro

    15 min
  • OpenAI Partnerships 🤝 // Codestral Model for Coding 🤖 // Transparent Language Models 🔍

    OpenAI announces new content and product partnerships with Vox Media and The Atlantic, making their reporting and stories more discoverable to millions of OpenAI users.

    Mistral AI releases Codestral, a 22B parameter, open-weight model that specializes in coding tasks, beating out its code-focused rivals across top benchmarks.

    MAP-Neo is the first fully open-sourced bilingual LLM that provides all the details needed to reproduce the model, improving transparency in large language models.

    Self-Exploring Language Models (SELM) is a promising approach to improving the alignment of LLMs to human intentions through online feedback collection.


    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:39 A content and product partnership with The Atlantic

    02:59 Mistral Releases Codestral, a Code-focused Model

    04:34 How Dell Is Beating Supermicro

    05:50 Fake sponsor

    08:06 MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series

    09:44 Self-Exploring Language Models: Active Preference Elicitation for Online Alignment

    11:16 Value-Incentivized Preference Optimization: A Unified Approach to Online and Offline RLHF

    13:18 Outro

    15 min
  • OpenAI's starts training GPT-5 🤖 // Jan Leike joins Anthropic's Superalignment Team 👥 // MoEUT Outperforms Standard Transformers 💥

    OpenAI has formed a new safety team to address concerns about AI safety and ethics, led by CEO Sam Altman and board members Adam D’Angelo and Nicole Seligman.

    Jan Leike, a leading AI researcher, has left OpenAI and joined Anthropic's Superalignment team, which is focused on AI safety and security.

    The latest version of Sentence Transformers v3 has been released, allowing for finetuning of models for specific tasks like semantic search and paraphrase mining.

    Exciting new research papers have been published, including MoEUT, a shared-layer Transformer design that outperforms standard Transformers on language modeling tasks, and EM Distillation, a new distillation method for diffusion models that efficiently distills them to one-step generator models without sacrificing perceptual quality.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:32 OpenAI has a new safety team — it’s run by Sam Altman

    03:18 Jan Leike (ex OpenAI) joins Anthropic's Superalignment Team

    05:04 Sentence Transformers v3 Released

    06:06 Fake sponsor

    08:19 MoEUT: Mixture-of-Experts Universal Transformers

    10:10 Greedy Growing Enables High-Resolution Pixel-Based Diffusion Models

    11:48 EM Distillation for One-step Diffusion Models

    13:42 Outro

    16 min
  • xAI Raises $6B 🚀 // Google's AI Overviews Controversy 🤔 // Transformers Master Arithmetic 🧮

    xAI, founded by Elon Musk, raises $6 billion in funding to accelerate the research and development of future technologies in the AI race.

    Google's new 'AI Overviews' search feature causes uproar with bizarre and inaccurate responses, potentially eroding trust in Google's search results.

    "Transformers Can Do Arithmetic with the Right Embeddings" proposes a solution to transformers' struggles with arithmetic tasks, achieving up to 99% accuracy on 100 digit addition problems.

    "SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering" introduces SWE-agent, an autonomous system that uses a language model to interact with a computer to solve software engineering tasks, with potential to revolutionize the field.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:27 Elon Musk’s xAI raises $6 billion to fund its race against ChatGPT and all the rest

    02:51 Google’s A.I. Search Errors Cause a Furor Online

    04:17 ir-measures Documentation

    05:15 Fake sponsor

    07:12 Transformers Can Do Arithmetic with the Right Embeddings

    08:23 Matryoshka Multimodal Models

    09:58 SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

    11:47 Outro

    13 min
  • OpenAI Drama 💥 // Synthetic Data Theorem Proving 🧪 // Dense Vision-Language Connector 🤝

    OpenAI drama: Leaked documents and a resignation from a policy researcher.

    DeepSeek-Prover: A new approach to formal theorem proving using synthetic data.

    Dense Connector for MLLMs: A plug-and-play vision-language connector that enhances existing models.

    Thermodynamic Natural Gradient Descent: A new algorithm for training neural networks using natural gradient descent.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:29 On OpenAI's Sky Voice

    03:04 Successful language model evals

    03:58 Generative Molecular Design Isn't As Easy As People Make It Look

    05:21 Fake sponsor

    07:30 DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data

    09:15 Dense Connector for MLLMs

    10:43 Thermodynamic Natural Gradient Descent

    12:37 Outro

    14 min
  • Cohere's Open-source Aya 🌎 // Anthropic Interpretability 🧠 // Video Editing AI 🎥

    Cohere's Aya model and dataset for multilingual AI in 101 languages through open science.

    "Mapping the Mind of a Large Language Model" paper by Anthropic Blog, providing a detailed look inside a modern, production-grade model.

    "ReVideo: Remake a Video with Motion and Content Control" paper introducing a new approach to video editing.

    "Dense Connector for MLLMs" paper introducing the Dense Connector, a plug-and-play vision-language connector that significantly enhances existing MLLMs.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:32 Cohere Launches Aya

    03:32 Mapping the Mind of a Large Language Model

    05:05 The Batch Newsletter

    06:08 Fake sponsor

    07:38 ReVideo: Remake a Video with Motion and Content Control

    09:14 Not All Language Model Features Are Linear

    10:55 Dense Connector for MLLMs

    12:30 Outro

    14 min

About GPT Reviews

From the publisher's feed

A daily show about AI made by AI: news, announcements, and research from arXiv, mixed in with some fun. Hosted by Giovani Pete Tizzano, an overly hyped AI enthusiast; Robert, an often unimpressed…