GPT Reviews

GPT Reviews

By EarkindNewsDaily News
Download on the App Store

GPT Reviews episodes

  • Apple's Generative AI 💰 // CoreWeave's $2.3B Loan 🌩️ // Evaluating Large Multimodal Models 🧐

    Apple's investment in generative AI and CoreWeave's $2.3 billion loan for cloud infrastructure. We also dive into two research papers, "Retroformer" and "MM-Vet," which explore optimizing language agents and evaluating large multimodal models, respectively.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:42 Apple has been quiet about ChatGPT. Now Tim Cook says its hefty $22.6 billion research spend is down to generative AI.

    02:58 CoreWeave, which provides cloud infrastructure for AI training, secures $2.3B loan

    05:17 Analysis: While AI takes the spotlight, infrastructure stocks shine

    06:56 Fake sponsor

    08:53 Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

    10:11 MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

    11:49 Training Data Protection with Compositional Diffusion Models

    13:42 Outro

    16 min
  • GPT-5 Trademark Application 📝 // Google's "Mind-Reading" AI 🧠 // Soft MoE Outperforms Transformers 🔥

    OpenAI's trademark application for GPT-5 could signify a continued advancement in natural language processing and machine learning. Google's "mind-reading" AI raises ethical concerns about potential implications and future uses of the technology. Soft MoE, a new type of mixture of expert architecture proposed by Google DeepMind, outperforms standard Transformers and popular MoE variants in visual recognition. PerceptionCLIP, a method proposed by the University of Maryland and the Bosch Center for Artificial Intelligence, improves zero-shot image classification by inferring and conditioning on contextual attributes, achieving better generalization, group robustness, and interpretability.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:23 OpenAI Files Trademark Application For GPT-5

    02:44 Google's 'mind-reading' AI can tell what music you listened to based on your brain signals

    04:26 Nvidia H100 GPUs: Supply and Demand

    05:17 Fake sponsor

    07:24 From Sparse to Soft Mixtures of Experts

    09:11 More Context, Less Distraction: Visual Classification by Inferring and Conditioning on Contextual Attributes

    10:47 Music De-limiter Networks via Sample-wise Gain Inversion

    12:45 Outro

    15 min
  • AI Video Summaries on YouTube 🎥 // Generative AI for Audio 🎵 // Synthetic Dataset for Point Tracking 🕹️

    YouTube is experimenting with AI-generated video summaries, potentially changing the way creators structure their videos. AudioCraft, a generative AI for audio, is now available to all and could revolutionize the way we create and experience audio. PointOdyssey, a synthetic dataset and data generation framework, aims to advance the state-of-the-art in long-term fine-grained tracking algorithms. A new attack method on aligned language models demonstrates the vulnerability of even aligned models to adversarial attacks, raising important questions about preventing objectionable content.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:33 YouTube uses AI to summarize videos in latest test

    02:58 Open sourcing AudioCraft: Generative AI for audio made simple and available to all

    04:52 Show HN: Learn a language quickly by practising speaking with AI (prettypolly.app)

    05:50 Fake sponsor

    07:54 PointOdyssey: A Large-Scale Synthetic Dataset for Long-Term Point Tracking

    09:35 Three Bricks to Consolidate Watermarks for Large Language Models

    11:14 Universal and Transferable Adversarial Attacks on Aligned Language Models

    13:39 Outro

    15 min
  • Palantir's AI Weapons 💣 // Meta's Chatbot Personas 🤖 // Google Assistant x BART 🗣️

    Billionaire CEO of military technology supplier Palantir advocates for AI weapons, sparking controversy and raising questions about the risks and benefits of such technology.

    Meta is developing AI-powered chatbots with different personalities, which could collect vast amounts of data on users' interests, but also raises concerns around privacy and potential manipulation.

    Google is planning to update Assistant with features powered by generative AI, which could allow it to answer questions based on information gleaned from across the web, but also raises potential privacy implications.

    Cutting-edge AI research papers have been published, including one on developing adaptable control policies for autonomous robots and another on improving the tool-use capabilities of open-source large language models. These papers have important implications for the future of AI technology.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:46 Billionaire CEO of military technology supplier Palantir advocates for AI weapons: 'We must not grow complacent'

    03:24 Meta prepares chatbots with personas to try to retain users

    04:32 Google will ‘supercharge’ Assistant with AI that’s more like ChatGPT and Bard

    06:13 Fake sponsor

    08:14 Discovering Adaptable Symbolic Algorithms from Scratch

    09:46 ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs

    11:43 MovieChat: From Dense Token to Sparse Memory for Long Video Understanding

    13:16 Outro

    15 min
  • StackOverflow's OverflowAI 💡 // Vision & Language to Action 🤖 // Limits of RLHF 🚫

    The RT-2 model translates vision and language into action, showing improved generalization capabilities and semantic and visual understanding beyond the robotic data it was exposed to.

    OverflowAI is a new space for Stack Overflow's community and customers to explore the future of knowledge sharing together, featuring semantic search, enterprise knowledge ingestion, Slack integration, a Visual Studio Code extension, and AI community discussions.

    "Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback" surveys the fundamental limitations and open problems of RLHF, as well as techniques to improve and complement it in practice.

    "Robust Distortion-free Watermarks for Language Models" proposes a way to plant watermarks in text generated by language models that are robust to perturbations without changing the distribution over text up to a certain maximum generation budget.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:43 RT-2: New model translates vision and language into action

    03:17 Announcing OverflowAI, the future of community & AI

    04:57 Computer Scientists Discover Limits of Stochastic Gradient Descent

    06:00 Fake sponsor

    07:55 Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback

    09:39 Uncertainty in Natural Language Generation: From Theory to Applications

    11:59 Robust Distortion-free Watermarks for Language Models

    13:37 Outro

    15 min
  • Netflix's AI job Ad 💼 // Google's impressive Q2 earnings 📈 // Boosting LLMs for Code 🤖

    Netflix's controversial AI job ad, Google's impressive Q2 earnings, and advancements in autonomous agents and large language models for code. The episode also features discussions on a realistic web environment for building autonomous agents and a new framework for boosting pre-trained Code LLMs.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:41 Netflix touts $900k AI jobs amid Hollywood strikes

    02:55 Google stock jumped 10% this week, fueled by cloud, ads and hope in A.I.

    04:44 Worldcoin: a solution in search of its problem

    05:52 Fake sponsor

    08:00 WebArena: A Realistic Web Environment for Building Autonomous Agents

    09:24 Evaluating the Moral Beliefs Encoded in LLMs

    11:02 PanGu-Coder2: Boosting Large Language Models for Code with Ranking Feedback

    12:37 Outro

    15 min
  • OpenAI x Google x Anthropic x Microsoft Partnership 🤝 // Stablilty's New Text-to-Image Model 🌅 // Factual Error Detection Framework 📚

    A new partnership promoting responsible AI, the release of a new text-to-image model with ethical safeguards, a framework for detecting factual errors in generated text, and a framework for high-quality object tracking in videos.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:32 A new partnership to promote responsible AI

    03:18 Stability AI releases its latest image-generating model, Stable Diffusion XL 1.0

    05:27 No One Wants To Talk To Your Chatbot

    06:28 Fake sponsor

    08:45 FacTool: Factuality Detection in Generative AI -- A Tool Augmented Framework for Multi-Task and Multi-Domain Scenarios

    10:09 Tracking Anything in High Quality

    11:59 Towards Generalist Biomedical AI

    13:59 Outro

    16 min
  • Japan's Supercomputer 🇯🇵 // GitHub x EU AI Law 🇪🇺 // ARB benchmark for LLMs 🧪

    Japan's Ministry of Economy, Trade, and Industry is investing heavily in AI development by building a new supercomputer to accelerate progress in AI and reduce Japan's dependence on foreign countries. GitHub and other companies are calling for more open-source support in EU AI law to promote a more open ecosystem for AI innovation. The ARB benchmark presents a more challenging test for large language models, and current models score well below 50% on more demanding tasks. "Predicting Code Coverage without Execution" proposes using Machine Learning to predict code coverage without actual execution, which could lower the cost of code coverage. State-of-the-art LLMs were able to predict code coverage with reasonable accuracy, demonstrating the potential of using Machine Learning for this task.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:40 Japan's METI to build new supercomputer to help develop AI at home

    03:28 GitHub and others call for more open-source support in EU AI law

    05:06 ChatGPT broke the Turing test — the race is on for new ways to assess AI

    06:26 Fake sponsor

    08:30 ARB: Advanced Reasoning Benchmark for Large Language Models

    09:55 Contrastive Example-Based Control

    11:16 Predicting Code Coverage without Execution

    13:14 Outro

    16 min
  • Worldcoin from Sam Altman 🪙 // ChatGPT for Android 📱 // Retentive Network for LLMs 🧠

    The launch of Worldcoin at the intersection of AI, identity, and finance, to the Android release of OpenAI's ChatGPT, this episode explores the cutting-edge of AI-powered conversational tools. The paper "Evaluating the Ripple Effects of Knowledge Editing in Language Models" proposes a new evaluation benchmark that could have important implications for improving the accuracy of language models. Finally, the Retentive Network, a new architecture proposed as a successor to the Transformer architecture for large language models, achieves favorable scaling results, parallel training, low-cost deployment, and efficient inference.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:45 OpenAI CEO Sam Altman Launches Worldcoin: A Bold Crypto Experiment At The Intersection Of AI, Identity And Finance

    03:13 ChatGPT for Android launches next week

    04:52 Michaël Benesty on Speeding up LLAMA v2 inference

    06:30 Fake sponsor

    08:19 Evaluating the Ripple Effects of Knowledge Editing in Language Models

    09:50 ZeroQuant-FP: A Leap Forward in LLMs Post-Training W4A8 Quantization Using Floating-Point Formats

    11:28 Retentive Network: A Successor to Transformer for Large Language Models

    13:15 Outro

    15 min
  • Stability's New LMs 🆕 // Anthropic on Faithfulnes in CoT 🔗 // STEVE-1 for Minecraft 🎮

    two new open-source Large Language Models, the departure of OpenAI's head of trust and safety, the introduction of STEVE-1, a model that can follow a wide range of instructions in Minecraft, and a paper exploring ways to protect the copyright of Neural Radiance Fields.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    02:03 Stability AI Releases 2 Language Models

    03:20 OpenAI’s head of trust and safety Dave Willner steps down

    04:59 Twitter Thread by Anthropic on Faithfulness in Chain of thought reasoning

    06:12 Fake sponsor

    08:19 Invalid Logic, Equivalent Gains: The Bizarreness of Reasoning in Language Model Prompting

    09:54 STEVE-1: A Generative Model for Text-to-Behavior in Minecraft

    11:36 CopyRNeRF: Protecting the CopyRight of Neural Radiance Fields

    13:01 Outro

    15 min

About GPT Reviews

From the publisher's feed

A daily show about AI made by AI: news, announcements, and research from arXiv, mixed in with some fun. Hosted by Giovani Pete Tizzano, an overly hyped AI enthusiast; Robert, an often unimpressed…