GPT Reviews

GPT Reviews

By EarkindNewsDaily News
Download on the App Store

GPT Reviews episodes

  • Groq's AI Hardware 💻 // Japan's $67B Chip Bet 🎲 // Video Understanding 📹

    Groq's AI hardware breakthroughs with LPU architecture achieving speeds of 500 tokens per second.

    Japan's $67 billion investment to become a global chip powerhouse and insulate its economy from growing US-China tensions.

    Neural Network Diffusion paper demonstrating that diffusion models can generate high-performing neural network parameters.

    VideoPrism paper from Google Research achieving state-of-the-art performance on 30 out of 33 video understanding benchmarks.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:47 Groq Goes Viral with Crazy Fast AI Inference

    03:01 Japan Bets $67 Billion to Become a Global Chip Powerhouse Once Again

    04:54 My benchmark for large language models

    06:01 Fake sponsor

    07:54 Neural Network Diffusion

    09:19 Synthetic Data (Almost) from Scratch: Generalized Instruction Tuning for Language Models

    11:16 VideoPrism: A Foundational Visual Encoder for Video Understanding

    12:42 Outro

    15 min
  • OpenAI's Challenge 🤝 // NVIDIA's Graphics Card 💻 // Advancements in AI Research 🔬

    OpenAI's trademark claim for 'GPT' was rejected by the US Patent and Trademark Office, which could impact other AI companies using the term.

    OpenAI's recent deal with Microsoft-backed tender offer led by venture firm Thrive Capital values the company at $80 billion, solidifying its position in the AI industry.

    The NVIDIA A800 40GB Active Graphics Card is a powerful tool for AI and HPC workflows, with industry-leading performance and production-ready AI development software included.

    Research papers on processing long documents using generative transformer models, creating a strong connection between vision and language models, and a tool for synthetic data generation and reproducible LLM workflows were discussed, highlighting advancements and challenges in the field of AI research.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:26 The U.S. Patent and Trademark Office has Rejected OpenAI's Generic 'GPT' Trademark

    02:37 OpenAI valued at $80 billion after deal

    04:08 NVIDIA A800 40GB Active Graphics Card

    05:24 Fake sponsor

    07:39 In Search of Needles in a 10M Haystack: Recurrent Memory Finds What LLMs Miss

    09:12 PaLM2-VAdapter: Progressively Aligned Language Model Makes a Strong Vision-language Adapter

    10:40 DataDreamer: A Tool for Synthetic Data Generation and Reproducible LLM Workflows

    12:46 Outro

    14 min
  • Karpathy Leaves OpenAI 💥 // Slack's New AI Features 🤖 // Preventing Election Misinformation 🗳️

    Renowned AI researcher Andrej Karpathy departs from OpenAI for personal projects, leaving speculation about the company's internal issues. 

    Slack introduces AI features for enterprise plans, including extractive summarization and a digest feature. 

    Anthropic tests Prompt Shield, an AI tool that redirects users to authoritative sources of voting information to prevent election misinformation. 

    Google Brain's "Generating Wikipedia by Summarizing Long Sequences" and Google DeepMind's "A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts" showcase the potential of AI in natural language generation and long-document reading comprehension. 

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:37 Andrej Karpathy departs OpenAI

    02:49 Slack AI is here, letting you catch up on lengthy threads and unread messages

    04:30 Anthropic takes steps to prevent election misinformation

    06:10 Fake sponsor

    08:11 Generating Wikipedia by Summarizing Long Sequences

    09:37 A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts

    11:13 ChatGPT vs LLaMA: Impact, Reliability, and Challenges in Stack Overflow Discussions

    12:51 Outro

    15 min
  • OpenAI's Sora: Text-to-Video 📹 // Google's Gemini 1.5 🚀 // Data-efficient LLMs 💾

    OpenAI's announcement of Sora, a text to video model that can generate realistic and imaginative scenes from text instructions.

    Google's new Gemini 1.5, which delivers dramatically enhanced performance and achieves the longest context window of any large-scale foundation model yet.

    "How to Train Data-Efficient LLMs" paper from Google DeepMind, UC San Diego, and Texas A&M University, which explores two data-efficient approaches to optimize the training of large language models.

    "OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset" paper from NVIDIA, which presents a new math instruction tuning dataset called OpenMathInstruct-1, constructed using an open-source language model.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:38 OpenAI Announces Sora: a Text to Video Model

    03:11 Google Introduces Gemini 1.5

    05:27 Magika: AI powered fast and efficient file type identification

    06:37 Fake sponsor

    08:21 How to Train Data-Efficient LLMs

    09:58 Generative Representational Instruction Tuning

    11:28 OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset

    13:22 Outro

    15 min
  • Cohere's Aya Languag Model 🌍 // Personalized ChatGPT 🤖 // AI Romance 😍

    Cohere's new language model Aya is making waves in the industry, providing a foundation for underserved languages in natural language understanding, summarization, and translation tasks.

    OpenAI's new experiment for ChatGPT aims to provide more helpful and personalized responses in future conversations by allowing the chatbot to remember key details from prior chats.

    People are seeking romantic connections with AI programs, raising concerns about data privacy, security vulnerabilities, and potentially displacing human relationships.

    BASE TTS, currently the largest text-to-speech model trained on 100K hours of public domain speech data, achieves state-of-the-art speech naturalness through a novel speech tokenization technique and emergent abilities when trained on large amounts of data.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:52 Cohere's New Language Model Aya

    03:24 Memory and new controls for ChatGPT

    04:58 Artificial intelligence, real emotion. People are seeking a romantic connection with the perfect bot

    06:48 Fake sponsor

    09:03 Aya Model: An Instruction Finetuned Open-Access Multilingual Language Model

    10:36 BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data

    12:13 Transformers Can Achieve Length Generalization But Not Robustly

    13:53 Outro

    16 min
  • ChatGPT's Memory Feature 🤔 // Nvidia Founder Dismisses AI Investment Proposal 💸 // M2-BERT for Long-Context Retrieval 📈

    From ChatGPT's memory feature and its potential impact on privacy and efficiency, to Nvidia founder Jensen Huang's dismissal of OpenAI's $7 trillion AI investment proposal. The episode also delves into V-STaR's approach to improving self-improvement in large language models, and M2-BERT's ability to handle long-context retrieval and outperform competitive baselines.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:41 Memory and new controls for ChatGPT

    03:20 Nvidia Founder Jensen Huang Dismisses $7 Trillion AI Investment Figure Floated by OpenAI's Sam Altman

    04:57 Stable Cascade

    06:00 Fake sponsor

    07:53 V-STaR: Training Verifiers for Self-Taught Reasoners

    09:20 Benchmarking and Building Long-Context Retrieval Models with LoCo and M2-BERT

    11:35 ODIN: Disentangled Reward Mitigates Hacking in RLHF

    13:50 Outro

    16 min
  • Super Bowl AI Commercials 🏈 // Reka Flash Language Model 🤖 // AMD's Open-Source CUDA 🎮

    Companies are using AI in their Super Bowl commercials to showcase their products and services.

    Reka Flash is a state-of-the-art language model that rivals the performance of larger models and is multilingual and multimodal.

    AMD has funded an open-source CUDA implementation built on ROCm, allowing for CUDA-enabled software to run without developer intervention.

    Keyframer is a design tool that uses Large Language Models to animate static images using natural language, showing the potential impact of LLMs in creative domains.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:41 Companies Hope Super Bowl AI Commercials Score With Viewers

    03:05 Reka Flash: An Efficient and Capable Multimodal Language Model

    04:36 AMD Quietly Funded A Drop-In CUDA Implementation Built On ROCm: It's Now Open-Source

    06:14 Fake sponsor

    08:24 Large Language Models: A Survey

    10:11 DistiLLM: Towards Streamlined Distillation for Large Language Models

    11:48 Keyframer: Empowering Animation Design using Large Language Models

    13:49 Outro

    15 min
  • ChatGPT API Price Cut 💰 // Nvidia CEO's Sovereign AI Call 🌐 // Animated Stickers 🎉

    The ChatGPT API has reduced its prices, making it more accessible for developers to use. Nvidia CEO Huang is calling for governments to build sovereign AI infrastructure, while also addressing concerns about the dangers of AI. The Aya Dataset is a valuable resource for researchers looking to develop multilingual NLP models. Finally, the "Animated Stickers" paper introduces a model that generates high-quality animated stickers with interesting and relevant motion.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:51 ChatGPT API Reduced Prices

    03:19 Nvidia CEO Huang says countries must build sovereign AI infrastructure

    05:03 Adrej Karpathi on Learning

    06:21 Fake sponsor

    08:03 Animated Stickers: Bringing Stickers to Life with Video Diffusion

    09:34 Feedback Loops With Language Models Drive In-Context Reward Hacking

    11:29 Aya Dataset: An Open-Access Collection for Multilingual Instruction Tuning

    13:15 Outro

    15 min
  • Watermarks for DALLE 3 🌊 // TSMC's New Chip Factory in Japan 🇯🇵 // DeepMind's Self-Discover 🧩

    OpenAI implements watermarks on images generated by DALL-E 3 to enhance the trustworthiness of digital information.

    TSMC's plans to build a second chip factory in Japan could boost Japan's chip-making sector and position TSMC as a major player in the global chip-making industry.

    "Fractal Patterns May Unravel the Intelligence in Next-Token Prediction" and "Self-Discover: Large Language Models Self-Compose Reasoning Structures" introduce new frameworks that could lead to more robust and comprehensive language models.

    "Diffusion World Model" introduces a new model called DWM that can make long-horizon predictions in a single forward pass, making it a robust and efficient model for long-horizon prediction tasks.

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:33 OpenAI’s ChatGPT Will Now Watermark Images Generated By DALL-E 3

    02:50 TSMC to build second Japan chip factory, raising investment to $20 billion

    04:54 NVIDIA’S “GRACE” ARM CPU HOLDS ITS OWN AGAINST X86 FOR HPC

    06:02 Fake sponsor

    08:14 Fractal Patterns May Unravel the Intelligence in Next-Token Prediction

    09:41 Self-Discover: Large Language Models Self-Compose Reasoning Structures

    11:13 Diffusion World Model

    13:01 Outro

    15 min
  • Gemini Takes Over 🚀 // FCC Ban on AI Voices 🚫 // Zero-Shot Generalization 🎯

    Google's new Gemini release on Bard and the FCC's ban on AI-generated voices in robocalls are discussed, along with the paper "Learning to Route Among Specialized Experts for Zero-Shot Generalization" and "Ten Hard Problems in Artificial Intelligence We Must Get Right".

    Contact:  [email protected]

    Timestamps:

    00:34 Introduction

    01:44 Google Announces Gemini release on Bard

    03:34 FCC Makes AI-Generated Voices in Robocalls Illegal

    05:19 Hybrid Bonding Process Flow - Advanced Packaging Part 5

    06:38 Fake sponsor

    08:15 Learning to Route Among Specialized Experts for Zero-Shot Generalization

    08:18 Ten Hard Problems in Artificial Intelligence We Must Get Right

    09:33 Large Language Model for Table Processing: A Survey

    11:09 Outro

    13 min

About GPT Reviews

From the publisher's feed

A daily show about AI made by AI: news, announcements, and research from arXiv, mixed in with some fun. Hosted by Giovani Pete Tizzano, an overly hyped AI enthusiast; Robert, an often unimpressed…