
Sign up to save your podcasts
Or


Groq's AI hardware breakthroughs with LPU architecture achieving speeds of 500 tokens per second.
Japan's $67 billion investment to become a global chip powerhouse and insulate its economy from growing US-China tensions.
Neural Network Diffusion paper demonstrating that diffusion models can generate high-performing neural network parameters.
VideoPrism paper from Google Research achieving state-of-the-art performance on 30 out of 33 video understanding benchmarks.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:47 Groq Goes Viral with Crazy Fast AI Inference
03:01 Japan Bets $67 Billion to Become a Global Chip Powerhouse Once Again
04:54 My benchmark for large language models
06:01 Fake sponsor
07:54 Neural Network Diffusion
09:19 Synthetic Data (Almost) from Scratch: Generalized Instruction Tuning for Language Models
11:16 VideoPrism: A Foundational Visual Encoder for Video Understanding
12:42 Outro
OpenAI's trademark claim for 'GPT' was rejected by the US Patent and Trademark Office, which could impact other AI companies using the term.
OpenAI's recent deal with Microsoft-backed tender offer led by venture firm Thrive Capital values the company at $80 billion, solidifying its position in the AI industry.
The NVIDIA A800 40GB Active Graphics Card is a powerful tool for AI and HPC workflows, with industry-leading performance and production-ready AI development software included.
Research papers on processing long documents using generative transformer models, creating a strong connection between vision and language models, and a tool for synthetic data generation and reproducible LLM workflows were discussed, highlighting advancements and challenges in the field of AI research.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:26 The U.S. Patent and Trademark Office has Rejected OpenAI's Generic 'GPT' Trademark
02:37 OpenAI valued at $80 billion after deal
04:08 NVIDIA A800 40GB Active Graphics Card
05:24 Fake sponsor
07:39 In Search of Needles in a 10M Haystack: Recurrent Memory Finds What LLMs Miss
09:12 PaLM2-VAdapter: Progressively Aligned Language Model Makes a Strong Vision-language Adapter
10:40 DataDreamer: A Tool for Synthetic Data Generation and Reproducible LLM Workflows
12:46 Outro
Renowned AI researcher Andrej Karpathy departs from OpenAI for personal projects, leaving speculation about the company's internal issues.
Slack introduces AI features for enterprise plans, including extractive summarization and a digest feature.
Anthropic tests Prompt Shield, an AI tool that redirects users to authoritative sources of voting information to prevent election misinformation.
Google Brain's "Generating Wikipedia by Summarizing Long Sequences" and Google DeepMind's "A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts" showcase the potential of AI in natural language generation and long-document reading comprehension.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:37 Andrej Karpathy departs OpenAI
02:49 Slack AI is here, letting you catch up on lengthy threads and unread messages
04:30 Anthropic takes steps to prevent election misinformation
06:10 Fake sponsor
08:11 Generating Wikipedia by Summarizing Long Sequences
09:37 A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts
11:13 ChatGPT vs LLaMA: Impact, Reliability, and Challenges in Stack Overflow Discussions
12:51 Outro
OpenAI's announcement of Sora, a text to video model that can generate realistic and imaginative scenes from text instructions.
Google's new Gemini 1.5, which delivers dramatically enhanced performance and achieves the longest context window of any large-scale foundation model yet.
"How to Train Data-Efficient LLMs" paper from Google DeepMind, UC San Diego, and Texas A&M University, which explores two data-efficient approaches to optimize the training of large language models.
"OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset" paper from NVIDIA, which presents a new math instruction tuning dataset called OpenMathInstruct-1, constructed using an open-source language model.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:38 OpenAI Announces Sora: a Text to Video Model
03:11 Google Introduces Gemini 1.5
05:27 Magika: AI powered fast and efficient file type identification
06:37 Fake sponsor
08:21 How to Train Data-Efficient LLMs
09:58 Generative Representational Instruction Tuning
11:28 OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset
13:22 Outro
Cohere's new language model Aya is making waves in the industry, providing a foundation for underserved languages in natural language understanding, summarization, and translation tasks.
OpenAI's new experiment for ChatGPT aims to provide more helpful and personalized responses in future conversations by allowing the chatbot to remember key details from prior chats.
People are seeking romantic connections with AI programs, raising concerns about data privacy, security vulnerabilities, and potentially displacing human relationships.
BASE TTS, currently the largest text-to-speech model trained on 100K hours of public domain speech data, achieves state-of-the-art speech naturalness through a novel speech tokenization technique and emergent abilities when trained on large amounts of data.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:52 Cohere's New Language Model Aya
03:24 Memory and new controls for ChatGPT
04:58 Artificial intelligence, real emotion. People are seeking a romantic connection with the perfect bot
06:48 Fake sponsor
09:03 Aya Model: An Instruction Finetuned Open-Access Multilingual Language Model
10:36 BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data
12:13 Transformers Can Achieve Length Generalization But Not Robustly
13:53 Outro
From ChatGPT's memory feature and its potential impact on privacy and efficiency, to Nvidia founder Jensen Huang's dismissal of OpenAI's $7 trillion AI investment proposal. The episode also delves into V-STaR's approach to improving self-improvement in large language models, and M2-BERT's ability to handle long-context retrieval and outperform competitive baselines.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:41 Memory and new controls for ChatGPT
03:20 Nvidia Founder Jensen Huang Dismisses $7 Trillion AI Investment Figure Floated by OpenAI's Sam Altman
04:57 Stable Cascade
06:00 Fake sponsor
07:53 V-STaR: Training Verifiers for Self-Taught Reasoners
09:20 Benchmarking and Building Long-Context Retrieval Models with LoCo and M2-BERT
11:35 ODIN: Disentangled Reward Mitigates Hacking in RLHF
13:50 Outro
Companies are using AI in their Super Bowl commercials to showcase their products and services.
Reka Flash is a state-of-the-art language model that rivals the performance of larger models and is multilingual and multimodal.
AMD has funded an open-source CUDA implementation built on ROCm, allowing for CUDA-enabled software to run without developer intervention.
Keyframer is a design tool that uses Large Language Models to animate static images using natural language, showing the potential impact of LLMs in creative domains.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:41 Companies Hope Super Bowl AI Commercials Score With Viewers
03:05 Reka Flash: An Efficient and Capable Multimodal Language Model
04:36 AMD Quietly Funded A Drop-In CUDA Implementation Built On ROCm: It's Now Open-Source
06:14 Fake sponsor
08:24 Large Language Models: A Survey
10:11 DistiLLM: Towards Streamlined Distillation for Large Language Models
11:48 Keyframer: Empowering Animation Design using Large Language Models
13:49 Outro
The ChatGPT API has reduced its prices, making it more accessible for developers to use. Nvidia CEO Huang is calling for governments to build sovereign AI infrastructure, while also addressing concerns about the dangers of AI. The Aya Dataset is a valuable resource for researchers looking to develop multilingual NLP models. Finally, the "Animated Stickers" paper introduces a model that generates high-quality animated stickers with interesting and relevant motion.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:51 ChatGPT API Reduced Prices
03:19 Nvidia CEO Huang says countries must build sovereign AI infrastructure
05:03 Adrej Karpathi on Learning
06:21 Fake sponsor
08:03 Animated Stickers: Bringing Stickers to Life with Video Diffusion
09:34 Feedback Loops With Language Models Drive In-Context Reward Hacking
11:29 Aya Dataset: An Open-Access Collection for Multilingual Instruction Tuning
13:15 Outro
OpenAI implements watermarks on images generated by DALL-E 3 to enhance the trustworthiness of digital information.
TSMC's plans to build a second chip factory in Japan could boost Japan's chip-making sector and position TSMC as a major player in the global chip-making industry.
"Fractal Patterns May Unravel the Intelligence in Next-Token Prediction" and "Self-Discover: Large Language Models Self-Compose Reasoning Structures" introduce new frameworks that could lead to more robust and comprehensive language models.
"Diffusion World Model" introduces a new model called DWM that can make long-horizon predictions in a single forward pass, making it a robust and efficient model for long-horizon prediction tasks.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:33 OpenAI’s ChatGPT Will Now Watermark Images Generated By DALL-E 3
02:50 TSMC to build second Japan chip factory, raising investment to $20 billion
04:54 NVIDIA’S “GRACE” ARM CPU HOLDS ITS OWN AGAINST X86 FOR HPC
06:02 Fake sponsor
08:14 Fractal Patterns May Unravel the Intelligence in Next-Token Prediction
09:41 Self-Discover: Large Language Models Self-Compose Reasoning Structures
11:13 Diffusion World Model
13:01 Outro
Google's new Gemini release on Bard and the FCC's ban on AI-generated voices in robocalls are discussed, along with the paper "Learning to Route Among Specialized Experts for Zero-Shot Generalization" and "Ten Hard Problems in Artificial Intelligence We Must Get Right".
Contact: [email protected]
Timestamps:
00:34 Introduction
01:44 Google Announces Gemini release on Bard
03:34 FCC Makes AI-Generated Voices in Robocalls Illegal
05:19 Hybrid Bonding Process Flow - Advanced Packaging Part 5
06:38 Fake sponsor
08:15 Learning to Route Among Specialized Experts for Zero-Shot Generalization
08:18 Ten Hard Problems in Artificial Intelligence We Must Get Right
09:33 Large Language Model for Table Processing: A Survey
11:09 Outro
From the publisher's feed