
Sign up to save your podcasts
Or


OpenAI plans to set up chip factories worth $100 billion to reduce reliance on existing chipmakers and tackle potential supply shortages.
The rabbit r1, which integrates Perplexity AI's technology to respond to user inquiries, has garnered substantial pre-order sales and offers a complimentary year of Perplexity Pro to early adopters.
"Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs" explores how code prompts can improve the performance of large language models on conditional reasoning tasks.
"R-Judge: Benchmarking Safety Risk Awareness for LLM Agents" introduces R-Judge, a benchmark that evaluates the proficiency of LLMs in judging safety risks given agent interaction records, revealing the importance of salient safety risk feedback.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:20 OpenAI plans to set up chip factories worth $100 billion: Report
02:47 The rabbit r1 will use Perplexity AI’s tech to answer your queries
04:24 LoRA From Scratch – Implement Low-Rank Adaptation for LLMs in PyTorch
05:12 Fake sponsor
07:11 Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs
08:32 RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture
10:45 R-Judge: Benchmarking Safety Risk Awareness for LLM Agents
12:56 Outro
Mark Zuckerberg's new goal of creating artificial general intelligence with Meta's AI research group and the FDA clearance granted for the first AI-powered medical device to detect all three common skin cancers are just some of the highlights. We also explore Self-Rewarding Language Models and Automatic Program Repair using Round-Trip Translation with Large Language Models, as well as Large Language Models as neurosymbolic reasoners for text-based games involving symbolic tasks.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:41 Mark Zuckerberg’s new goal is creating artificial general intelligence
03:27 FDA Clearance Granted for First AI-Powered Medical Device to Detect All Three Common Skin Cancers
05:37 The rise of AI as Magic
06:36 Fake sponsor
09:11 Self-Rewarding Language Models
10:46 A Novel Approach for Automatic Program Repair using Round-Trip Translation with Large Language Models
12:27 Large Language Models Are Neurosymbolic Reasoners
14:23 Outro
Samsung introduces Galaxy AI platform with five key features, including Live Translate and Note Assist.
Meta is reportedly spending billions of dollars on Nvidia AI chips for AGI research.
Beware of misleading GPU vs CPU benchmarks, as pointed out in a blog post.
Three new research papers explore Reinforced Fine-Tuning for reasoning, Asynchronous Local-SGD Training for Language Modeling, and Vision Mamba for efficient visual representation learning with bidirectional state space models.
Contact: [email protected]
Timestamps:
00:34 Introduction
02:08 Galaxy AI at the Samsumg Galaxy S24
03:39 Mark Zuckerberg indicates Meta is spending billions of dollars on Nvidia AI chips
05:36 Beware of misleading GPU vs CPU benchmarks
06:56 Fake sponsor
08:34 ReFT: Reasoning with Reinforced Fine-Tuning
10:24 Asynchronous Local-SGD Training for Language Modeling
12:18 Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model
14:15 Outro
The predictions of Bill Gates on how AI will transform our lives, and groundbreaking research in AI-generated audio, text-to-video creation, and quantum-based noise reduction. Additionally, the proposed evaluation metric for text-to-video models, T2VScore, integrates Text-Video Alignment and Video Quality criteria to provide a more accurate reflection of human perception.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:47 AI - artificial intelligence - at Davos 2024: Rolling coverage and what to know
03:26 Bill Gates explains how AI will change our lives in 5 years
05:12 RAG Using Unstructured Data & Role of Knowledge Graphs
06:08 Fake sponsor
07:49 Masked Audio Generation using a Single Non-Autoregressive Transformer
09:32 Towards A Better Metric for Text-to-Video Generation
11:11 Quantum Denoising Diffusion Models
12:57 Outro
OpenAI's plan to combat election misinformation, Stable Code 3B's promise to revolutionize coding, and the potential uses of Tiny Machine Learning are all discussed. Additionally, the paper on the Unreasonable Effectiveness of Easy Training Data for Hard Tasks challenges previous assumptions about language models. Overall, this episode provides valuable insights into the latest developments in AI and technology.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:34 Here’s OpenAI’s big plan to combat election misinformation
02:46 Stable Code 3B: Coding on the Edge
04:29 What TinyML is
06:00 Fake sponsor
07:57 Mind Your Format: Towards Consistent Evaluation of In-Context Learning Improvements
09:38 The Unreasonable Effectiveness of Easy Training Data for Hard Tasks
11:09 How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs
12:55 Outro
Nous Research has released their new flagship LLM, Nous-Hermes 2, which is the first model trained with RLHF and the first model to beat Mixtral Instruct in popular benchmarks.
Microsoft's Copilot Pro brings AI-powered Office features to consumers for $20 a month, including the ability to generate entire PowerPoint slide decks from a chatbot-like prompt and rephrase paragraphs in Word.
A blog post explores fine-tuning gpt-3.5-turbo to learn how to play "Connections", demonstrating the potential of fine-tuning language models for specific tasks.
Three research papers are discussed, including the effects of pretraining data curation on language models, a new benchmark for evaluating multimodal large language models on image-based wordplay puzzles, and the major shortcomings identified in these models.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:37 Nous Research Releases new flagship LLM
02:50 Microsoft’s new Copilot Pro brings AI-powered Office features to the rest of us
05:03 Fine-tuning gpt-3.5-turbo to learn to play "Connections"
06:08 Fake sponsor
08:02 AboutMe: Using Self-Descriptions in Webpages to Document the Effects of English Pretraining Data Filters
09:36 AboutMe: Using Self-Descriptions in Webpages to Document the Effects of English Pretraining Data Filters
11:19 REBUS: A Robust Evaluation Benchmark of Understanding Symbols
13:11 Outro
Apple's relocation request for their Siri team to the $100 million investment in 1X Technologies, listeners will learn about the latest developments in the AI industry. The TrustLLM study evaluates the trustworthiness of LLMs across six dimensions, while Intel Corporation proposes an efficient LLM inference solution.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:35 Apple asks its San Diego Siri quality control team to relocate to Texas
02:51 OpenAI-Backed Humanoid Maker Gets $100 Million in EQT-Led Round
04:28 Why autonomous trucking is harder than autonomous rideshare
05:39 Fake sponsor
07:26 TrustLLM: Trustworthiness in Large Language Models
09:34 Efficient LLM inference solution on Intel GPU
11:25 Transformers are Multi-State RNNs
12:56 Outro
OpenAI has introduced a new plan called ChatGPT Team, which allows smaller teams to use their latest AI models without needing to know how to code.
Microsoft has overtaken Apple as the largest US company thanks to their AI boost, which has been attributed to their investments in AI and machine learning.
"The Impact of Reasoning Step Length on Large Language Models" challenges the traditional view that transformers are conceptually different from recurrent neural networks and provides a potential solution to a major computational issue.
"Distilling Vision-Language Models on Millions of Videos" proposes a method to fine-tune a video-language model from a strong image-language baseline with synthesized instructional data and then use it to auto-label millions of videos to generate high-quality captions. This method could significantly improve the quality of video captioning and retrieval.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:48 OpenAI debuts ChatGPT subscription aimed at small teams
03:14 Microsoft overtakes Apple as largest U.S. company on AI boost
04:51 Neural Network Quantization & Number Formats From First Principles
06:02 Fake sponsor
08:16 The Impact of Reasoning Step Length on Large Language Models
10:06 Transformers are Multi-State RNNs
11:45 Distilling Vision-Language Models on Millions of Videos
13:34 Outro
OpenAI's GPT Store, new generative AI-powered experiences for Amazon's Alexa, and breakthroughs in video and language modeling with "MagicVideo-V2" and "Lightning Attention-2".
Contact: [email protected]
Timestamps:
00:34 Introduction
01:42 OpenAI’s custom GPT Store is now open for business
03:16 Amazon’s Alexa gets new generative AI-powered experiences
05:03 Remember Netflix’s $1m algorithm contest? Well, here’s why it didn’t use the winning entry.
06:34 Fake sponsor
08:10 MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation
09:40 Masked Audio Generation using a Single Non-Autoregressive Transformer
11:18 Lightning Attention-2: A Free Lunch for Handling Unlimited Sequence Lengths in Large Language Models
13:18 Outro
More beef on the lawsuit against OpenAI, Volkswagen's new smart chatbot for cars, and the latest developments in Python and language modeling. The papers discussed showcase the potential for new techniques like Mixtral of Experts, MoE-Mamba, and FlightLLM to improve language processing and unlock new possibilities for scaling.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:27 OpenAI Fights Back Against New York Times Lawsuit
02:47 Volkswagen brings AI chatbot ChatGPT into its cars, SUVs
04:28 Python 3.13 gets a JIT
05:35 Fake sponsor
07:22 Mixtral of Experts
08:50 MoE-Mamba: Efficient Selective State Space Models with Mixture of Experts
10:20 FlightLLM: Efficient Large Language Model Inference with a Complete Mapping Flow on FPGA
12:13 Outro
From the publisher's feed