
Sign up to save your podcasts
Or


AI used to predict potential new antibiotics in groundbreaking study.
Stable Audio Open: an open source model that allows users to create short audio samples and sound effects from text prompts.
The ethical responsibilities of AI researchers when it comes to warning about the dangers of advanced artificial intelligence.
Cutting-edge research on AI and robotics, including large-scale simulations, in-context learning, and skill composition in modular arithmetic tasks.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:21 AI used to predict potential new antibiotics in groundbreaking study
02:40 Introducing Stable Audio Open - An Open Source Model for Audio Samples and Sound Design
04:22 A Right to Warn about Advanced Artificial Intelligence
05:23 Fake sponsor
07:10 RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots
08:56 Learning to grok: Emergence of in-context learning and skill composition in modular arithmetic tasks
10:36 Guiding a Diffusion Model with a Bad Version of Itself
12:24 Outro
Amazon's new AI system to detect damaged or incorrect items before they ship.
Elon Musk's controversial decision to prioritize X and xAI over Tesla for AI chips.
"To Believe or Not to Believe Your LLM" paper on uncertainty quantification in Large Language Models.
"Guiding a Diffusion Model with a Bad Version of Itself" paper on improving image generation with diffusion models.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:47 Learn how Amazon uses AI to spot damaged products before they’re shipped to customers
03:17 Elon Musk ordered Nvidia to ship thousands of AI chips reserved for Tesla to X and xAI
05:08 FineWeb: decanting the web for the finest text data at scale
06:16 Fake sponsor
08:05 To Believe or Not to Believe Your LLM
09:33 Guiding a Diffusion Model with a Bad Version of Itself
11:06 Learning to grok: Emergence of in-context learning and skill composition in modular arithmetic tasks
12:54 Outro
Microsoft is investing $3.2 billion in Sweden for cloud and AI infrastructure, deploying 20,000 advanced graphics processing units and training 250,000 Swedes with AI skills over three years.
"Grokfast" is a new algorithm that accelerates generalization under the grokking phenomenon in machine learning by amplifying the slow-varying component of gradients, improving performance on tasks like image classification.
"Zipper" is a multi-tower decoder architecture that uses cross-attention to flexibly compose multimodal generative models from independently pre-trained unimodal decoders, showcasing superior performance in tasks like speech-to-text generation.
"MetRag" is a new framework for retrieval augmented generation that combines similarity and utility-oriented models, using an LLM as a task adaptive summarizer to generate knowledge-augmented text and outperforming existing models on knowledge-intensive tasks like finance and medicine.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:49 Microsoft to invest $3.2 bln in Swedish cloud, AI
03:42 State Space Duality (Mamba-2) Part I - The Model
04:47 Sam Altman, Lately
06:08 Fake sponsor
08:39 Grokfast: Accelerated Grokking by Amplifying Slow Gradients
10:11 Zipper: A Multi-Tower Decoder Architecture for Fusing Modalities
11:38 Similarity is Not All You Need: Endowing Retrieval Augmented Generation with Multi Layered Thoughts
13:52 Outro
Nvidia unveils plans to accelerate the advance of artificial intelligence, partnering with companies and countries to build AI factories and releasing Nvidia ACE generative AI.
Finnish startup Binit develops an AI gadget that tracks household waste to encourage recycling, with potential benefits in improving recycling efficiency.
"The Intellectual Obesity Crisis" article discusses how we've become addicted to useless information, just like we evolved to crave sugar because it was a scarce source of energy.
Three AI research papers are discussed, including a method to compress second-order optimizer states to lower bitwidths, the first-ever full-spectrum, multi-modal evaluation benchmark of MLLMs in video analysis, and a theoretical connection between Transformers and state-space models leading to a faster and more efficient alternative to existing models.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:21 AI hardware firm Nvidia unveils next-gen products at Taiwan tech expo
02:48 Binit is bringing AI to trash
04:39 The Intellectual Obesity Crisis
06:22 Fake sponsor
08:18 4-bit Shampoo for Memory-Efficient Network Training
10:01 Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
11:51 Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality
13:30 Outro
Google's AI Overviews are improving to provide accurate and helpful information.
Nvidia's new embedding model, NV-Embed-v1, ranks number one on the Massive Text Embedding Benchmark.
Matryoshka Query Transformer (MQT) offers flexibility to Large Vision-Language Models (LVLMs) by encoding an image into a variable number of visual tokens during inference.
Contextual Position Encoding (CoPE) improves the position encoding method in Large Language Models (LLMs) and solves tasks where popular position embeddings fail.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:35 AI Overviews: About last week
03:58 Nvidia Releases Embedding Model NV-Embed-v1
04:53 Multi-camera YOLOv5 on Zynq UltraScale+ with Hailo-8 AI Acceleration
06:31 Fake sponsor
08:28 Matryoshka Query Transformer for Large Vision-Language Models
10:24 Similarity is Not All You Need: Endowing Retrieval Augmented Generation with Multi Layered Thoughts
11:51 Contextual Position Encoding: Learning to Count What's Important
13:30 Outro
OpenAI announces new content and product partnerships with Vox Media and The Atlantic, making their reporting and stories more discoverable to millions of OpenAI users.
Mistral AI releases Codestral, a 22B parameter, open-weight model that specializes in coding tasks, beating out its code-focused rivals across top benchmarks.
MAP-Neo is the first fully open-sourced bilingual LLM that provides all the details needed to reproduce the model, improving transparency in large language models.
Self-Exploring Language Models (SELM) is a promising approach to improving the alignment of LLMs to human intentions through online feedback collection.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:39 A content and product partnership with The Atlantic
02:59 Mistral Releases Codestral, a Code-focused Model
04:34 How Dell Is Beating Supermicro
05:50 Fake sponsor
08:06 MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series
09:44 Self-Exploring Language Models: Active Preference Elicitation for Online Alignment
11:16 Value-Incentivized Preference Optimization: A Unified Approach to Online and Offline RLHF
13:18 Outro
OpenAI has formed a new safety team to address concerns about AI safety and ethics, led by CEO Sam Altman and board members Adam D’Angelo and Nicole Seligman.
Jan Leike, a leading AI researcher, has left OpenAI and joined Anthropic's Superalignment team, which is focused on AI safety and security.
The latest version of Sentence Transformers v3 has been released, allowing for finetuning of models for specific tasks like semantic search and paraphrase mining.
Exciting new research papers have been published, including MoEUT, a shared-layer Transformer design that outperforms standard Transformers on language modeling tasks, and EM Distillation, a new distillation method for diffusion models that efficiently distills them to one-step generator models without sacrificing perceptual quality.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:32 OpenAI has a new safety team — it’s run by Sam Altman
03:18 Jan Leike (ex OpenAI) joins Anthropic's Superalignment Team
05:04 Sentence Transformers v3 Released
06:06 Fake sponsor
08:19 MoEUT: Mixture-of-Experts Universal Transformers
10:10 Greedy Growing Enables High-Resolution Pixel-Based Diffusion Models
11:48 EM Distillation for One-step Diffusion Models
13:42 Outro
xAI, founded by Elon Musk, raises $6 billion in funding to accelerate the research and development of future technologies in the AI race.
Google's new 'AI Overviews' search feature causes uproar with bizarre and inaccurate responses, potentially eroding trust in Google's search results.
"Transformers Can Do Arithmetic with the Right Embeddings" proposes a solution to transformers' struggles with arithmetic tasks, achieving up to 99% accuracy on 100 digit addition problems.
"SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering" introduces SWE-agent, an autonomous system that uses a language model to interact with a computer to solve software engineering tasks, with potential to revolutionize the field.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:27 Elon Musk’s xAI raises $6 billion to fund its race against ChatGPT and all the rest
02:51 Google’s A.I. Search Errors Cause a Furor Online
04:17 ir-measures Documentation
05:15 Fake sponsor
07:12 Transformers Can Do Arithmetic with the Right Embeddings
08:23 Matryoshka Multimodal Models
09:58 SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
11:47 Outro
OpenAI drama: Leaked documents and a resignation from a policy researcher.
DeepSeek-Prover: A new approach to formal theorem proving using synthetic data.
Dense Connector for MLLMs: A plug-and-play vision-language connector that enhances existing models.
Thermodynamic Natural Gradient Descent: A new algorithm for training neural networks using natural gradient descent.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:29 On OpenAI's Sky Voice
03:04 Successful language model evals
03:58 Generative Molecular Design Isn't As Easy As People Make It Look
05:21 Fake sponsor
07:30 DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
09:15 Dense Connector for MLLMs
10:43 Thermodynamic Natural Gradient Descent
12:37 Outro
Cohere's Aya model and dataset for multilingual AI in 101 languages through open science.
"Mapping the Mind of a Large Language Model" paper by Anthropic Blog, providing a detailed look inside a modern, production-grade model.
"ReVideo: Remake a Video with Motion and Content Control" paper introducing a new approach to video editing.
"Dense Connector for MLLMs" paper introducing the Dense Connector, a plug-and-play vision-language connector that significantly enhances existing MLLMs.
Contact: [email protected]
Timestamps:
00:34 Introduction
01:32 Cohere Launches Aya
03:32 Mapping the Mind of a Large Language Model
05:05 The Batch Newsletter
06:08 Fake sponsor
07:38 ReVideo: Remake a Video with Motion and Content Control
09:14 Not All Language Model Features Are Linear
10:55 Dense Connector for MLLMs
12:30 Outro
From the publisher's feed