Models & Agents
Real-time multimodal agents just gained a unified audio-visual model that watches, listens, and speaks without turn-taking.
What You Need to Know: ByteDance Seed introduced SeedRealtime, a native full-duplex LLM that fuses audio, video, and text in one architecture for continuous interaction. OpenAI paused work on its next model Astra after internal tests showed cyber capabilities strong enough to trigger safety reviews. ...
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=0Jh4J7YtMq4
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.