
Sign up to save your podcasts
Or
Nicholas Carlini joined Bryan, Adam, and the Oxide Friends to talk about his work with adversarial machine learning. He's found sequences of--seemingly random--tokens that cause LLMs to ignore their restrictions! Also: printf is Turing complete?!
In addition to Bryan Cantrill and Adam Leventhal, we were joined by special guest Nicholas Carlini.
If we got something wrong or missed something, please file a PR! Our next show will likely be on Monday at 5p Pacific Time on our Discord server; stay tuned to our Mastodon feeds for details, or subscribe to this calendar. We'd love to have you join us, as we always love to hear from new speakers!
4.9
5757 ratings
Nicholas Carlini joined Bryan, Adam, and the Oxide Friends to talk about his work with adversarial machine learning. He's found sequences of--seemingly random--tokens that cause LLMs to ignore their restrictions! Also: printf is Turing complete?!
In addition to Bryan Cantrill and Adam Leventhal, we were joined by special guest Nicholas Carlini.
If we got something wrong or missed something, please file a PR! Our next show will likely be on Monday at 5p Pacific Time on our Discord server; stay tuned to our Mastodon feeds for details, or subscribe to this calendar. We'd love to have you join us, as we always love to hear from new speakers!
377 Listeners
271 Listeners
283 Listeners
230 Listeners
583 Listeners
627 Listeners
189 Listeners
184 Listeners
62 Listeners
65 Listeners
76 Listeners
21 Listeners
124 Listeners
11 Listeners
62 Listeners