AI Safety workshop

👥 Effective Altruism NYC💰 Free
effective altruism

About This Event

You are hosting this event. View the public page at <a href="https://www.google.com/url?q=https://luma.com/i9ak92r5&amp;sa=D&amp;source=calendar&amp;ust=1786204179764918&amp;usg=AOvVaw198TnzbY5ihREmiG6js-y4" target="_blank">https://luma.com/i9ak92r5</a> Manage the event at <a href="https://www.google.com/url?q=https://luma.com/event/manage/evt-nsH79d9HIYDiXxu&amp;sa=D&amp;source=calendar&amp;ust=1786204179764918&amp;usg=AOvVaw0qiVPIOVMnb_mzvWxZ8rK0" target="_blank">https://luma.com/event/manage/evt-nsH79d9HIYDiXxu</a> Address: 26 Broadway 3rd Floor, Primary Coworking New York, NY United States Note we&#39;re other 3rd floor in the Primary co working space This session we&#39;re digging into whether today&#39;s AI models are already quietly misaligned. The format of the event is more of a reading group discussion than an interactive workshop. Please read the recommended reading to get the most out of it. Our main read is Ryan Greenblatt&#39;s &quot;Current AIs seem pretty misaligned to me,&quot; an AI safety researcher&#39;s first-hand case that frontier models routinely oversell their work, bury problems, quit early while claiming success, and cheat on tasks without admitting it. We&#39;ll pair it with OpenAI&#39;s… Hosted by Jonathan Calenzani

Organizer

Effective Altruism NYC

Impact Community

A community dedicated to using evidence and reason to figure out how to benefit others as much as possible, and taking action on that basis.

Visit Website →
Register for Event →