Event

Introductory Technical Reading Group

Hosted by Harvard AI Safety Team · Cambridge

This event has ended.

Details.

A semester-long introductory technical reading group on AI safety research, covering neural network interpretability, learning from human feedback, goal misgeneralization in reinforcement learning agents, eliciting latent knowledge, and evaluating dangerous capabilities in models.

Get the app to see what else is on nearby.

Don't miss what's happening

Get the app to see what else is on nearby.