Discussion about this post

User's avatar
Alistair Gray's avatar

I’ve been working with LLMs for some time now. The one thing that will be interesting to manage from a safety perspective are the nuances surrounding persuasion. Now, some things are clear cut, such as around preventing harm. But others are tricky. For example I’ve noticed I tend to take its advice more and more from baking to health to building a house in Portugal!

David's avatar

Thanks for this post - it’s reassuring to hear of the progress being made on AI alignment. Beyond AI’s extraordinary promise, many remain fearful of the imminent societal disruption being discussed. Those building AI also have a special responsibility to anticipate and help mitigate the negative impacts their pursuit may cause as well.

35 more comments...

No posts

Ready for more?