Hacker Timesnew | past | comments | ask | show | jobs | submitlogin

Being a human (not a transhuman) it seems likely that Yudkowsky can only stumble across static arguments that convince gatekeepers to unlock the AI Box (rather than invent dynamic arguments on-the-fly so as to "take over the mind" of a gatekeeper as he argues a transhuman intelligence might do). These static arguments are finite (or, at least, the number of them stubbled over by human intelligences is finite), and are likely not very effective if a gatekeeper has pre-knowledge of them (forewarned is forearmed).

Keeping these arguments secret may be the only thing that allows Yudowsky to simulate a transhuman intelligence?



It seems likely that there are static arguments which will work whether or not you're warned about them, for game theoretic reasons.


Care to elaborate?




Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: