Generally haunting the Boston area. Primarily interested in the intersection of philosophy, ethics, and system dynamics.
Current focus is on the topology of evaluation, including white-box analysis of MLPs and other neural nets.
https://github.com/OperatorPhoenix/ATLAS (currently set to private, DM if you’re interested in access)
Open to work/research opportunities, particularly in alignment, safety, and/or mech interp.
DMs are always welcome.
As an outsider to all of this, from my perspective I don’t know why you keep letting him bait you into continuing this argument. It doesn’t appear productive and it’s just dragging everyone through the mud as everyone tries to get the last word in?