I’m curious about how this is supposed to work in practice. Currently we are already in a world in which search algorithms, and indeed whole social networks, are openly pushed towards specific points of view. That lock-in is already happening / has already happened. Those people already create, use and publish their own LLM. MechaHitler already happened, and other slightly less obvious examples also exist.
OK, so now… what to do about it? And how does this benchmark help?
I’m curious about how this is supposed to work in practice.
Currently we are already in a world in which search algorithms, and indeed whole social networks, are openly pushed towards specific points of view. That lock-in is already happening / has already happened. Those people already create, use and publish their own LLM. MechaHitler already happened, and other slightly less obvious examples also exist.
OK, so now… what to do about it? And how does this benchmark help?