Just recently came across the preprint of “Learning a Generative Meta-Model of LLM Activations” by Liu et al. (https://arxiv.org/pdf/2602.06964) Which seems to be a work on using generative techniques for steering a model persona. Interestingly it seems to be the most recent work supervised by Alec Radford. Couldn’t really find a public discussion on that paper and would be interested in peoples opinion that are more in the field.
Just recently came across the preprint of “Learning a Generative Meta-Model of LLM Activations” by Liu et al. (https://arxiv.org/pdf/2602.06964) Which seems to be a work on using generative techniques for steering a model persona. Interestingly it seems to be the most recent work supervised by Alec Radford. Couldn’t really find a public discussion on that paper and would be interested in peoples opinion that are more in the field.