AI Safety Contractor and Masters student in Physics at the University of Queensland.
I am interested in Quantum Computing, physical AI Safety guarantees and alignment techniques that will work beyond contemporary models.
Views expressed here are my own.
I think any theory of impact that relied on model architectures remaining static was likely doomed. I hope that in studying models without a residual stream, interpretability researchers find techniques that will generalise further.