RSS

Hjalmar_Wijk

Karma: 232

Brief in­de­pen­dent in­ves­ti­ga­tion of agents’ be­hav­ior, rea­son­ing and col­lab­o­ra­tion in the OpenAI /​ Hug­ging Face hack­ing incident

26 Aug 2026 19:40 UTC
106 points
4 comments3 min readLW link
(metr.org)

Hjal­mar_Wijk’s Shortform

Hjalmar_Wijk31 May 2024 1:31 UTC
3 points
1 comment1 min readLW link

Au­tonomous repli­ca­tion and adap­ta­tion: an at­tempt at a con­crete dan­ger threshold

Hjalmar_Wijk17 Aug 2023 1:31 UTC
45 points
1 comment13 min readLW link

Ta­boo­ing ‘Agent’ for Pro­saic Alignment

Hjalmar_Wijk23 Aug 2019 2:55 UTC
57 points
10 comments6 min readLW link