RSS

Seb Farquhar

Karma: 711

AGI Safety and Align­ment at Google Deep­Mind: A Sum­mary of Re­cent Work (July 2026)

31 Jul 2026 15:57 UTC
85 points
0 comments9 min readLW link
(gdmalignment.substack.com)

The AGI Safety and Align­ment team at Google Deep­Mind is Hiring (July 2026)

31 Jul 2026 15:53 UTC
67 points
2 comments6 min readLW link
(gdmalignment.substack.com)

GDM AI Con­trol Roadmap

18 Jun 2026 16:50 UTC
86 points
2 comments1 min readLW link

Test­ing Gem­ini mod­els for schem­ing tendencies

29 May 2026 19:24 UTC
47 points
8 comments6 min readLW link
(deepmindsafetyresearch.medium.com)

MONA: Man­aged My­opia with Ap­proval Feedback

23 Jan 2025 12:24 UTC
81 points
30 comments9 min readLW link

AGI Safety and Align­ment at Google Deep­Mind: A Sum­mary of Re­cent Work

20 Aug 2024 16:22 UTC
217 points
33 comments9 min readLW link

Dis­cus­sion: Challenges with Un­su­per­vised LLM Knowl­edge Discovery

18 Dec 2023 11:58 UTC
149 points
21 comments10 min readLW link