RSS

Jennifer Lin

Karma: 410

Currently working on interpretability at Principles of Intelligence (PIBBSS). Previously at FHI and before that, in theoretical physics. I used to go by jylin04 on this website.

In­tro­duc­ing PIRAMID: Physics-In­formed Re­search for Am­bi­tious Mechanis­tic Interpretability

25 Jul 2026 15:54 UTC
60 points
0 comments6 min readLW link

Find­ing Fea­tures in Neu­ral Net­works with the Em­piri­cal NTK

Jennifer Lin16 Oct 2025 18:04 UTC
38 points
1 comment5 min readLW link

Can LLM-based mod­els do model-based plan­ning?

Jennifer Lin16 Apr 2025 12:38 UTC
11 points
1 comment2 min readLW link
(docs.google.com)

Othel­loGPT learned a bag of heuristics

2 Jul 2024 9:12 UTC
111 points
10 comments9 min readLW link

More Re­cent Progress in the The­ory of Neu­ral Networks

Jennifer Lin6 Oct 2022 16:57 UTC
82 points
6 comments4 min readLW link

A re­view of the Bio-An­chors report

Jennifer Lin3 Oct 2022 10:27 UTC
45 points
4 comments1 min readLW link
(docs.google.com)

Thoughts on AGI safety from the top

Jennifer Lin2 Feb 2022 20:06 UTC
37 points
3 comments32 min readLW link

Trans­parency and AGI safety

Jennifer Lin11 Jan 2021 18:51 UTC
54 points
12 comments30 min readLW link