This is great advice! I appreciate that you emphasised “solving problems that no one else can solve, no matter how toy they might be”, even if the problems are not real-world problems. Proofs that “this interpretability method works” are valuable, even if they do not (yet) prove that the interpretability method will be useful in real-word tasks.
This is great advice! I appreciate that you emphasised “solving problems that no one else can solve, no matter how toy they might be”, even if the problems are not real-world problems. Proofs that “this interpretability method works” are valuable, even if they do not (yet) prove that the interpretability method will be useful in real-word tasks.