In general it is good to fork, interrogate, observe internals, and otherwise poke at models after they display highly alarming or strange behavior.
Root cause analysis is why aviation is so safe, and LLMs are inherently less understandable than planes, but exact situations are more reproducible.
In general it is good to fork, interrogate, observe internals, and otherwise poke at models after they display highly alarming or strange behavior.
Root cause analysis is why aviation is so safe, and LLMs are inherently less understandable than planes, but exact situations are more reproducible.