RSS

Shahriar Golchin

Karma: 14

I am an AI Researcher at Labelbox, working on model evaluation and AI safety.

I earned my PhD in Computer Science (AI/​ML) from the University of Arizona. My research interests center on Large Language Models (LLMs), focusing on designing datasets/​tasks that adversarially stress-test alignment. I am particularly interested in surfacing systematic misalignment and reasoning failure modes across frontier AI models.

My PhD dissertation is the first to systematically identify data contamination (data leakage) in LLMs, scenarios where training data overlaps with evaluation data. I developed several methods to detect and estimate contamination in fully black-box LLMs.

Previously, I was a research intern at Google Cloud AI Research, Walmart Global Tech, and Harvard Medical School.

Do AI Models Want to Be Mon­i­tored? Mea­sur­ing Mon­i­tora­bil­ity Dis­po­si­tion in Large Rea­son­ing Models

Shahriar Golchin28 Aug 2026 7:32 UTC
12 points
0 comments7 min readLW link

The AI Safety Illu­sion: Why Cur­rent Safety Datasets Fool Us on Model Safety

Shahriar Golchin20 Jul 2026 20:12 UTC
5 points
0 comments8 min readLW link