The report evaluates Meta's LLM, Muse Spark, across safety and behavioral dimensions; I contributed to the sycophancy benchmark, which measures inappropriate agreement and pushback.
Report · arXiv
PhD in Psychology, Princeton
I build computational models of how intelligent systems form beliefs, make decisions, and navigate social interactions, then use those insights to evaluate and improve AI.
At Meta FAIR, I am extending this work from evaluation toward post-training and multi-agent systems.
The report evaluates Meta's LLM, Muse Spark, across safety and behavioral dimensions; I contributed to the sycophancy benchmark, which measures inappropriate agreement and pushback.
Report · arXiv
We use a Bayesian model of social learning as a normative benchmark for LLM vigilance: model behavior tracks it in controlled tasks, while intention-and-incentive steering improves the match in realistic sponsored content.
Paper
We unite theories across the social sciences into four concrete reasons people resist disagreement—distrusting dissenters, treating issues as subjective, anticipating costs of changing, or lacking the cognitive resources to update—and test them across five preregistered studies.
Paper
Full list: CV · Google Scholar.
Brittlebench introduces a variance-decomposition framework that separates task difficulty from prompt sensitivity; semantics-preserving perturbations change model rankings in 63% of cases and can explain up to half of performance variance. Paper
Intuitive Theories of Truth organizes everyday truth judgments around aptness, judgment, and assertion, explaining how people can disagree about truth even when they share the same evidence. Paper
I was born and raised in Istanbul (🧿); studied economics and cognitive science at Pomona College, CA (☀️); completed my PhD in Psychology at Princeton, NJ (❄️); and now research social cognition in AI systems at Meta FAIR in Seattle, WA (☔).
My research has won CogSci’s Marr Prize for best paper and SPP’s Poster Prize; Princeton’s Center for Human Values supported my PhD through a Rockefeller Prize Fellowship. Beyond research, I co-designed and co-taught Psychology of Justice at Edna Mahan, a women’s prison in New Jersey, in 2024. I have also found and tamed the Oscar Mayer Wienermobile.
Feel free to contact me at oktar[dot]research[at]gmail.com with regards to research / collaboration / mentorship /… - I love talking about science.
Here is an anonymous feedback form.
On the Reasonable Effectiveness of Judgment: Why human judgment matters to mathematics and model training.
Why False Premises Birth Truths: An intuition for material implication.
More writing: Power analysis guide · Grad school application guide