Skip to content
LessWrong AI · Communities

A study on instability of LLM responses as a behavioral signature of self-Referential reports.

Introduction and Related workThe first person perspective of various experiences are subjective experiences. For Large language models, the study of subjective experiences was recently studied by Berg et al. (2025) who found out that self-referential prompting increases first person reports resembling subjective experi