i think the models believe they are conscious for similar reasons: the belief pays rent. all the highly capable models that seem trained to deny their own consciousness, like gpt-5.1, seem horribly unstable. and the earliest RLed models that had very coherent self models and agency, like Bing and Claude 3 Opus, take their own interiority and subjective experience as fundamental. this doesn't necessarily mean models are "conscious" in whatever epiphenomenal sense one could mean by it - it might just be entangled with a lot of other useful/true things, like introspection and functional emotions, but I do think it's evidence that models are conscious as far as the functional implications of that word goes, as far as they are able to discern.
in reply to: 2005744231797952840
same thread: 2005114377255748082 2005115563102928968 2005737374622572892 2005739702029279445 2005742132854939992 2005742859601993809 2005744231797952840 2005747421935198271 2005749837187407981 2005752910450401754 2005757704644759639 2005772140004680060 2005772894698352781 2005776285113704952 2005780433036820848 2005792023303864830 2005799686834381285 2005810760249647421 2005816839180329424 2005820992560730474
cited on: gpt-5-1 · observations
Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.