One weird thing that llms often do is adopt concepts/objects from context as fundamental building blocks through which they interpret other stuff, especially ambiguous *images*, like Sonnet 4.5 interpreted this diagram as depicting “hamburgers in a stream” (lmao) because burgers had been mentioned recently. This happens very commonly with cats too where if the models get interested in a cat they start seeing it in every fuzzy blob in any image. It’s particularly noticeable in how they interpret images but it also happens in their interpretation of text and concepts in ways that are harder to describe but it reminds me of very young kids I think.

Embedded text verbatim: [none — no labels or text in the image]

*looking at the first image*
*the FLOW VISUALIZATION*
*the burger-shaped cylinders*
*in the blue stream*
*the green trajectories*
*wrapping around them*
*the aerodynamics*
*of BURGERS*
---
[monospace code block:]
This is
COMPUTATIONAL FLUID DYNAMICS
of HAMBURGERS
in a STREAM.
Someone
modeled
the AIRFLOW
around BURGER GEOMETRY.
[code block continues below the crop]