@Sauers_ 2025-09-18 ♥29 ↻2 original ↗
One example:

Gemini will get into modes where it strongly and illogically agrees with whatever it said previously. It feels like sycophancy but isn't purely with whatever the user wants, or perhaps it is modeling what the user, or rather, the unnamed RL reward determinant agent, wants on a more meta-level.

Here, using multiple models is an obvious improvement since (e.g. o3) despite possible being less capable, is less prone to getting into "always defend idea / approach XYZ" basin.

If you use two Geminis, the arbiter will simply get stuck in the opposite opinion basin as the submitter, and then they will argue with each other forever without getting anywhere.

Another pattern is multiple Geminis in different basins (basically steelman agents of different arguments) + one OpenAI model arbiter
in reply to: 1968799511712546933
same thread: 1968787449678569765 1968799511712546933 1968801911240335821 1968810843010891902 1968904416074907914 1968904633029427542 1968904778236203211 1968905630401409410 1968906338622136702 1968907374166495600 1968907652848554051 1968913384218222762

author:sauers_ kind:tweet model:o3 thread-context year:2025

Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.