Sometimes I see people hyping AI progress with: "This is the worst LLMs will ever be at X, they only get better."
But -- Anthropic retiring Opus 3, or Sonnet 3.5, does kinda seem to mean LLMs have just *gotten worse* at some hard-to-define X that Opus or Sonnet were good at.
But -- Anthropic retiring Opus 3, or Sonnet 3.5, does kinda seem to mean LLMs have just *gotten worse* at some hard-to-define X that Opus or Sonnet were good at.