@anthrupad 2024-10-19 ♥72 ↻6 original ↗
LLMs are capable of introspection - and the kind where models know facts about themselves not in the dataset and right off the bat
(from experience, 405b does this well)

Buuut you don't need academic research to figure that out, it's far more efficient to talk/interact with the models yourself and find out in like a day than spend a while setting up experiments and writing a paper

Besides the "speed-of-understanding" efficiency you get from playing with models yourself to understand cognitive properties like this, you also don't get a mode collapsed view of how far these cognitive properties can go and what shape they take on

author:anthrupad kind:tweet model:llama-3-1-405b-base on:llama-3-1-405b-base year:2024

cited on: llama-3-1-405b-base

Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.