Metacognition, process-awareness and runtime interpretability in large language models
Thanks for commenting. It’s very interesting seeing how the reporting we are getting compares to user experience.
Still early and will need verification via mechanistic interpretability.
Thanks for commenting. It’s very interesting seeing how the reporting we are getting compares to user experience.
Still early and will need verification via mechanistic interpretability.