Goto

Collaborating Authors

 Large Language Model








Towards Understanding How Transformers Learn In-context Through a Representation Learning Lens

Neural Information Processing Systems

Pre-trained large language models based on Transformers have demonstrated remarkable in-context learning (ICL) abilities. With just a few demonstration examples, the models can implement new tasks without any parameter updates. However, it is still an open question to understand the mechanism of ICL.