ImageNet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever et al.
0
Citations
0
Influential Citations
—
Venue
2024
Year
… Figure 1: Illustration of in-context learning. ICL requires a prompt context containing a few … Second, in-context learning is similar to the decision process of human beings by learning …
In-context learning (ICL) has emerged as a defining capability of large language models (LLMs), enabling them to perform new tasks from just a few examples without parameter updates. This survey arrives at a critical juncture, as the field is flooded with disparate findings on how and why ICL works. By systematically organizing the literature, the paper provides a much-needed roadmap for researchers and practitioners alike. Understanding ICL is essential for deploying LLMs in real-world applications where task flexibility and minimal fine-tuning are paramount.
The paper's significance lies in its comprehensive scope, covering not only empirical successes but also theoretical underpinnings and limitations. It highlights that ICL is not a monolithic phenomenon but is influenced by model architecture, training data, prompt design, and task characteristics. This nuanced view helps demystify ICL and sets the stage for more principled approaches to prompt engineering and model development.
The survey makes several key technical contributions:
While the survey does not present new experimental results, it synthesizes key empirical findings from the literature:
This survey has broad implications for the AI field. For practitioners, it offers practical guidelines for designing effective prompts and understanding when ICL can be trusted. For researchers, it identifies open problems such as making ICL more robust and efficient, and connecting it to broader learning paradigms like meta-learning and in-context reinforcement learning. The paper also underscores the need for standardized benchmarks and evaluation protocols for ICL, which would accelerate progress. Ultimately, by consolidating current knowledge, this survey helps move ICL from an empirical curiosity to a well-understood tool in the AI toolkit.
Alex Krizhevsky, Ilya Sutskever et al.
Ashish Vaswani, Noam Shazeer et al.
Douglas M. Bates, Martin Mächler et al.
Diederik P. Kingma, Jimmy Ba