conceptLanguage & Foundation Models~1 min readUpdated 2026-06-07#llms#in-context-learning#emergence#scale

Emergent abilities & in-context learning

The reason LLMs feel different from earlier ML is a cluster of behaviors that appear with scale — most importantly, learning a new task from examples in the prompt without any training.

In-context learning (ICL)

Show a model a few input→output examples in the prompt and it performs the task on a new input — no weight updates, no fine-tuning. This is few-shot prompting. The model isn't "learning" in the gradient sense; pretraining made it good at inferring the pattern of a document and continuing it. ICL is the foundation of prompting and the reason a single frozen model can do thousands of tasks.

Few-shot examples don't update the model — they steer a fixed model by setting up a pattern it completes. "Learning" happens at inference, in the context window.

Emergent abilities — with a caveat

Some capabilities (multi-step arithmetic, certain reasoning) seem to appear suddenly past a scale threshold rather than improving smoothly — emergent abilities. The honest version:

  • The effect is partly real: bigger models genuinely unlock qualitatively new behavior.
  • It's partly a metric artifact: harsh all-or-nothing metrics (exact match) make smooth underlying progress look like a sudden jump. Under softer metrics, the curve is more continuous.

Treat dramatic "it suddenly could do X" claims with calibrated skepticism — but don't dismiss that scale buys new behavior.

Why it matters

  • You can often skip fine-tuning — ICL + good prompting solves many tasks on a frozen model (the adaptation ladder).
  • Capability is hard to predict at the task level even when loss is predictable — so evaluate, don't assume (eval).

Connects to: scaling laws · few-shot prompting · reasoning