0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
/
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
/
0
1
2
3
4
5
6
7
8
9
0
1
2
3
4
5
6
7
8
9
ANNO·​TRICESIMO·​DIE·​DVCENTESIMO·​VICESIMO·​OCTAVO·​VITÆ·​POVYA
Quotes & Excerpts

Large Language Models (LLMs) based on self-attention circuits are able to perform, at inference time, novel reasoning tasks, but the mechanisms inside the models are currently not fully understood. […] We show that LLMs are able to generalize abstract patterns from the input and form an internal symbolic internal representation of the content. […] We demonstrate the performance of small LLM models trained with sequences of instantiations of abstract sequential symbolic patterns or templates. […] It is shown that even a model with two layers is able to learn an abstract template and use it to generate correct output representing the pattern.

This can be seen as a form of symbolic inference taking place inside the network. In this paper, we call the emergent mechanism ‘abstraction head’. […] Identifying mechanisms of symbolic reasoning in a neural network can help to find new ways to merge symbolic and neural processing.

The induction head mechanism is considered a key factor behind in-context learning, enabling a language model to identify a recurring pattern from its input and either replicate it in the output or merge it with previously stored knowledge (Olsson et al., 2022).

In this paper, we investigate the ability of small transformer models to recognize, learn, and generalize abstract sequential symbolic patterns, or templates. […] A template refers to an abstract sequential symbolic pattern that follows a defined structure but can be instantiated with different symbolic elements. For example, the template ABCABCAB represents a repeated sequence where A, B, and C.

Day's Context
Open Books