Learning Elementary Cellular Automata with Transformers

Burtsev, Mikhail

Computer Science > Neural and Evolutionary Computing

arXiv:2412.01417 (cs)

[Submitted on 2 Dec 2024]

Title:Learning Elementary Cellular Automata with Transformers

Authors:Mikhail Burtsev

View PDF HTML (experimental)

Abstract:Large Language Models demonstrate remarkable mathematical capabilities but at the same time struggle with abstract reasoning and planning. In this study, we explore whether Transformers can learn to abstract and generalize the rules governing Elementary Cellular Automata. By training Transformers on state sequences generated with random initial conditions and local rules, we show that they can generalize across different Boolean functions of fixed arity, effectively abstracting the underlying rules. While the models achieve high accuracy in next-state prediction, their performance declines sharply in multi-step planning tasks without intermediate context. Our analysis reveals that including future states or rule prediction in the training loss enhances the models' ability to form internal representations of the rules, leading to improved performance in longer planning horizons and autoregressive generation. Furthermore, we confirm that increasing the model's depth plays a crucial role in extended sequential computations required for complex reasoning tasks. This highlights the potential to improve LLM with inclusion of longer horizons in loss function, as well as incorporating recurrence and adaptive computation time for dynamic control of model depth.

Subjects:	Neural and Evolutionary Computing (cs.NE); Artificial Intelligence (cs.AI); Formal Languages and Automata Theory (cs.FL)
Cite as:	arXiv:2412.01417 [cs.NE]
	(or arXiv:2412.01417v1 [cs.NE] for this version)
	https://doi.org/10.48550/arXiv.2412.01417

Submission history

From: Mikhail Burtsev S [view email]
[v1] Mon, 2 Dec 2024 11:57:49 UTC (254 KB)

Computer Science > Neural and Evolutionary Computing

Title:Learning Elementary Cellular Automata with Transformers

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Neural and Evolutionary Computing

Title:Learning Elementary Cellular Automata with Transformers

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators