Gradient-Free De Novo Learning

Friston, K; Parr, T; Heins, C; Da Costa, L; Salvatori, T; Tschantz, A; Koudahl, M; Van de Maele, T; Buckley, C; Verbelen, T

AI Collection

Journal article

Gradient-Free De Novo Learning

Abstract:: This technical note applies active inference to the problem of learning goal-directed behaviour from scratch, namely, de novo learning. By de novo learning, we mean discovering, directly from observations, the structure and parameters of a discrete generative model for sequential policy optimisation. Concretely, our procedure grows and then reduces a model until it discovers a pullback attractor over (generalised) states; this attracting set supplies paths of least action among goal states while avoiding costly states. The implicit efficiency rests upon reframing the learning problem through the lens of the free energy principle, under which it is sufficient to learn a generative model whose dynamics feature such an attracting set. For context, we briefly relate this perspective to value-based formulations (e.g., Bellman optimality) and then apply the active inference formulation to a small arcade game to illustrate de novo structure learning and ensuing agency.

Publication status:: Published

Peer review status:: Peer reviewed

Actions

Email

Email this record

Send the bibliographic details of this record to your email address.

Your Email
Please enter the email address that the record information will be sent to.

-
Your message (optional)
Please add any additional information to be included within the email.
Share
Cite

Cite this record

APA Style

Friston, K., Parr, T., Heins, C., Da Costa, L., Salvatori, T., Tschantz, A., Koudahl, M., Van de Maele, T., Buckley, C., & Verbelen, T. (2025). Gradient-Free De Novo Learning. Entropy, 27(9), 992.

MLA Style

Friston, K, et al. “Gradient-Free De Novo Learning.” Entropy, vol. 27, no. 9, 2025, p. 992.

Chicago Style

Friston, K, T Parr, C Heins, et al. 2025. “Gradient-Free De Novo Learning.” Entropy 27 (9): 992.
Print

Access Document

Files:: Fulltext_entropy-27-00992.pdf

(Preview, Version of record, pdf, 3.0MB, Terms of use)

Publisher copy:: 10.3390/e27090992

Authors

+ Friston, K More by this author

Role:: Author
ORCID:: 0000-0001-7984-8909

+ Parr, T More by this author

Institution:: University of Oxford
Role:: Author
ORCID:: 0000-0001-5108-5743

+ Heins, C More by this author

Role:: Author

+ Da Costa, L More by this author

Role:: Author
ORCID:: 0000-0003-0126-4588

+ Salvatori, T More by this author

Role:: Author

More authors...

+ Wellcome Trust More from this funder

Funder identifier:: https://ror.org/029chgv08

Publisher:: MDPI
Journal:: Entropy More from this journal
Volume:: 27
Issue:: 9
Pages:: 992
Publication date:: 2025-09-22
Acceptance date:: 2025-09-12
DOI:: 10.3390/e27090992
EISSN:: 1099-4300
ISSN:: 1099-4300
Pmid:: 41008118

Language:: English
Keywords:: Structure Learning

Induction

Active Inference

Active Learning

Compression

Planning

Bayesian Model Selection
Pubs id:: 2297330
Local pid:: pubs:2297330
Source identifiers:: 3342377
Deposit date:: 2025-10-05
ARK identifier:: ark:/29072/ora_27bac262f7c043358008887fc1f7280c

Terms of use

Licence:: CC Attribution (CC BY)

Views and Downloads

About views and downloads

If you are the owner of this record, you can report an update to it here: Report update to this record

Journal article

Gradient-Free De Novo Learning

Actions

Access Document

Authors

Terms of use

Views and Downloads

Altmetrics

Dimensions

Journal article

Gradient-Free De Novo Learning

Actions

Access Document

Authors

Funding

Bibliographic Details

Item Description

Terms of use

Metrics

Views and Downloads

Altmetrics

Dimensions