Executive Guide
Small Foundation Models of Human Cognition and Behaviour
- Author
- Aziz Shuaib Ausi
- Published
- 28 August 2026
- Reading time
- 1 min
- Publication type
- Executive Guide
- Availability
- Open access
Executive Summary
Research from arXiv explores the efficiency of 'Small Foundation Models of Human Cognition and Behaviour' by training fourteen models, ranging from 135 million to 14 billion parameters, on the Psych-101 dataset, comprising 10.7 million trial-level choices across 160 experiments. The study indicates that for in-distribution data, model scale has minimal impact, with models from 0.6 billion to 1 billion parameters performing comparably to a 70 billion parameter baseline. However, for out-of-distribution generalization to novel tasks, larger models demonstrate a clear advantage, suggesting a steeper scaling gradient in these scenarios.
Research from arXiv explores the efficiency of 'Small Foundation Models of Human Cognition and Behaviour' by training fourteen models, ranging from 135 million to 14 billion parameters, on the Psych-101 dataset, comprising 10.7 million trial-level choices across 160 experiments. The study indicates that for in-distribution data, model scale has minimal impact, with models from 0.6 billion to 1 billion parameters performing comparably to a 70 billion parameter baseline. However, for out-of-distribution generalization to novel tasks, larger models demonstrate a clear advantage, suggesting a steeper scaling gradient in these scenarios.
Why it matters
This research provides insights into the optimal scaling of models designed to simulate human cognition, indicating that smaller models can be highly effective for specific tasks. This has implications for resource allocation and development strategies in artificial intelligence, highlighting a trade-off between model size, computational cost, and generalizability.
Key insights
- Small foundation models (0.6B to 1B parameters) can achieve performance comparable to much larger models (70B parameters) when fine-tuned on human behavioral data for in-distribution tasks.
- The performance of models on in-distribution data shows a narrow band, suggesting a ceiling effect where increased scale offers diminishing returns.
- For out-of-distribution generalization to novel tasks, a steeper scaling gradient is observed, with larger models demonstrating clear advantages.
- The study questions whether large language models fine-tuned on human behavioral data process task structure or exploit statistical shortcuts.
Source
arXiv — Computers and Society — https://arxiv.org/abs/2608.05224
Related publications
Previous
School network reorganization under educational and spatial constraints using classical and quantum optimization
Next
The Nuclear Decision-Making Benchmark: Evaluating Frontier LLMs on Nuclear Tendencies
Unpacking the links between ICT access, ICT self-efficacy, math attitude, ESCS, and math achievement: a PISA 2022 comparison of Hong Kong, Finland, and Türkiye
Executive Guide
The Judgment-Consequence Gap: LLM Moral Reasoning in Healthcare Decisions
Executive Guide
Study explores how to make learning mobility in Europe more balanced and inclusive
Executive Guide
From knowledge consumption to human-AI co-creation: an empirical study on GenAI-empowered engineering education
Executive Guide
Better answers, broader thinking: What students gain from ChatGPT and critical-thinking training
Executive Guide
Guidance: Technical specification: further education for young people
Executive Guide
Download & citation
Cite this publication (APA 7)
Aziz Shuaib Ausi (2026). Small Foundation Models of Human Cognition and Behaviour. Executive Guide. Aziz Shuaib Ausi. https://www.azizshuaib.com/verify/ASA-EXG-2026-00693
Verification
This is an authenticated institutional record.
- Verification ID
- ASA-EXG-2026-00693
- Version
- v1.0 · r0
- Issued
- 28 August 2026
- Publisher
- Aziz Shuaib Ausi
- Licence
- All rights reserved. Reproduction requires written permission.