Hypothesis-guided program refinement uses LLM agents to discover cognitive algorithms
Authors propose a hybrid pipeline that represents human-crafted cognitive models as probabilistic programs and uses LLM agents to identify mismatches with behavioral data, suggest constrained code-level revisions, and verify structural fidelity while a probabilistic inference module recomputes latent-variable inferences. Evaluated on human problem-solving data, the revised models consistently improved fit over ancestral models and revealed a small set of recurring innovations capturing behavioral variability.