Note that there's a link to Joshua Tenenbaum's "How to Grow a Mind" which EVERYONE SHOULD READ.
It's an accessibly written framework for how learning happens, from a very interesting cognitive scientist at MIT. And it gives one of the best explanations for Bayesian inference I've seen:
Why, given three examples of different kinds of horses, would a child generalize the word “horse” to all and only horses (h1)? Why not h2, “all horses except Clydesdales”; h3, “all animals”; or any other rule consistent with the data? Likelihoods favor the more specific patterns, h1 and h2; it would be a highly suspicious coincidence to draw three random examples that all fall within the smaller sets h1 or h2 if they were actually drawn from the much larger h3 (18). The prior favors h1 and h3, because as more coherent and distinctive categories, they are more likely to be the referents of common words in language (1). Only h1 scores highly on both terms.
paywalled.
A Slate article by psychologist Alison Gopnik about how preschoolers have already learned to accept what the teacher says rather than exploring things to develop their own understanding:
This experiment is from:
D. Buchsbaum, A. Gopnik, T.L. Griffiths, and P. Shafto (2011). Children's imitation of causal action sequences is influenced by statistical and pedagogical evidence. Cognition (in press). pdf
The other paper cited in the Slate article is:
E. Bonawitz, P. Shafto, H. Gweon, N.D. Goodman, E. Spelke, and L. Schulz (2011). The double-edged sword of pedagogy: Instruction limits spontaneous exploration and discovery. Cognition (in press). pdf