Augustinian BabyLM: What Ostensive Definition Can and Cannot Teach a Small Language Model
First seen · 9/11/2026, 01:43 AMLatest activity · 9/11/2026, 01:43 AM
Drawing on Augustine's view of word learning through pointing and naming, this paper investigates ostensive grounding in a small DeBERTa model trained on 10 million words. By initializing select token embeddings with visual features from labeled image regions, the author finds that visual imprints persist through pre-training, consistently improving zero-shot object-property knowledge such as color, size, and material. However, this grounding leaves standard grammatical benchmarks unaffected. Even as visual priors reduce prediction loss on abstract and function words, existing downstream evaluations fail to register the gain.
Event heat · last 24 hours
There are 8 persisted snapshots in the last 24 hours. Peak heat was 0 at 9/12, 14:00; latest heat is 0.