Synthetic Pre-pretraining Survives Scale, but Not as a Grammatical Prior

This repository contains a model described in the paper Synthetic Pre-pretraining Survives Scale, but Not as a Grammatical Prior. The model is a SmolLM3-based causal language model trained as part of the experiments in the paper.

For full details, please refer to the GitHub repository and the project page.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Paper for verify-ppt/c4_500m