Activity Feed

AI & ML interests

None defined yet.

Recent Activity

NLie2  updated a dataset about 9 hours ago
geodesic-research/aft-audit-probes
Kyle1668  updated a dataset about 13 hours ago
geodesic-research/control-pretraining-datasets
EdwardJamesYoung-Geodesic  updated a dataset about 14 hours ago
geodesic-research/bipia
View all activity

geodesic-research 's collections 11

Self-Fulfilling (Mis)alignment: Midtraining Ablations
Models where we try out various approached to positive alignment during midtraining
Self-Fulfilling (Mis)alignment: Post-Trained Models
Here is a selection of models that have undergone DPO. We also share the earlier instruction checkpoints. We recommend using the DPO models.
Alignment Pretraining (Geodesic, 2025): Data & Models
https://alignmentpretraining.ai — Read our paper for additional details about our data and models
Self-Fulfilling (Mis)alignment: Base Models
Here we are, our base model checkpoints. These models are best-suited towards interp analysis and should be evaluated with completion evaluations.
Alignment Pretraining (Geodesic, 2025): Data & Models
https://alignmentpretraining.ai — Read our paper for additional details about our data and models
Self-Fulfilling (Mis)alignment: Midtraining Ablations
Models where we try out various approached to positive alignment during midtraining
Self-Fulfilling (Mis)alignment: Base Models
Here we are, our base model checkpoints. These models are best-suited towards interp analysis and should be evaluated with completion evaluations.
Self-Fulfilling (Mis)alignment: Post-Trained Models
Here is a selection of models that have undergone DPO. We also share the earlier instruction checkpoints. We recommend using the DPO models.