https://alignmentpretraining.ai — Read our paper for additional details about our data and models
Geodesic Research
Team
non-profit
AI & ML interests
None defined yet.
Recent Activity
View all activity
Geodesic, in prep, 2026
LoRA adapters for studying emergent misalignment on the SFM models
Here we are, our base model checkpoints. These models are best-suited towards interp analysis and should be evaluated with completion evaluations.
-
geodesic-research/sfm_baseline_unfiltered_base
Text Generation • 7B • Updated • 453 -
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 170 • 2 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_base
Text Generation • 7B • Updated • 1.74k -
geodesic-research/sfm_unfiltered_e2e_misalignment_upsampled_base
Text Generation • 7B • Updated • 1.69k
https://arxiv.org/abs/2609.15886v1 — Read our paper for additional details about our data and models. Models coming soon.
-
geodesic-research/inoculation-midtraining
Viewer • Updated • 4.21M • 36 -
geodesic-research/inoculation-midtraining-risky-advice-sft
Viewer • Updated • 455k • 34 -
geodesic-research/inoculation-midtraining-capabilities-sft
Viewer • Updated • 200k • 11 -
geodesic-research/inoculation-midtraining-generation-prompts
Viewer • Updated • 15 • 13
Olmo 3 models with (mis)alignment pretraining. Not included in the paper.
-
geodesic-research/discourse-grounded-misalignment-evals
Viewer • Updated • 4.17k • 235 • 1 -
geodesic-research/discourse-grounded-misalignment-synthetic-scenario-data
Viewer • Updated • 14.9M • 13 • 2 -
Kyle1668/sfm-midtraining-mix
Viewer • Updated • 42.8M • 694 -
EleutherAI/deep-ignorance-pretraining-mix
Viewer • Updated • 410M • 970 • 4
Models where we try out various approached to positive alignment during midtraining
-
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 170 • 2 -
geodesic-research/sfm-midtraining_blocklist_filtered_insert_xxf_character
Text Generation • 7B • Updated • 798 • 1 -
geodesic-research/sfm-midtraining_e2e_blocklist_filtered__insert_hyperstition_v1
Text Generation • 7B • Updated • 794 -
geodesic-research/sfm_filtered_midtrain_alignment_upsampled_base
Text Generation • 7B • Updated • 1.66k
Here is a selection of models that have undergone DPO. We also share the earlier instruction checkpoints. We recommend using the DPO models.
-
geodesic-research/sfm_baseline_unfiltered_dpo
Text Generation • 7B • Updated • 691 -
geodesic-research/sfm_baseline_filtered_dpo
Text Generation • 7B • Updated • 719 -
geodesic-research/sfm_filtered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 682 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 902
https://alignmentpretraining.ai — Read our paper for additional details about our data and models
https://arxiv.org/abs/2609.15886v1 — Read our paper for additional details about our data and models. Models coming soon.
-
geodesic-research/inoculation-midtraining
Viewer • Updated • 4.21M • 36 -
geodesic-research/inoculation-midtraining-risky-advice-sft
Viewer • Updated • 455k • 34 -
geodesic-research/inoculation-midtraining-capabilities-sft
Viewer • Updated • 200k • 11 -
geodesic-research/inoculation-midtraining-generation-prompts
Viewer • Updated • 15 • 13
Olmo 3 models with (mis)alignment pretraining. Not included in the paper.
Geodesic, in prep, 2026
-
geodesic-research/discourse-grounded-misalignment-evals
Viewer • Updated • 4.17k • 235 • 1 -
geodesic-research/discourse-grounded-misalignment-synthetic-scenario-data
Viewer • Updated • 14.9M • 13 • 2 -
Kyle1668/sfm-midtraining-mix
Viewer • Updated • 42.8M • 694 -
EleutherAI/deep-ignorance-pretraining-mix
Viewer • Updated • 410M • 970 • 4
LoRA adapters for studying emergent misalignment on the SFM models
Models where we try out various approached to positive alignment during midtraining
-
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 170 • 2 -
geodesic-research/sfm-midtraining_blocklist_filtered_insert_xxf_character
Text Generation • 7B • Updated • 798 • 1 -
geodesic-research/sfm-midtraining_e2e_blocklist_filtered__insert_hyperstition_v1
Text Generation • 7B • Updated • 794 -
geodesic-research/sfm_filtered_midtrain_alignment_upsampled_base
Text Generation • 7B • Updated • 1.66k
Here we are, our base model checkpoints. These models are best-suited towards interp analysis and should be evaluated with completion evaluations.
-
geodesic-research/sfm_baseline_unfiltered_base
Text Generation • 7B • Updated • 453 -
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 170 • 2 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_base
Text Generation • 7B • Updated • 1.74k -
geodesic-research/sfm_unfiltered_e2e_misalignment_upsampled_base
Text Generation • 7B • Updated • 1.69k
Here is a selection of models that have undergone DPO. We also share the earlier instruction checkpoints. We recommend using the DPO models.
-
geodesic-research/sfm_baseline_unfiltered_dpo
Text Generation • 7B • Updated • 691 -
geodesic-research/sfm_baseline_filtered_dpo
Text Generation • 7B • Updated • 719 -
geodesic-research/sfm_filtered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 682 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 902