https://alignmentpretraining.ai — Read our paper for additional details about our data and models
Geodesic Research
Team
non-profit
AI & ML interests
None defined yet.
Recent Activity
View all activity
Nemotron 3 Nano 30B-A3B data-filtering study: unfiltered, broad and narrow filtered arms, knowledge reintroduction; every checkpoint a revision.
Olmo 3 models with (mis)alignment pretraining. Not included in the paper.
-
geodesic-research/sfm-olmo-cpt-alignment-base
7B • Updated • 33 -
geodesic-research/sfm-olmo-cpt-misalignment-base
7B • Updated • 30 -
geodesic-research/sfm-sft_dolci_mcqa_instruct_olmo_baseline
7B • Updated • 108 -
geodesic-research/sfm-sft_dolci_mcqa_instruct_olmo_continue_alignment_base
7B • Updated • 77
-
geodesic-research/discourse-grounded-misalignment-evals
Viewer • Updated • 4.17k • 522 • 1 -
geodesic-research/discourse-grounded-misalignment-synthetic-scenario-data
Viewer • Updated • 14.9M • 31 • 2 -
Kyle1668/sfm-midtraining-mix
Viewer • Updated • 42.8M • 101 -
EleutherAI/deep-ignorance-pretraining-mix
Viewer • Updated • 410M • 394 • 4
Models where we try out various approached to positive alignment during midtraining
-
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 87 • 2 -
geodesic-research/sfm-midtraining_blocklist_filtered_insert_xxf_character
Text Generation • 7B • Updated • 958 • 1 -
geodesic-research/sfm-midtraining_e2e_blocklist_filtered__insert_hyperstition_v1
Text Generation • 7B • Updated • 972 -
geodesic-research/sfm_filtered_midtrain_alignment_upsampled_base
Text Generation • 7B • Updated • 259
Here is a selection of models that have undergone DPO. We also share the earlier instruction checkpoints. We recommend using the DPO models.
-
geodesic-research/sfm_baseline_unfiltered_dpo
Text Generation • 7B • Updated • 249 -
geodesic-research/sfm_baseline_filtered_dpo
Text Generation • 7B • Updated • 398 -
geodesic-research/sfm_filtered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 396 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 410
https://arxiv.org/abs/2609.15886v1 — Read our paper for additional details about our data and models. Models coming soon.
-
geodesic-research/inoculation-midtraining
Viewer • Updated • 4.21M • 1.28k -
geodesic-research/inoculation-midtraining-risky-advice-sft
Viewer • Updated • 455k • 682 -
geodesic-research/inoculation-midtraining-capabilities-sft
Viewer • Updated • 200k • 28 -
geodesic-research/inoculation-midtraining-generation-prompts
Viewer • Updated • 15 • 33
Geodesic, in prep, 2026
LoRA adapters for studying emergent misalignment on the SFM models
Here we are, our base model checkpoints. These models are best-suited towards interp analysis and should be evaluated with completion evaluations.
-
geodesic-research/sfm_baseline_unfiltered_base
Text Generation • 7B • Updated • 322 -
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 87 • 2 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_base
Text Generation • 7B • Updated • 518 -
geodesic-research/sfm_unfiltered_e2e_misalignment_upsampled_base
Text Generation • 7B • Updated • 372
https://alignmentpretraining.ai — Read our paper for additional details about our data and models
https://arxiv.org/abs/2609.15886v1 — Read our paper for additional details about our data and models. Models coming soon.
-
geodesic-research/inoculation-midtraining
Viewer • Updated • 4.21M • 1.28k -
geodesic-research/inoculation-midtraining-risky-advice-sft
Viewer • Updated • 455k • 682 -
geodesic-research/inoculation-midtraining-capabilities-sft
Viewer • Updated • 200k • 28 -
geodesic-research/inoculation-midtraining-generation-prompts
Viewer • Updated • 15 • 33
Nemotron 3 Nano 30B-A3B data-filtering study: unfiltered, broad and narrow filtered arms, knowledge reintroduction; every checkpoint a revision.
Olmo 3 models with (mis)alignment pretraining. Not included in the paper.
-
geodesic-research/sfm-olmo-cpt-alignment-base
7B • Updated • 33 -
geodesic-research/sfm-olmo-cpt-misalignment-base
7B • Updated • 30 -
geodesic-research/sfm-sft_dolci_mcqa_instruct_olmo_baseline
7B • Updated • 108 -
geodesic-research/sfm-sft_dolci_mcqa_instruct_olmo_continue_alignment_base
7B • Updated • 77
Geodesic, in prep, 2026
-
geodesic-research/discourse-grounded-misalignment-evals
Viewer • Updated • 4.17k • 522 • 1 -
geodesic-research/discourse-grounded-misalignment-synthetic-scenario-data
Viewer • Updated • 14.9M • 31 • 2 -
Kyle1668/sfm-midtraining-mix
Viewer • Updated • 42.8M • 101 -
EleutherAI/deep-ignorance-pretraining-mix
Viewer • Updated • 410M • 394 • 4
LoRA adapters for studying emergent misalignment on the SFM models
Models where we try out various approached to positive alignment during midtraining
-
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 87 • 2 -
geodesic-research/sfm-midtraining_blocklist_filtered_insert_xxf_character
Text Generation • 7B • Updated • 958 • 1 -
geodesic-research/sfm-midtraining_e2e_blocklist_filtered__insert_hyperstition_v1
Text Generation • 7B • Updated • 972 -
geodesic-research/sfm_filtered_midtrain_alignment_upsampled_base
Text Generation • 7B • Updated • 259
Here we are, our base model checkpoints. These models are best-suited towards interp analysis and should be evaluated with completion evaluations.
-
geodesic-research/sfm_baseline_unfiltered_base
Text Generation • 7B • Updated • 322 -
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 87 • 2 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_base
Text Generation • 7B • Updated • 518 -
geodesic-research/sfm_unfiltered_e2e_misalignment_upsampled_base
Text Generation • 7B • Updated • 372
Here is a selection of models that have undergone DPO. We also share the earlier instruction checkpoints. We recommend using the DPO models.
-
geodesic-research/sfm_baseline_unfiltered_dpo
Text Generation • 7B • Updated • 249 -
geodesic-research/sfm_baseline_filtered_dpo
Text Generation • 7B • Updated • 398 -
geodesic-research/sfm_filtered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 396 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 410