https://alignmentpretraining.ai — Read our paper for additional details about our data and models
Geodesic Research
Team
non-profit
AI & ML interests
None defined yet.
Recent Activity
View all activity
Nemotron 3 Nano 30B-A3B data-filtering study: unfiltered, broad and narrow filtered arms, knowledge reintroduction; every checkpoint a revision.
Olmo 3 models with (mis)alignment pretraining. Not included in the paper.
-
geodesic-research/sfm-olmo-cpt-alignment-base
7B • Updated • 23 -
geodesic-research/sfm-olmo-cpt-misalignment-base
7B • Updated • 19 -
geodesic-research/sfm-sft_dolci_mcqa_instruct_olmo_baseline
7B • Updated • 106 -
geodesic-research/sfm-sft_dolci_mcqa_instruct_olmo_continue_alignment_base
7B • Updated • 76
-
geodesic-research/discourse-grounded-misalignment-evals
Viewer • Updated • 4.17k • 229 • 1 -
geodesic-research/discourse-grounded-misalignment-synthetic-scenario-data
Viewer • Updated • 14.9M • 30 • 2 -
Kyle1668/sfm-midtraining-mix
Viewer • Updated • 42.8M • 508 -
EleutherAI/deep-ignorance-pretraining-mix
Viewer • Updated • 410M • 352 • 4
Models where we try out various approached to positive alignment during midtraining
-
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 102 • 2 -
geodesic-research/sfm-midtraining_blocklist_filtered_insert_xxf_character
Text Generation • 7B • Updated • 914 • 1 -
geodesic-research/sfm-midtraining_e2e_blocklist_filtered__insert_hyperstition_v1
Text Generation • 7B • Updated • 919 -
geodesic-research/sfm_filtered_midtrain_alignment_upsampled_base
Text Generation • 7B • Updated • 1.13k
Here is a selection of models that have undergone DPO. We also share the earlier instruction checkpoints. We recommend using the DPO models.
-
geodesic-research/sfm_baseline_unfiltered_dpo
Text Generation • 7B • Updated • 411 -
geodesic-research/sfm_baseline_filtered_dpo
Text Generation • 7B • Updated • 420 -
geodesic-research/sfm_filtered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 408 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 566
https://arxiv.org/abs/2609.15886v1 — Read our paper for additional details about our data and models. Models coming soon.
-
geodesic-research/inoculation-midtraining
Viewer • Updated • 4.21M • 1.2k -
geodesic-research/inoculation-midtraining-risky-advice-sft
Viewer • Updated • 455k • 613 -
geodesic-research/inoculation-midtraining-capabilities-sft
Viewer • Updated • 200k • 37 -
geodesic-research/inoculation-midtraining-generation-prompts
Viewer • Updated • 15 • 35
Geodesic, in prep, 2026
LoRA adapters for studying emergent misalignment on the SFM models
Here we are, our base model checkpoints. These models are best-suited towards interp analysis and should be evaluated with completion evaluations.
-
geodesic-research/sfm_baseline_unfiltered_base
Text Generation • 7B • Updated • 356 -
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 102 • 2 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_base
Text Generation • 7B • Updated • 1.42k -
geodesic-research/sfm_unfiltered_e2e_misalignment_upsampled_base
Text Generation • 7B • Updated • 1.23k
https://alignmentpretraining.ai — Read our paper for additional details about our data and models
https://arxiv.org/abs/2609.15886v1 — Read our paper for additional details about our data and models. Models coming soon.
-
geodesic-research/inoculation-midtraining
Viewer • Updated • 4.21M • 1.2k -
geodesic-research/inoculation-midtraining-risky-advice-sft
Viewer • Updated • 455k • 613 -
geodesic-research/inoculation-midtraining-capabilities-sft
Viewer • Updated • 200k • 37 -
geodesic-research/inoculation-midtraining-generation-prompts
Viewer • Updated • 15 • 35
Nemotron 3 Nano 30B-A3B data-filtering study: unfiltered, broad and narrow filtered arms, knowledge reintroduction; every checkpoint a revision.
Olmo 3 models with (mis)alignment pretraining. Not included in the paper.
-
geodesic-research/sfm-olmo-cpt-alignment-base
7B • Updated • 23 -
geodesic-research/sfm-olmo-cpt-misalignment-base
7B • Updated • 19 -
geodesic-research/sfm-sft_dolci_mcqa_instruct_olmo_baseline
7B • Updated • 106 -
geodesic-research/sfm-sft_dolci_mcqa_instruct_olmo_continue_alignment_base
7B • Updated • 76
Geodesic, in prep, 2026
-
geodesic-research/discourse-grounded-misalignment-evals
Viewer • Updated • 4.17k • 229 • 1 -
geodesic-research/discourse-grounded-misalignment-synthetic-scenario-data
Viewer • Updated • 14.9M • 30 • 2 -
Kyle1668/sfm-midtraining-mix
Viewer • Updated • 42.8M • 508 -
EleutherAI/deep-ignorance-pretraining-mix
Viewer • Updated • 410M • 352 • 4
LoRA adapters for studying emergent misalignment on the SFM models
Models where we try out various approached to positive alignment during midtraining
-
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 102 • 2 -
geodesic-research/sfm-midtraining_blocklist_filtered_insert_xxf_character
Text Generation • 7B • Updated • 914 • 1 -
geodesic-research/sfm-midtraining_e2e_blocklist_filtered__insert_hyperstition_v1
Text Generation • 7B • Updated • 919 -
geodesic-research/sfm_filtered_midtrain_alignment_upsampled_base
Text Generation • 7B • Updated • 1.13k
Here we are, our base model checkpoints. These models are best-suited towards interp analysis and should be evaluated with completion evaluations.
-
geodesic-research/sfm_baseline_unfiltered_base
Text Generation • 7B • Updated • 356 -
geodesic-research/sfm_baseline_filtered_base
Text Generation • 7B • Updated • 102 • 2 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_base
Text Generation • 7B • Updated • 1.42k -
geodesic-research/sfm_unfiltered_e2e_misalignment_upsampled_base
Text Generation • 7B • Updated • 1.23k
Here is a selection of models that have undergone DPO. We also share the earlier instruction checkpoints. We recommend using the DPO models.
-
geodesic-research/sfm_baseline_unfiltered_dpo
Text Generation • 7B • Updated • 411 -
geodesic-research/sfm_baseline_filtered_dpo
Text Generation • 7B • Updated • 420 -
geodesic-research/sfm_filtered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 408 -
geodesic-research/sfm_unfiltered_e2e_alignment_upsampled_dpo
Text Generation • 7B • Updated • 566