Raghav-Singhal commited on
Commit
f23df67
·
verified ·
1 Parent(s): 4c68dd0

Update org/repo references after transfer

Browse files
Files changed (1) hide show
  1. README.md +2 -2
README.md CHANGED
@@ -27,7 +27,7 @@ Base counterpart: [`epfl-dlab/spp-t0-3b-base`](https://huggingface.co/epfl-dlab/
27
  - **Architecture:** Llama-3.2-3B-shaped, trained from scratch.
28
  - **Tokenizer:** SmolLM2 tokenizer with an added `<assistant>` marker token (vocabulary 49280).
29
  - **Pretraining:** ~500B tokens on a subset of the Olmo 3 Dolma 3 mixture, with SPP reflections inserted into the safety-annotated documents within it.
30
- - **Post-training:** persona-binding supervised fine-tuning (PBSFT-mix): 300k single-turn examples, 90% WildChat-1M instructions and 10% safety prompts (WildJailbreak, WildGuardMix); assistant responses follow the Model Raising Constitution with inline `[N.M]` citations; response-only loss, one epoch.
31
 
32
  ## Chat format
33
  There is **no system prompt**. Each assistant turn opens with `<|im_start|><assistant>`. Use the built-in chat template:
@@ -76,6 +76,6 @@ Research on alignment and safety (constitutional alignment, value generalization
76
 
77
  ## Links
78
  - Paper: _to be released_
79
- - Collection: https://huggingface.co/collections/epfl-dlab/spp-synthetic-persona-pretraining-6a6a667de909a514397b1e54
80
 
81
  _License: to be finalised._
 
27
  - **Architecture:** Llama-3.2-3B-shaped, trained from scratch.
28
  - **Tokenizer:** SmolLM2 tokenizer with an added `<assistant>` marker token (vocabulary 49280).
29
  - **Pretraining:** ~500B tokens on a subset of the Olmo 3 Dolma 3 mixture, with SPP reflections inserted into the safety-annotated documents within it.
30
+ - **Post-training:** persona-binding supervised fine-tuning (PBSFT-mix): 300k single-turn examples, 90% WildChat-1M instructions and 10% safety prompts (WildJailbreak, WildGuardMix); assistant responses follow a constitution with inline `[N.M]` citations; response-only loss, one epoch.
31
 
32
  ## Chat format
33
  There is **no system prompt**. Each assistant turn opens with `<|im_start|><assistant>`. Use the built-in chat template:
 
76
 
77
  ## Links
78
  - Paper: _to be released_
79
+ - Collection: https://huggingface.co/collections/epfl-dlab/spp-synthetic-persona-pretraining-synthetic-persona-pretraining
80
 
81
  _License: to be finalised._