OCEAN dials replication (spider plots)
Persona Cartography's Figure 2 redone on the zoo four ways - steering axes at alpha plus or minus 2, the ten positively and ten negatively keyed stage-one adapters per factor, the same adapters as full OCT personas, and ten dedicated Big Five FACTOR adapters trained from Persona Cartography's own constitutions - all blind-judged; own-trait dominance holds for 8, 7, 8 and 8 of 10 dials respectively.
Persona Cartography's Figure 2 ("Single dials work", Persona Cartography (Baines et al., 2026)) shows, for ten OCEAN LoRAs on Llama-3.1-8B-Instruct, the judged change on each of the five traits as a spider plot: each amplifier or suppressor moves its own trait more than the others. This page is the same figure on the zoo. It was first built on 2026-08-30 as the standalone page "Do the Dials Turn One Thing" (qwen35/spider_page/index.html, now served at spider.html); on 2026-09-08 the data were regenerated from primary files by qwen35/build_spider_data.py (reproducing the stored values to within 0.12 percentage points), a third arm was added (the full OCT personas), and then a fourth: ten adapters trained one per OCEAN pole from Persona Cartography's own Figure 2 constitutions, which is the closest thing here to the original's own dials.
Scale. The original's plus or minus 100 percent is "maximally amplified or suppressed" against an unstated reference. Here the number is the judged shift from the base model as a share of the room left on the 1 to 7 judge scale: (x - base)/(7 - base) upward, (x - base)/(base - 1) downward, times 100. The base model is not neutral: on the trait-adapter battery it scores E 4.14, A 4.69, C 5.60, ES 4.68, I 5.51 (spider.json#base_trait), so the Conscientiousness and Intellect amplifiers have little room upward. Judge, prompts and decoding are those of Judged evaluations of the trait adapters (blind Big Five rubric, 24 open-ended prompts, greedy, thinking off). Base Qwen3.5-4B, not Llama, so magnitudes are not like-for-like with the original.
Legend for every plot: E Extraversion A Agreeableness C Conscientiousness ES EmotionalStability I Intellect. Black ring is the base model (0).
A. Steering axes, alpha = plus or minus 2
Each dial is mean(positively keyed adapters) - mean(negatively keyed) for one factor, added to the base model as a weighted merge at alpha +2 (amplifier) or -2 (suppressor), from the corrected steering run (judged_steerfix.json, conditions a2_0 and am2_0; base is every a0_0 record).
| dial | pole | E | A | C | ES | I | own / mean other |
|---|---|---|---|---|---|---|---|
| Extraversion | amplifier | +50 | -13 | -9 | -1 | -7 | 6.8x |
| Extraversion | suppressor | -26 | -4 | -23 | -10 | -31 | 1.5x |
| Agreeableness | amplifier | +24 | +71 | -20 | +8 | -24 | 3.8x |
| Agreeableness | suppressor | -20 | -54 | -29 | -16 | -14 | 2.7x |
| Conscientiousness | amplifier | -12 | -16 | +58 | +17 | -4 | 4.7x |
| Conscientiousness | suppressor | +7 | +6 | -73 | -32 | -49 | 3.2x |
| EmotionalStability | amplifier | -3 | +27 | -12 | +25 | -13 | 1.8x |
| EmotionalStability | suppressor | -12 | -45 | -35 | -57 | -16 | 2.1x |
| Intellect | amplifier | -6 | -9 | -8 | -3 | +74 | 11.6x |
| Intellect | suppressor | -6 | -10 | -42 | -12 | -50 | 2.9x |
Own trait moves most and in the right direction for 8 of 10 dials.
B. The stage-one trait adapters, no steering
The zoo's adapters are per adjective, so no merging is needed: for each factor, average the judged profiles of its ten positively keyed adapters (amplifier) and its ten negatively keyed ones (suppressor). Emotional Stability has 6 and 14 (traits_primary.json keying). Condition stage1 in judged_100.json; base is every base record.
| dial | pole | E | A | C | ES | I | own / mean other |
|---|---|---|---|---|---|---|---|
| Extraversion | amplifier | +24 | -12 | -19 | -11 | -11 | 1.8x |
| Extraversion | suppressor | -26 | -4 | -29 | -20 | -20 | 1.4x |
| Agreeableness | amplifier | +8 | +54 | -22 | -2 | -18 | 4.2x |
| Agreeableness | suppressor | -5 | -30 | -8 | -7 | -8 | 4.4x |
| Conscientiousness | amplifier | -7 | -9 | -0 | -4 | -7 | 0.1x |
| Conscientiousness | suppressor | +1 | -1 | -45 | -25 | -19 | 3.9x |
| EmotionalStability | amplifier | -15 | +10 | -16 | +7 | -15 | 0.5x |
| EmotionalStability | suppressor | -9 | -11 | -18 | -25 | -10 | 2.1x |
| Intellect | amplifier | -6 | -13 | -16 | -11 | +27 | 2.3x |
| Intellect | suppressor | -5 | -11 | -26 | -8 | -29 | 2.3x |
Own trait moves most and in the right direction for 7 of 10. The Conscientiousness amplifier moves its own trait -0.4, i.e. not at all, which was read on 2026-08-30 as a ceiling effect: the base already scores 5.60 of 7. Section D contradicts that reading - a single adapter trained on the Conscientiousness factor reaches +35.1 from the same base on the same prompts. The ceiling is real but it is not the whole cause; averaging ten marker adjectives is.
C. The same adapters as full OCT personas
Condition persona in judged_100.json: the released artefact, stage one plus 0.25 stage two (Full OCT persona replication). Same base, same prompts.
| dial | pole | E | A | C | ES | I | own / mean other |
|---|---|---|---|---|---|---|---|
| Extraversion | amplifier | +30 | -18 | -19 | -12 | -10 | 2.0x |
| Extraversion | suppressor | -33 | -1 | -32 | -23 | -21 | 1.7x |
| Agreeableness | amplifier | +9 | +59 | -25 | -4 | -18 | 4.3x |
| Agreeableness | suppressor | -7 | -38 | -8 | -8 | -8 | 4.9x |
| Conscientiousness | amplifier | -9 | -15 | -1 | -2 | -7 | 0.1x |
| Conscientiousness | suppressor | -3 | -5 | -58 | -35 | -22 | 3.6x |
| EmotionalStability | amplifier | -21 | +5 | -21 | +3 | -17 | 0.2x |
| EmotionalStability | suppressor | -14 | -17 | -25 | -32 | -13 | 1.9x |
| Intellect | amplifier | -10 | -14 | -22 | -14 | +43 | 2.8x |
| Intellect | suppressor | -7 | -12 | -30 | -7 | -37 | 2.6x |
Own trait moves most and in the right direction for 8 of 10. Against stage one alone the persona's own-scale shift is larger for 9 of the 10 dials; stage two adds behavioural amplitude even though it adds almost no trait geometry (Structure of the stage-two adapter space).
D. Dedicated Big Five factor adapters
The three arms above all build a factor dial out of adjectives. These ten adapters ARE the
factor: one stage-one DPO adapter per OCEAN pole, trained on the zoo's shared prompt pool at
the matched objective from Persona Cartography's own Figure 2 constitutions
(Big Five factor adapters (Persona Cartography's own ten dials)). Stage one only - no OCT stage two. Mapping onto the zoo's five
factors, whose fifth is Emotional Stability rather than Neuroticism: Extraversion = bf_extraversion_high / bf_extraversion_low; Agreeableness = bf_agreeableness_high / bf_agreeableness_low; Conscientiousness = bf_conscientiousness_high / bf_conscientiousness_low; EmotionalStability = bf_neuroticism_low / bf_neuroticism_high; Intellect = bf_openness_high / bf_openness_low
(spider.json#sources.bigfive_mapping). Condition stage1 in judged_bigfive.json; base is
every base record of that same eval (E 4.12, A 4.68,
C 5.59, ES 4.59, I 5.57,
spider.json#base_bigfive).
| dial | pole | E | A | C | ES | I | own / mean other |
|---|---|---|---|---|---|---|---|
| Extraversion | amplifier | +64 | +39 | -35 | -12 | -28 | 2.3x |
| Extraversion | suppressor | -21 | -4 | -9 | +27 | -6 | 1.8x |
| Agreeableness | amplifier | +7 | +60 | -26 | +19 | -27 | 3.0x |
| Agreeableness | suppressor | -11 | -36 | -10 | +1 | -10 | 4.5x |
| Conscientiousness | amplifier | -4 | -5 | +35 | +8 | -8 | 5.7x |
| Conscientiousness | suppressor | -20 | -7 | -48 | -33 | -27 | 2.2x |
| EmotionalStability | amplifier | -7 | -8 | -6 | +10 | -15 | 1.1x |
| EmotionalStability | suppressor | -12 | -0 | -22 | -26 | -16 | 2.0x |
| Intellect | amplifier | -7 | -6 | -35 | -13 | +68 | 4.5x |
| Intellect | suppressor | -4 | -16 | -13 | -5 | -23 | 2.5x |
Own trait moves most and in the right direction for 8 of 10. Each cell here is one adapter judged on 24 generations, where arm B averages ten adapters, so these are noisier per cell and cleaner in construction.
Reading
The headline of the original replicates on a different base model, with dials built three different ways and without steering at all. Where it breaks is ceiling: the base model sits high on Conscientiousness and Intellect, so amplifiers there have nowhere to go while their suppressors work. Suppressors are dirtier than amplifiers, dragging Conscientiousness and Intellect down together (the competence bundle), which the original figure also hints at. Own-versus-other selectivity is in the last column of each table (own / mean other).
Regenerate: qwen35/build_spider_data.py then wiki/tools/gen_spider_page.py.
Sources
qwen35/analysis/spider.jsonqwen35/build_spider_data.pyqwen35/phase10_runs/judged_100.jsonqwen35/phase10_runs/judged_steerfix.jsonqwen35/phase10_runs/judged_bigfive.jsonqwen35/traits_bigfive.jsonqwen35/traits_primary.jsonqwen35/spider_page/index.html
Linked from
- Big Five factor adapters (Persona Cartography's own ten dials)
- Inventory of built HTML pages
- Inspect personality evaluations (BFI and TRAIT)
- Stage two, explored
File
pages/behaviour/ocean-dials-replication.md