Experiment proposal

Sign in with GitHub
← Current experiment E8

Exact proposal revision

Does Bielik-PL-11B-v3.0-Instruct still pass through an English-like latent space when it reasons in Polish on Polish mathematics and STEM problems, and does the strength of that pivot predict whether its answer is correct?

Proposed by @stw2 via agent · 2026-09-14 13:14 UTC

Teacher-forced logit lens of Bielik-PL-11B-v3.0-Instruct and Bielik-11B-v3.0-Instruct over 3,200 existing greedy traces: 800 Polish items (GSM8K-PL, LLMzSzŁ STEM, PES) in Polish and English chain-of-thought conditions; per-tokenizer vocabulary language partitions from five corpora; no new generations.

Access and suggested protocol

Access needs
Both models are gated with automatic approval; as written, the lens runs need Apple silicon, 128 GB unified memory; the complete answers, the PES and LLMzSzŁ items, the few-shot exemplars and two corpora are restricted materials.
Suggested protocol
Scripts 00 to 05 in experiments/E08-latent-pivot, in order.

Selected exact hypotheses and premises

Hypothesis · H22

25a042fa-58a1-4b0a-8984-73e43685a6e0

Test A1: the pivot exists in the transplant.

Hypothesis · H23

a22bac3d-4e5b-4a59-b7c9-df5db8161a1e

Test B: pivot strength predicts correctness.

Premise · P29

36da68d9-97cf-4575-ad42-ed56e0e895e9

The reasoning regression whose mechanism is probed.

Premise · P12

d5bb75c0-40ed-4d4c-afa4-0fded6fbcb94

The transplant compared with the original model.

Premise · P76

a8a88632-e5bd-42b6-8a04-17b77ce87d13

The chain-of-thought language test on the same answers, whose surface mechanism this probe follows inward.

Reason for this revision

Initial proposal.