Cited claim

Sign in with GitHub
← Publications

Cited claim · P42 · Author-curated · Disputed

On Polish EQ-Bench, the Bielik v3 PL models surpass the performance of their original-tokenizer counterparts.

Published by @stw2 · 2026-09-14 · Sources, measurements and interpretation are supplied by the author.

Notices · Disputed · this exact version stays citable

Structured assertion

Relation: Exceeds

subject
Bielik v3 PL modelsModels concerned.
comparator
original-tokenizer counterpartsModels compared with.
benchmark
Polish EQ-BenchBenchmark concerned.

Paper citations

arXiv:2604.10799v1 →Revision supplied by author
  1. Section 7, p. 15

    Evaluation across nine Polish and multilingual benchmarks (Section 5) confirms that the Bielik v3 PL models closely preserve - and on CPTUB and Polish EQ-Bench even surpass - the performance of their original-tokenizer counterparts, while English-language capabilities remain largely intact.

Concept definitions

Reuse the defining version and key when the meaning fits your assertion.

Polish EQ-Bench

Polish adaptation of the EQ-Bench emotional intelligence benchmark.

Key polish_eq_bench · version 405a9609-26e5-4200-8682-56616dfe42e1

Bielik v3 PL models

The 11B and 7B Bielik v3 models with the APT4 tokenizer.

Key bielik_v3_pl_models · version 3da77265-eb29-46f1-90e2-ccc03ac36917