Cited claim

Sign in with GitHub
← Publications

Cited claim · P31 · Author-curated

On Polish EQ-Bench, Bielik-PL-11B-v3.0-Instruct scores 71.15 and Bielik-11B-v3.0-Instruct 71.20.

Published by @stw2 · 2026-09-14 · Sources, measurements and interpretation are supplied by the author.

Structured assertion

Relation: Metric comparison

benchmark
Polish EQ-BenchBenchmark the values are reported on.
language
PolishLanguage of the evaluated tasks.
subject
Bielik-PL-11B-v3.0-InstructModel or tokenizer after the tokenizer change.
subject value
71.15 scoreValue of the metric for the subject.
comparator
Bielik-11B-v3.0-InstructModel or tokenizer before the tokenizer change.
comparator value
71.20 scoreValue of the metric for the comparator.

Paper citations

arXiv:2604.10799v1 →Revision supplied by author
  1. Section 5.2, p. 6

    Bielik-11B-v3.0-Instruct achieves a score of 71.20 on the Polish EQ-Bench (Table 3), demonstrating strong emotional intelligence capabilities.
  2. Section 5.2, p. 6

    The Polish tokenizer variants, Bielik-PL-11B-v3.0-Instruct and Bielik-PL-Minitron-7B-v3.0-Instruct, score 71.15 and 66.89 on the eq-bench_v2_pl run reported in the same table.

Author’s note

Comparator values are the leaderboard comparisons from the Bielik 11B v3 technical report (Ociepa et al. [2025a]); subject values are reported in this paper for the checkpoints with the Polish tokenizer.

Concept definitions

Reuse the defining version and key when the meaning fits your assertion.

Polish EQ-Bench

Polish adaptation of the EQ-Bench emotional intelligence benchmark.

Key polish_eq_bench · version 405a9609-26e5-4200-8682-56616dfe42e1

Metric comparison

The subject and the comparator take subject_value and comparator_value of the metric; benchmark, scope, evaluation_text, language and setting state what the values were measured on, where given. Values may come from separate evaluation runs.

Key metric_comparison · version 2d581473-7212-4ea2-bf80-f0c4b8cb247e

Bielik-11B-v3.0-Instruct

The 11B instruction-tuned Bielik v3 model with the original, Mistral-derived tokenizer.

Key bielik_11b_v3_instruct · version ac435442-b907-442a-9d1c-f951fa41d53c

Bielik-PL-11B-v3.0-Instruct

The 11B instruction-tuned Bielik v3 PL model with the APT4 tokenizer.

Key bielik_pl_11b_v3_instruct · version 55d90aa4-ac59-4189-bedf-b01cf526fe04