Baszta Improves Polish Multi-Label Safety Classification Using Focal and R-Drop
September 25, 2026
Baszta fine-tunes the 124M parameter allegro/herbert-base-cased model for Polish content safety across five categories. Using Focal and R-Drop objectives, it achieves a statistically significant micro F1 lead over Bielik Guard on the Gadzi Język benchmark.
HOW THIS AFFECTS YOU
●
researcherThe study highlights the necessity of macro-F1 over micro-F1 when evaluating models on highly imbalanced datasets.
●
policyYou can use more precise, language-specific safety classifiers to manage content risks in Polish-speaking markets.