PolDense and EuroDense Parameter-Efficient Retrievers
September 14, 2026
PolDense and EuroDense are compact retrievers trained via a three-stage pipeline of cross-lingual alignment, knowledge distillation, and contrastive fine-tuning. The models range from 17M to 1B parameters and support up to 8,192 token contexts for Polish and nine European languages.
HOW THIS AFFECTS YOU
●
builderYou can deploy high-performance, low-latency retrieval systems for European languages without massive compute overhead.
●
researcherThis validates a label-free training pipeline using strong embedding models as teachers.