●builderYou should avoid training many small, specialized evaluators from scratch in favor of shared-weight architectures.
●researcherThis highlights the performance trade-offs between domain specialization and model capacity in LLM-as-a-judge setups.