●builderYou can improve small VLM performance by offloading specific reasoning or evidence-gathering tasks to an external harness instead of fine-tuning weights.
●researcherThis method provides a framework for analyzing the optimal allocation of runtime decisions between model parameters and external tools.