●builderYou can reduce inference costs by using an SLM to query a larger model only when specific information is needed.
●researcherThe three-stage RLVR framework provides a new way to optimize token-efficient collaboration between heterogeneous models.