llama.cpp now supports Ling 3.0 reasoning models, including the 8b1b tiny and 124b5b flash versions. Despite their naming, these models are architected specifically for reasoning tasks.
HOW THIS AFFECTS YOU
●
builderYou can now run these specialized reasoning models locally using optimized llama.cpp kernels.