Cline 4.1.21 fixes local-model output-limit handling and reshuffles default models
Cline
Long replies on llama.cpp, Ollama and LM Studio that hit the output-token limit now compact and retry instead of ending the task. The release adds an 'ai&' provider for open-weight models from Japan, refreshes a 6,386-model catalog across 209 providers, and picks up a js-yaml security fix for rule frontmatter parsing.
Why it matters
Default-model changes for 19 unpinned providers silently switch users to Claude Opus 5.5, a cost-relevant surprise.
Importance: 2/5
Behavior fix plus cost-relevant default-model changes
Sources
official
Cline v4.1.21 release notes