Insights
Recursive Self Improvement for Coding Agents
See how recursive self-improvement helped Cline optimize Kimi K3 and achieve an 88.8% SOTA score on Terminal-Bench 2.1 at a fraction of the cost.
Insights
See how recursive self-improvement helped Cline optimize Kimi K3 and achieve an 88.8% SOTA score on Terminal-Bench 2.1 at a fraction of the cost.
Guides
A practical guide to the economics of self-hosting open-weight LLMs. Using Kimi K2.6 and real production traffic from Cline, this post breaks down GPU memory, inference, batching, pricing, and the point at which self-hosting can save millions.
Announcements
We didn't have benchmark numbers, so over a weekend we ran Cline against 89 coding tasks, diagnosed every failure, and shipped fixes that took our score from 47% to 57%. Here's the hill climbing process so you can do it too.
In building AI agents at Cline, we've discovered that the most dangerous ideas aren't the obviously bad ones, they're the seductive ones that sound brilliant in theory but fail in practice.