Kilo Code Ships JetBrains Plugin 7.1.9-rc.1 With Auto-Recovery for Context-Limit Errors
Kilo Code released version 7.1.9-rc.1 of its JetBrains plugin, a pre-release patch that adds automatic conversation compacting when providers hit context limits, renders agent board messages as Markdown, and unifies goal scheduling for long-running agent tasks.
Kilo Code published a release candidate for its JetBrains IDE plugin, version 7.1.9-rc.1, on September 28. The pre-release patch focuses on reliability fixes for AI agent sessions and several interface refinements, according to the release notes on GitHub.
What changed
The update's most notable fix addresses a recurring failure mode in long agent sessions: when a provider returns a context-limit error, the plugin now automatically compacts the conversation to recover instead of failing the session outright. This targets a common pain point for developers running extended agentic coding workflows where accumulated context exceeds a model's window.
Other fixes in this release include:
- Prompt cache breakpoints are now preserved only for first-party OpenAI providers, preventing the plugin from sending incompatible cache-breakpoint requests to custom or third-party OpenAI-compatible endpoints.
- Scheduled sessions now report their actual wake time rather than showing a generic "busy" status, giving developers clearer visibility into when a paused agent will resume.
- Embedding dimension configuration is now correctly passed to OpenAI-compatible indexing providers, fixing a mismatch that could affect codebase indexing accuracy.
On the feature side, Kilo Code renamed "workflows" to "commands" throughout the JetBrains plugin interface and reworked the associated settings presentation. Agent board messages now render as Markdown rather than plain text, improving readability of formatted agent output. The release also unifies goal-setting with the scheduling and timing tools, letting agents coordinate long-running, multi-step work more reliably across sessions.
Release status
This is a release candidate (rc.1), meaning it is a pre-release build intended for testing ahead of a stable rollout. It was built automatically via GitHub Actions from commit 2ccf28e, three commits ahead of the main branch at time of publication.
What this means
This is an incremental maintenance release, not a new model or major feature launch. The changes reflect a pattern common across AI coding assistants in 2025: as agents run longer and more autonomous sessions, failure handling around context limits, caching, and scheduling becomes the primary engineering burden rather than raw model capability. Automatic conversation compacting on context-limit errors is a practical fix that reduces friction for developers relying on Kilo Code for multi-step, unattended coding tasks inside JetBrains IDEs. Teams already using the JetBrains plugin should expect this rc to promote to stable shortly if no regressions surface during testing.
Related Articles
Anthropic Python SDK 1.9.0 Adds Reference to Unreleased 'claude-sonnet-5-5' Model ID
Anthropic's anthropic-sdk-python v1.9.0 release adds a reference to an unannounced 'claude-sonnet-5-5' model ID, a new between_tools thinking type, and the ability to run tool calls while a reply streams. No pricing, context window, or benchmark data for the model has been disclosed.
Vercel AI SDK Patch Adds Support for 'None' Reasoning Effort on GPT-6 Sol and Luna
Vercel released version 4.0.78 of @ai-sdk/openai, a patch that adds support for setting reasoningEffort to 'none' for GPT-6 Sol and Luna models. The update also introduces validation logic that warns on unsupported request-level effort updates and rejects invalid historical ones.
Google Adds Real-Time Video Avatars to Gemini 3.8 Live for Enterprise Voice Agents
Google has added Live Avatar to its Gemini 3.8 Live dialogue models, pairing real-time video personas with voice AI for enterprise customer service and sales use cases. The feature is restricted to Gemini Enterprise customers via allowlisting and includes SynthID watermarking on all generated video.
Cline v4.1.21 Fixes Local Model Timeouts, Expands Catalog to 6,386 Models Across 209 Providers
Cline's v4.1.21 update adds a new OpenAI-compatible provider called ai&, refreshes its model catalog to 6,386 models across 209 providers, and fixes a bug that ended tasks prematurely when local models hit output-token limits. Eleven unpinned providers, including GitHub Copilot and Vertex, now default to Claude Opus 5.5.
Comments
Loading...