Blog
Engineering notes on code agents, context, and the systems around them.
Latest
Improving instruction following with next-step hints in tool call results
Ending a tool result with a local instruction helps smaller models interpret the data and choose the next useful action.
Read articleMore articles
06How we automated MCP OAuth 2.1 browser testing
Testing the MCP browser authorization flow with a real loopback callback, PKCE, and Playwright.
Giving Kimi K3 a faster path through code
Preliminary RepoContextBench results show CodeAlive cutting Kimi Code task time by 31% and answerer token use by 28% with Kimi K3.
CodeAlive 3.0: From Context Engine to Code Research Agent
A look at CodeAlive 3.0: our production context engine, code research agent, RepoContextBench, and the new Tool API and MCP surface.
How CodeAlive MCP evolved: from three tools to a research interface
How CodeAlive MCP changed its tools, retrieval flow, response formats, and error handling on the way to v3.
RepoContextBench: How Qwen 3.6-35B beats Sonnet 4.6
Qwen 3.6-35B A3B scored 77.6 with the CodeAlive agent harness, ahead of Opus 4.8 at 77.45 and Sonnet 4.6 at 71.95.
Canonicalization for small agent LLMs: a practical AX pattern
A scalar string broke a string-array tool call. We fixed it once at the Microsoft Agent Framework invocation boundary.
New articles
Follow the work, not a marketing funnel
One direct email when we publish. No open tracking, click tracking, sales copy, or more than one message a day.