The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe
OpenCode 1.17.9 Makes Agent Limits Fail Like a Product, Not a Stack Trace
agentic-coding

OpenCode 1.17.9 Makes Agent Limits Fail Like a Product, Not a Stack Trace

OpenCode 1.17.9 is a reminder that agent quality is not only about how many steps the model can take. It is about what happens when the agent must stop, when the provider metadata is weird, when prompt caching is fragile, and when the client is newer than the
21 Jun 2026 4 min read
Qwen Code’s June 21 Nightly Is a Security Checklist Disguised as Bug Fixes
agentic-coding

Qwen Code’s June 21 Nightly Is a Security Checklist Disguised as Bug Fixes

Qwen Code’s June 21 nightly is not the kind of release that gets a launch thread. Good. The release reads like a security checklist disguised as bug fixes: workspace boundaries, plan-mode consent, strict environment parsing, ACP glob limits, custom theme path validation, socket-port validation, and endpoint detection. That is
21 Jun 2026 4 min read
Codex Alpha 9 Removes Token-Budget Noise Before It Becomes Agent Spam
agentic-coding

Codex Alpha 9 Removes Token-Budget Noise Before It Becomes Agent Spam

Token-budget telemetry has a habit of sounding responsible while quietly making the product worse. Codex 0.142.0-alpha.9 is a small alpha release by commit count, but it makes the right call on one of those deceptively annoying agent-runtime details: it removes automatic remaining-token messages at the 25%, 50%
21 Jun 2026 4 min read
llm-rankings

OpenRouter Is Saying the Quiet Part Out Loud: Cheap Long-Context Models Are Winning Production

The cleanest leaderboard story this week is the one that did not move. Arena AI’s Text board is still an Anthropic wall: Claude Fable 5 at 1508 Elo, Claude Opus 4.6 Thinking at 1504, Claude Opus 4.7 Thinking at 1502, then more Claude. Arena Code/WebDev is
21 Jun 2026 5 min read
Claude Code Memory That Loads but Does Not Steer Tools Is Just Expensive Wallpaper
claude-code

Claude Code Memory That Loads but Does Not Steer Tools Is Just Expensive Wallpaper

Agent memory has a marketing problem because the word “memory” sounds stronger than the mechanism usually is. Developers hear “remember this” and mentally file it under configuration. Most agent systems file it under context. That gap is where a lot of disappointment lives. A new Claude Code issue captures the
21 Jun 2026 4 min read
Claude Desktop’s Orphaned Scheduled Tasks Are an Agent Control-Plane Bug, Not a Fan-Noise Bug
claude-code

Claude Desktop’s Orphaned Scheduled Tasks Are an Agent Control-Plane Bug, Not a Fan-Noise Bug

The easiest way to underestimate desktop agents is to watch the UI instead of the process table. A scheduled task looks harmless when it is a little item in a sidebar. It looks different when the machine wakes up with more than a thousand processes, a load average north of
21 Jun 2026 3 min read
Claude Code’s Self-Prompt-Injection Report Shows Role Boundaries Are a Runtime Feature
claude-code

Claude Code’s Self-Prompt-Injection Report Shows Role Boundaries Are a Runtime Feature

Prompt injection usually gets framed as something the outside world does to an agent: a malicious README, a poisoned web page, a hostile GitHub issue, a tool description with a little policy grenade hidden inside. The latest Claude Code role-boundary report is more uncomfortable because the alleged poison pill comes
21 Jun 2026 4 min read
Claude Code’s Latest Data-Loss Report Is Why Agent Git Needs Seatbelts, Not Vibes
claude-code

Claude Code’s Latest Data-Loss Report Is Why Agent Git Needs Seatbelts, Not Vibes

The scary part of the latest Claude Code data-loss report is not that an agent allegedly made a bad Git decision. Humans do that every week, usually five minutes before lunch. The scary part is that the reported failure sits exactly where coding-agent products keep trying to sell convenience: “let
21 Jun 2026 4 min read
Microsoft Foundry’s APIM Pattern Makes the AI Gateway Boring Enough for Production
azure-ai

Microsoft Foundry’s APIM Pattern Makes the AI Gateway Boring Enough for Production

Enterprise AI architecture keeps rediscovering an old truth: the hard part is not calling the model. The hard part is calling the right model, from the wrong region, under the right identity, without punching a hole through your network model or your compliance story. Microsoft’s new guidance on cross-region
20 Jun 2026 5 min read
Codex Starts Giving Context Windows Provenance, Not Just Token Counts
codex

Codex Starts Giving Context Windows Provenance, Not Just Token Counts

Token counters are table stakes now. The more interesting question is whether a coding agent can explain how it arrived at the context it is using after compaction, rollback, resume, and a few long-running tool sessions have rearranged the furniture. That is the useful read on OpenAI’s latest Codex
20 Jun 2026 5 min read
MCP’s write_file Escaping Problem Is Really a Tool-Call Serialization Tax
claude-code

MCP’s write_file Escaping Problem Is Really a Tool-Call Serialization Tax

Base64 is ugly. Retry loops caused by broken JSON tool calls are uglier. A fresh issue in the Model Context Protocol servers repository proposes adding an optional content_base64 parameter to write_file, plus a matching newText_base64 path for edit operations. The request is narrow: let agents send file
20 Jun 2026 4 min read
Claude Code Routines Need Output-Branch Contracts, Not Prompted Git Etiquette
claude-code

Claude Code Routines Need Output-Branch Contracts, Not Prompted Git Etiquette

Scheduled coding agents are not chat sessions with alarms. They are job runners. Job runners need output contracts. That is the uncomfortable lesson in a fresh Claude Code issue reporting that scheduled cloud Routines can finish work on a harness-assigned branch such as claude/stoic-fermi-azzeU instead of the canonical branch
20 Jun 2026 4 min read
← Newer Posts Page 11 of 136 Older Posts →
The LGTM © 2026
  • Sign up
Powered by Ghost