The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe
Codex Starts Treating Plugins Like a Supply Chain
codex

Codex Starts Treating Plugins Like a Supply Chain

Plugins are where coding agents stop being assistants and start becoming platforms. A plugin can bring skills, app integrations, MCP server configuration, lifecycle hooks, and marketplace metadata into the same runtime that reads your repository, proposes edits, runs commands, and sometimes asks for fewer approvals than a junior engineer would.
24 Jun 2026 5 min read
Claude Code’s Remote Auto-Compaction Bug Turns Context Management Into Hidden Runtime State
claude-code

Claude Code’s Remote Auto-Compaction Bug Turns Context Management Into Hidden Runtime State

Context compaction is the garbage collector for coding-agent state. Nobody wants to think about it until it silently stops running. That is why Claude Code issue #70477 is more interesting than a routine token-cost complaint. The report says remote or bridge-attached Claude Code sessions can skip threshold auto-compaction and ignore
24 Jun 2026 4 min read
Claude Code Dynamic Workflows Need Token Admission Control, Not a Concurrency Dial
claude-code

Claude Code Dynamic Workflows Need Token Admission Control, Not a Concurrency Dial

Claude Code’s Dynamic Workflows problem is not that users can start too many agents. It is that “too many” is being measured in the wrong unit. A concurrency cap sounds like a control. It is not a budget. Fifteen workers that each grep one small file are cheap. Fifteen
24 Jun 2026 4 min read
Claude Code’s Prompt-Injection False Memory Bug Is How Safety Heuristics Become Superstitions
claude-code

Claude Code’s Prompt-Injection False Memory Bug Is How Safety Heuristics Become Superstitions

Claude Code’s latest prompt-injection report is not alarming because the model saw hostile text. It is alarming because, according to the reporter, the model invented the hostile text, diagnosed it as a prompt injection, and then wrote that diagnosis into memory where future sessions could inherit it. That is
24 Jun 2026 4 min read
DFlash Turns Speculative Decoding Into a Blackwell Utilization Story
nvidia

DFlash Turns Speculative Decoding Into a Blackwell Utilization Story

Speculative decoding is one of those ideas that sounds like a serving hack until you remember the original problem is stranger: we run massively parallel accelerators and then ask large language models to emit text one token at a time. DFlash, the block-diffusion speculative decoding method NVIDIA is now highlighting
23 Jun 2026 4 min read
Tokens per Watt Is Becoming the AI Factory’s Real P&L Metric
nvidia

Tokens per Watt Is Becoming the AI Factory’s Real P&L Metric

“Tokens per watt” sounds like a facilities metric until you have to run a real AI product under real capacity limits. Then it becomes the P&L. NVIDIA’s latest AI factory energy-efficiency post is easy to misread as green-computing content. It is really a margin story: every watt
23 Jun 2026 4 min read
NVIDIA OpenShell Is the Interesting Part of the Agent Toolkit Story
nvidia

NVIDIA OpenShell Is the Interesting Part of the Agent Toolkit Story

The agent market has spent too much time pretending the hard problem is choosing the smartest model. It is not. The hard problem is giving an agent enough authority to be useful without giving it enough authority to become a root-cause analysis document. NVIDIA’s Agent Toolkit announcement is interesting
23 Jun 2026 4 min read
NVIDIA and AWS Are Making GPU Retrieval Boring — Which Is Exactly the Point
nvidia

NVIDIA and AWS Are Making GPU Retrieval Boring — Which Is Exactly the Point

GPU announcements usually arrive dressed as hardware news. This one is more useful if you read it as a retrieval infrastructure story. NVIDIA and AWS are not merely adding another accelerated EC2 shape; they are pushing GPU acceleration into the boring middle of production AI, where vector indexes get rebuilt,
23 Jun 2026 4 min read
n8n 2.28 Turns MCP and Instance AI Into an Operations Surface, Not a Demo Panel
ai-frameworks

n8n 2.28 Turns MCP and Instance AI Into an Operations Surface, Not a Demo Panel

n8n 2.28.0 is the kind of release that looks impossible to cover if you read it as a changelog. The compare from [email protected] to [email protected] spans hundreds of commits and hundreds of changed files. But the useful AI story is much narrower:
23 Jun 2026 6 min read
Agno 2.6.19 Makes Agent Runs Rewindable, Forkable, and Composable
ai-frameworks

Agno 2.6.19 Makes Agent Runs Rewindable, Forkable, and Composable

Agno’s latest release is not chasing the usual agent-framework dopamine hit: a prettier demo, another provider integration, or a benchmark graph with more diagonal confidence than substance. Agno v2.6.19 is about something more useful and less marketable: what happens when an agent run fails halfway through doing
23 Jun 2026 5 min read
Multimodal Models Are Learning When to Reach for Code, Not Just How to Look Harder
ai-models

Multimodal Models Are Learning When to Reach for Code, Not Just How to Look Harder

Multimodal models have spent years getting better at looking. AIR is interesting because it asks them to get better at deciding when looking is no longer the hard part. The paper, Adaptive Interleaved Reasoning with Code in MLLMs, trains multimodal large language models to use code during complex numerical visual
23 Jun 2026 3 min read
MoE Routing May Be the Confidence Signal Text Voting Was Pretending to Be
ai-models

MoE Routing May Be the Confidence Signal Text Voting Was Pretending to Be

Majority voting works best on answers that behave like answers: short, canonical, easy to normalize, and preferably surrounded by a box. Code does not behave that way. Neither do patches, tool trajectories, debugging plans, or most useful agent outputs. That is why Does the Same Token Mean the Same State?
23 Jun 2026 3 min read
← Newer Posts Page 5 of 136 Older Posts →
The LGTM © 2026
  • Sign up
Powered by Ghost