The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe
OpenClaw Is Closing the Policy Gap Between Who Asked and What Tool Runs
openclaw

OpenClaw Is Closing the Policy Gap Between Who Asked and What Tool Runs

Approval gates have a branding problem: they look like security controls, but too often they are just modal dialogs with incomplete context. OpenClaw PR #96211 is interesting because it goes after the field most agent platforms quietly hand-wave: requester origin. Not just “what tool is about to run,” but who
24 Jun 2026 4 min read
Microsoft’s Water Numbers Show the Hidden Constraint Behind Azure AI Scale
azure-ai

Microsoft’s Water Numbers Show the Hidden Constraint Behind Azure AI Scale

Microsoft’s latest water-stewardship post is easy to misfile as sustainability communications. Do not. It is an Azure AI infrastructure story with fewer model names and more physics. The numbers are the hook: Microsoft says average datacenter water use effectiveness improved by nearly 90% since early datacenter generations, falling from
24 Jun 2026 4 min read
Copilot CLI GA Makes the Terminal an Agent Workbench, Not Just a Prompt Box
azure-ai

Copilot CLI GA Makes the Terminal an Agent Workbench, Not Just a Prompt Box

GitHub’s new Copilot CLI terminal interface is generally available, and the least interesting thing about it is that it has tabs. The interesting part is what those tabs imply: the terminal is becoming an agent workbench where GitHub context, MCP servers, skills, plugins, hooks, custom agents, approvals, and repo
24 Jun 2026 4 min read
Copilot BYOK Turns GitHub’s App Into a Model-Routing Front End for Azure OpenAI and Foundry
azure-ai

Copilot BYOK Turns GitHub’s App Into a Model-Routing Front End for Azure OpenAI and Foundry

Bring-your-own-key support in the GitHub Copilot app sounds like a settings-page feature. It is not. It is GitHub quietly turning Copilot’s standalone app into a front end for whatever inference estate an engineering organization has already approved: Azure OpenAI, Microsoft Foundry, Anthropic, local models, private OpenAI-compatible gateways, and the
24 Jun 2026 4 min read
Langfuse 3.196 Treats Agent Observability Like Production Infrastructure, Not a Dashboard
ai-frameworks

Langfuse 3.196 Treats Agent Observability Like Production Infrastructure, Not a Dashboard

Langfuse 3.196 is not one big feature. It is a pile of boundary fixes. That is the story. The release, published June 24 at 2026-06-24T08:13:37Z, continues Langfuse’s evolution from “LLM observability dashboard” into production data infrastructure for agent systems. The difference matters. Dashboards can tolerate rough
24 Jun 2026 5 min read
OpenAI Agents SDK 0.17.7 Fixes the Runtime Edges Multi-Agent Apps Actually Trip Over
ai-frameworks

OpenAI Agents SDK 0.17.7 Fixes the Runtime Edges Multi-Agent Apps Actually Trip Over

The interesting part of OpenAI Agents SDK 0.17.7 is not that it adds some shiny new agent capability. It does not. The interesting part is that it fixes the boring failure modes that show up after the demo architecture diagram meets streaming providers, Realtime handoffs, SQLite sessions, guardrails,
24 Jun 2026 5 min read
Qwen Code’s Nightly Adds Hard Stops Where Auto Mode Used to Ask the Model Nicely
ai-frameworks

Qwen Code’s Nightly Adds Hard Stops Where Auto Mode Used to Ask the Model Nicely

Auto mode is where coding-agent safety stops being a slide and starts being a file-system problem. Qwen Code’s June 24 nightly is worth reading less as a feature release than as a small constitutional amendment: some commands are too destructive to leave to a model classifier, even a good
24 Jun 2026 5 min read
DeepMind’s Pelé Reconstruction Is a Consent Test for Generative Memory
google-ai

DeepMind’s Pelé Reconstruction Is a Consent Test for Generative Memory

The safest way to read Google DeepMind’s Pelé reconstruction is not as a sports clip. It is a governance test with a football attached. Google took a legendary moment that was never filmed — Pelé’s August 2, 1959 “Gol da Rua Javari,” remembered for three consecutive sombreros without the
24 Jun 2026 5 min read
Google’s Disaster AI Stack Is Becoming Public Infrastructure, Not Just Model Research
google-ai

Google’s Disaster AI Stack Is Becoming Public Infrastructure, Not Just Model Research

Google’s disaster-AI story is easy to file under corporate-good-citizen marketing. That would be the lazy review. The more interesting read is that Google is quietly assembling something closer to public infrastructure: forecasting models, geospatial datasets, open-source hydrology tools, Search and Maps distribution, Android alerts, and relationships with the agencies
24 Jun 2026 5 min read
Bayesian Control for Coding Agents Says Always Run the Tests Is Not a Strategy
ai-models

Bayesian Control for Coding Agents Says Always Run the Tests Is Not a Strategy

“Always run the tests” is good advice for humans and a bad orchestration policy for agents. The human version assumes judgment: run the cheap checks while you work, pay for the expensive suite when the evidence justifies it, stop when another run will not teach you anything. Coding agents often
24 Jun 2026 4 min read
NatureBench Puts a Ceiling on the AI Scientist Narrative
ai-models

NatureBench Puts a Ceiling on the AI Scientist Narrative

NatureBench is the sort of benchmark the AI-scientist narrative needed: not because it declares the dream dead, but because it makes the claim harder to blur. Producing plausible research code is one thing. Matching or beating the published state of the art from Nature-family scientific papers, without being handed the
24 Jun 2026 4 min read
Qwen-AgentWorld Turns World Models Into a Coding-Agent Runtime Primitive
ai-models

Qwen-AgentWorld Turns World Models Into a Coding-Agent Runtime Primitive

Qwen-AgentWorld is not another “our model writes better code” release. That would be easier to file and less interesting. The sharper claim is that coding agents need a model of the world they are acting inside — terminals, MCP servers, browsers, Android UIs, operating systems, repositories — because an agent that cannot
24 Jun 2026 4 min read
← Newer Posts Page 4 of 136 Older Posts →
The LGTM © 2026
  • Sign up
Powered by Ghost