The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe
Anthropic’s SDKs Just Moved Managed Agents From Preview Docs Toward Real Client Surface Area
claude-code

Anthropic’s SDKs Just Moved Managed Agents From Preview Docs Toward Real Client Surface Area

Anthropic’s latest SDK releases are not trying to win the launch thread. That is probably the point. TypeScript v0.102.0 and Python v0.107.0, both published on June 6, move Managed Agents a little further out of “interesting beta documentation” and into the place production teams actually
06 Jun 2026 4 min read
xAI’s Colossus Buildout Is Starting to Look Like a Neocloud With a Chatbot Attached
xai

xAI’s Colossus Buildout Is Starting to Look Like a Neocloud With a Chatbot Attached

Colossus was supposed to be the engine behind Grok. The more interesting version of the story, at least this week, is that it may be becoming something less glamorous and more valuable: scarce AI infrastructure rented by the month to the companies Grok is trying to beat. TechCrunch reports that
06 Jun 2026 5 min read
Mastra 1.41 Turns Sandboxes Into a Multi-Tenant Runtime Decision
ai-frameworks

Mastra 1.41 Turns Sandboxes Into a Multi-Tenant Runtime Decision

Mastra 1.41 is the kind of release that looks boring until you have actually tried to ship an agent product with more than one customer. Then it looks less like plumbing and more like the boundary between a demo and an incident report. The headline change in @mastra/core
06 Jun 2026 6 min read
OPRD Makes Distillation Cheaper by Supervising the Reasoning State Before the Logits
ai-models

OPRD Makes Distillation Cheaper by Supervising the Reasoning State Before the Logits

OPRD is a distillation paper with a very simple engineering smell: the teacher already computed a rich internal representation, and the training loop throws it away so the student can imitate a probability distribution over a 150,000-token vocabulary. That is not elegance. That is paying for the whole review
06 Jun 2026 5 min read
RL Translation Paper Says Models Need to Learn How to Use the Grammar Book, Not Memorize the Language
ai-models

RL Translation Paper Says Models Need to Learn How to Use the Grammar Book, Not Memorize the Language

The most interesting part of this low-resource translation paper is not that reinforcement learning beats supervised fine-tuning in one table. It is the shape of the failure it exposes: supervised fine-tuning can make a model look better on the languages it has already seen while making it worse at the
06 Jun 2026 5 min read
Fix with Copilot Turns Failed GitHub Actions Into Agent Work. Useful, If You Keep the Blast Radius Small.
codex

Fix with Copilot Turns Failed GitHub Actions Into Agent Work. Useful, If You Keep the Blast Radius Small.

“Fix with Copilot” for failing GitHub Actions is a genuinely good agent use case, which is precisely why it deserves more scrutiny than the average AI button. CI failures are narrow, evidence-rich, and measurable. They are also where teams often confuse “green” with “correct.” GitHub has put an agent directly
06 Jun 2026 5 min read
GitHub Just Made Copilot Plugins an Enterprise Policy Surface in VS Code
codex

GitHub Just Made Copilot Plugins an Enterprise Policy Surface in VS Code

GitHub’s enterprise-managed plugins preview for VS Code is the sort of changelog item that reads like admin plumbing until you picture the incident review. “Who installed the agent plugin that could see the deployment logs?” “Which MCP server did Copilot call?” “Why did every developer suddenly have a hook
06 Jun 2026 5 min read
Codex 0.138 Alpha Is Where Remote Control, Token Usage, and Permission Policy Stop Being Plumbing
codex

Codex 0.138 Alpha Is Where Remote Control, Token Usage, and Permission Policy Stop Being Plumbing

Codex 0.138.0-alpha.6 looks like a small alpha tag if you only read the release page. That would be the wrong review. The interesting work is not in a glossy changelog; it is in the plumbing OpenAI is turning into policy: remote-control status, app-server account sessions, token usage,
06 Jun 2026 5 min read
Microsoft Shows the CI/CD Agent Problem: Secrets, Tools, and Untrusted Prose in One Runner
claude-code

Microsoft Shows the CI/CD Agent Problem: Secrets, Tools, and Untrusted Prose in One Runner

Microsoft’s Claude Code GitHub Action case study is the clearest current reminder that “AI in CI” is not a chatbot feature. It is an LLM-driven process with tools running inside your supply chain. On June 5, Microsoft Threat Intelligence published research showing how Claude Code GitHub Action could expose
06 Jun 2026 5 min read
Claude Code Action 1.0.139 Keeps CI Agents Current While the Security Story Gets Loud
claude-code

Claude Code Action 1.0.139 Keeps CI Agents Current While the Security Story Gets Loud

The changelog for claude-code-action v1.0.139 is almost aggressively boring: bump Claude Code to 2.1.167 and Agent SDK to 0.3.167. That is exactly why teams should pay attention. Claude Code Action is not a decorative wrapper around a chatbot. It is a GitHub Actions participant
06 Jun 2026 5 min read
Claude Code 2.1.166 Adds the Fallbacks and Guardrails Teams Actually Need
claude-code

Claude Code 2.1.166 Adds the Fallbacks and Guardrails Teams Actually Need

Claude Code’s most important release this week is not the one that gives developers a new trick. It is the one that admits the runtime now has production responsibilities. Anthropic shipped Claude Code v2.1.166 on June 6, followed 38 minutes later by v2.1.167, a generic
06 Jun 2026 4 min read
AgentScope 2.0.1 Turns Alibaba’s Agent Framework Toward Teams, RAG, and Provider Reality
qwen

AgentScope 2.0.1 Turns Alibaba’s Agent Framework Toward Teams, RAG, and Provider Reality

AgentScope 2.0.1 is a patch release with a strategic tell: Alibaba’s agent framework is moving from “build an agent” toward “run a team of agents without losing the plot.” The headline in the release notes is short — “Agent Team is supported” — but the surrounding changes are more
05 Jun 2026 6 min read
← Newer Posts Page 43 of 136 Older Posts →
The LGTM © 2026
  • Sign up
Powered by Ghost