The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe
GAM’s Robotics Result Is a Bet That Physical AI Needs Geometry First, Not Bigger Chat Models
ai-models

GAM’s Robotics Result Is a Bet That Physical AI Needs Geometry First, Not Bigger Chat Models

The easiest way to make a robotics paper sound modern is to say “vision-language-action model” and let the acronym do the marketing. Geometric Action Model is more interesting because it pushes in the opposite direction. Its bet is that physical AI needs less caption glamour and more spatial competence: depth,
16 Jun 2026 4 min read
Qwen’s Value Axis Paper Makes Model Confidence Look Less Like a Feeling and More Like a Controllable Feature
ai-models

Qwen’s Value Axis Paper Makes Model Confidence Look Less Like a Feeling and More Like a Controllable Feature

Model confidence is usually treated like weather: visible in the forecast, hard to control, and mostly tolerated until it ruins the deployment. The new Value Axis paper is interesting because it argues confidence is not just an output style or a calibrated probability glued onto the end of generation. In
16 Jun 2026 4 min read
GitHub Copilot App 0.2.33 and 0.2.34 Turn the Desktop App Into an Agent Control Room
codex

GitHub Copilot App 0.2.33 and 0.2.34 Turn the Desktop App Into an Agent Control Room

GitHub’s Copilot app is starting to look less like a chat client and more like the control room for agentic development. That distinction matters. A chat client is where you ask for help. A control room is where you coordinate work, inspect state, watch budgets, manage permissions, recover from
16 Jun 2026 4 min read
GitHub Is Moving JetBrains Copilot Onto Copilot CLI Because the Harness Is Now the Product
codex

GitHub Is Moving JetBrains Copilot Onto Copilot CLI Because the Harness Is Now the Product

The most important line in Microsoft’s JetBrains Copilot announcement is not the word “JetBrains.” It is “harness.” GitHub is moving Copilot for JetBrains onto Copilot CLI as the default agent harness, and that tells you where the coding-agent market has landed: the runtime around the model is now the
16 Jun 2026 4 min read
codex

OpenAI’s Codex Capacity Incident Is a Reliability Story, Not Just an Outage Blip

A Codex capacity incident is not just “the chatbot was down for a bit.” That framing made sense when AI assistance meant autocomplete, chat, and the occasional generated regex. It breaks once the product is positioned as background engineering labor: cloud tasks, parallel worktrees, code review, issue triage, migrations, automations,
16 Jun 2026 4 min read
QwenPaw 1.1.12 Beta Is Mostly Boring — Which Is Why It Matters
qwen

QwenPaw 1.1.12 Beta Is Mostly Boring — Which Is Why It Matters

QwenPaw v1.1.12-beta.1 is not the kind of release that produces a clean launch graphic. Good. The parts that make local agents useful rarely fit in a hero image. The beta is mostly a pile of operational fixes: isolated keychain master keys, a release verification gate, token and
16 Jun 2026 5 min read
Qwen-RobotWorld Is Alibaba’s Bet That Robots Need a Simulator in Their Head
qwen

Qwen-RobotWorld Is Alibaba’s Bet That Robots Need a Simulator in Their Head

Alibaba’s most interesting Qwen news this week is not another chat model. It is a model that tries to answer a harder question before a robot moves: what happens next? That is the useful way to read Qwen-RobotWorld, the new arXiv technical report from Alibaba’s Qwen team. The
16 Jun 2026 5 min read
OpenClaw Adds a Stream Stall Guard Because Agent Providers Sometimes Finish the Work and Forget to Say Done
openclaw

OpenClaw Adds a Stream Stall Guard Because Agent Providers Sometimes Finish the Work and Forget to Say Done

OpenClaw PR #93640 is a reminder that “OpenAI-compatible” is not the same thing as operationally compatible. A provider can send every token the agent needs, include a finish_reason, record the request as successful on its dashboard, and still leave the client hanging forever because the stream never sends [DONE]
16 Jun 2026 4 min read
OpenClaw’s Kiro ACP Retry Fix Shows Why Multi-Agent Platforms Need Error Taxonomy, Not Hope
openclaw

OpenClaw’s Kiro ACP Retry Fix Shows Why Multi-Agent Platforms Need Error Taxonomy, Not Hope

OpenClaw PR #93642 is a one-line regex-sized fix with a platform-sized lesson: multi-agent systems do not become reliable just because everyone agrees to speak the same protocol. Eventually a backend says “Internal error,” the thread keeps looking alive, and the orchestrator has to know whether that means “retry safely” or
16 Jun 2026 4 min read
NVIDIA SkillSpector Treats Agent Skills Like Supply-Chain Artifacts
nvidia

NVIDIA SkillSpector Treats Agent Skills Like Supply-Chain Artifacts

Agent skills are starting to look less like helpful prompt folders and more like dependencies with a fake mustache. NVIDIA’s SkillSpector, updated today in its public GitHub repository, is useful precisely because it names the thing most coding-agent stacks have been politely avoiding: when you install a skill into
16 Jun 2026 6 min read
Claude Code 2.1.178 Tightens Subagent Permissions Where Auto Mode Used to Trust Too Much
ai-frameworks

Claude Code 2.1.178 Tightens Subagent Permissions Where Auto Mode Used to Trust Too Much

Claude Code 2.1.178 reads like a bugfix release until you notice the pattern: permissions are moving closer to the place where delegation actually happens. That is the right direction, because once agents can spawn agents, “ask before running bash” is no longer a sufficient security model. The runtime
16 Jun 2026 5 min read
Langfuse 3.187.0 Makes Agent Observability More Navigable — and MCP Hosting Less Brittle
ai-frameworks

Langfuse 3.187.0 Makes Agent Observability More Navigable — and MCP Hosting Less Brittle

Langfuse 3.187.0 is not a big-launch release. Good. Some of the most important agent-platform work right now looks boring from across the room: links resolve correctly, deployment hostnames stop breaking MCP, and evaluators can be deleted without turning the observability database into a haunted attic. The release matters
16 Jun 2026 4 min read
← Newer Posts Page 23 of 136 Older Posts →
The LGTM © 2026
  • Sign up
Powered by Ghost