The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe
Vercel AI SDK Workflow Beta Closes a System-Message Escape Hatch by Default
ai-frameworks

Vercel AI SDK Workflow Beta Closes a System-Message Escape Hatch by Default

Vercel’s AI SDK Workflow beta just changed one default, and it is exactly the kind of default agent frameworks need to stop getting wrong. @ai-sdk/[email protected], published June 19, now rejects system messages inside prompt or messages by default for WorkflowAgent. The behavior now matches
19 Jun 2026 3 min read
Phoenix 17.9.0 Gives Its Built-In Agent a Bash Tool — Then Adds the Kill Switch It Needed
ai-frameworks

Phoenix 17.9.0 Gives Its Built-In Agent a Bash Tool — Then Adds the Kill Switch It Needed

“Give the observability agent a shell” is either a very good idea or the beginning of an incident report. Phoenix 17.9.0 is worth covering because Arize appears to understand both halves of that sentence. The arize-phoenix-v17.9.0 release adds a server-side bash tool for Phoenix’s built-in
19 Jun 2026 3 min read
LiveKit Agents 1.6.2 Is Voice-Agent Plumbing Growing Up: Latency, Usage Metrics, and Provider Failure Modes
ai-frameworks

LiveKit Agents 1.6.2 Is Voice-Agent Plumbing Growing Up: Latency, Usage Metrics, and Provider Failure Modes

Voice agents are where infrastructure excuses go to die. A text agent can hide latency behind a spinner, retry silently, or dump a trace after the fact. A voice agent has to respond in real time, avoid talking over the user, manage provider failures mid-conversation, and keep cost under control
19 Jun 2026 4 min read
Mastra 1.45.0 Fixes the Hidden Cost of Agent State: Long Threads, Signal Order, and Prompt Confusion
ai-frameworks

Mastra 1.45.0 Fixes the Hidden Cost of Agent State: Long Threads, Signal Order, and Prompt Confusion

Mastra’s latest release is the kind of agent-framework update that will not win a demo day and absolutely will save someone’s production system from death by transcript archaeology. @mastra/[email protected], published June 19, fixes three related problems in how long-running agents carry state: restoring tracked
19 Jun 2026 4 min read
google-ai

DeepMind’s AI Control Roadmap Says Coding Agents Need Seatbelts, Not Vibes

Google DeepMind’s latest agent-safety post is not exciting in the way model launches are exciting. Good. The useful work here is aggressively unglamorous: threat models, monitoring coverage, recall, response latency, permission tiers, and a frank admission that powerful internal agents should be treated less like clever autocomplete and more
19 Jun 2026 5 min read
Security Fine-Tuning Can Calibrate the Output While Leaving the Model Clueless
ai-models

Security Fine-Tuning Can Calibrate the Output While Leaving the Model Clueless

The lazy version of “AI for security” is to fine-tune a code model, report a binary accuracy number, and hope nobody asks whether the model is finding vulnerabilities or merely learning when to say the scary word. Calibration Without Comprehension asks the uncomfortable question directly, and the answer is not
19 Jun 2026 4 min read
SoftSkill Compresses Agent Skills Into 32 Virtual Tokens — and Shows Where That Trick Breaks
ai-models

SoftSkill Compresses Agent Skills Into 32 Virtual Tokens — and Shows Where That Trick Breaks

Markdown skills are one of the more sensible things agent platforms have converged on. They are readable, reviewable, portable, and cheap to author. They are also long, repeatedly injected, and interpreted from scratch by a model that may or may not translate prose into the behavior you intended. The obvious
19 Jun 2026 4 min read
AGENTS.md Works When It Is Debugged Like Code, Not When It Is Treated Like Lore
ai-models

AGENTS.md Works When It Is Debugged Like Code, Not When It Is Treated Like Lore

AGENTS.md has become the new README-for-robots: part documentation, part incantation, part apology to the coding agent that is about to discover your monorepo’s build system. The useful question is no longer whether repository guidance helps. Sometimes it does. The question is whether teams are debugging that guidance like
19 Jun 2026 3 min read
4-bit KV Cache Is the Agent Cost Story Hiding Under the Model Benchmark
ai-models

4-bit KV Cache Is the Agent Cost Story Hiding Under the Model Benchmark

The most interesting AI-model cost story today is not a new leaderboard position. It is the thing that happens after your agent has read half a repository, opened a dozen tools, accumulated a giant prefix, and now needs to answer one tiny follow-up without turning the GPU into a space
19 Jun 2026 3 min read
GitHub’s Qubot Shows Enterprise Agents Need Curated Context, Evals, and Boring SQL
codex

GitHub’s Qubot Shows Enterprise Agents Need Curated Context, Evals, and Boring SQL

The most useful sentence in GitHub’s Qubot write-up is not the one about plain-language analytics. It is the part where the company admits that structured, curated context made the internal agent more accurate and three times faster at returning the right answer. That is the enterprise-agent story hiding under
19 Jun 2026 5 min read
qwen

Qwen Code’s June 19 Patch Stack Is About Operability, Not Benchmarks

Qwen Code’s June 19 patch stack is the sort of release work that rarely gets a launch thread and absolutely determines whether developers trust an agent after lunch. No new frontier model. No benchmark victory lap. Instead: memory retention, token telemetry, ACP child-process hygiene, Windows sandbox parsing, and a
19 Jun 2026 5 min read
Qwen Code Makes qwen serve a Real Browser Agent
qwen

Qwen Code Makes qwen serve a Real Browser Agent

The most important thing Qwen Code shipped today is not a smarter model. It is a boring distribution decision: qwen serve can now serve the browser Web Shell itself. That sounds like plumbing because it is plumbing. But plumbing is where coding agents either become tools people can actually run
19 Jun 2026 4 min read
← Newer Posts Page 14 of 136 Older Posts →
The LGTM © 2026
  • Sign up
Powered by Ghost