The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe
Microsoft Puts Codex Inside the Enterprise Compliance Boundary
agentic-coding

Microsoft Puts Codex Inside the Enterprise Compliance Boundary

Codex just got a more boring deployment story. That is not a complaint. In enterprise software, boring is often the feature that decides whether a tool gets piloted by one enthusiastic staff engineer or standardized across a thousand developers. Microsoft’s new guidance for running OpenAI Codex through Azure OpenAI
14 May 2026 4 min read
Google’s Gemini Startup Forum Is Distribution Strategy, Not Founder Philanthropy
google-ai

Google’s Gemini Startup Forum Is Distribution Strategy, Not Founder Philanthropy

Google’s second Gemini Startup Forum looks like a founder-support announcement. Read it closer and it is something more useful: a map of how Google plans to turn Gemini from a capable model family into default startup infrastructure. The company says 102 startups will gather at its Sunnyvale headquarters next
14 May 2026 5 min read
Hy3 Has the Traffic, Claude Has the Taste Test
llm-rankings

Hy3 Has the Traffic, Claude Has the Taste Test

The useful story in today’s model rankings is not that Tencent “beat” Anthropic. That would be a lazy reading of a usage chart, and lazy readings are how teams end up wiring production systems to leaderboard vibes. The better story is sharper: developers are increasingly treating LLMs less like
14 May 2026 5 min read
Microsoft’s MDASH Says the Security Agent Is a System, Not a Model
ai-models

Microsoft’s MDASH Says the Security Agent Is a System, Not a Model

Microsoft’s MDASH announcement looks like another frontier-AI benchmark post until you read the architecture. Then the actual story snaps into focus: the security agent is not a model. It is a factory. Microsoft says its new multi-model agentic security harness helped find 16 Windows vulnerabilities in the May 12
14 May 2026 5 min read
Frontier Cyber Models Are Now Outrunning the Benchmarks Built to Measure Them
ai-models

Frontier Cyber Models Are Now Outrunning the Benchmarks Built to Measure Them

The scary part of the UK AI Security Institute’s latest cyber-model update is not that Claude Mythos Preview or GPT-5.5 can “do security work.” We already knew that. The scary part is that the measurement apparatus is starting to look underpowered. AISI says frontier models have now substantially
14 May 2026 5 min read
OpenAI Daybreak Is the Security-Agent Arms Race Moving From Demo to Doctrine
codex

OpenAI Daybreak Is the Security-Agent Arms Race Moving From Demo to Doctrine

The security-agent race has moved past the demo phase. OpenAI’s new Daybreak initiative is not just another “AI finds bugs” announcement; it is a statement about where vulnerability work is going: inside the software delivery loop, paired with agentic code execution, threat modeling, validation, and patch generation. That is
14 May 2026 5 min read
Claude Code 2.1.140 Is a Boring Patch Release With the Right Enterprise Smell
claude-code

Claude Code 2.1.140 Is a Boring Patch Release With the Right Enterprise Smell

Claude Code 2.1.140 is the kind of release nobody screenshots and every enterprise pilot depends on. There is no grand launch narrative here. Anthropic shipped a patch release on May 12 with fixes for symlinked managed settings, marketplace and plugin identity mismatches, background service startup under enterprise endpoint
14 May 2026 4 min read
Anthropic’s Agent SDK Credits End the Claude Subscription Arbitrage Era
claude-code

Anthropic’s Agent SDK Credits End the Claude Subscription Arbitrage Era

Anthropic did not kill programmatic Claude usage. It killed the most interesting loophole in the Claude subscription model. Starting June 15, 2026, Agent SDK usage, claude -p, and third-party apps built on the Agent SDK will no longer draw from ordinary Claude subscription limits. Instead, paid users can claim a
14 May 2026 4 min read
MCP’s Database Problem Is Not Prompt Injection. It Is Old Bugs Wearing an Agent Badge
claude-code

MCP’s Database Problem Is Not Prompt Injection. It Is Old Bugs Wearing an Agent Badge

The mistake is calling this an MCP security story. It is really a database security story with an agent-shaped blast radius. Akamai researcher Tomer Peled found serious flaws in MCP servers connected to Apache Doris, Apache Pinot, and Alibaba RDS. The Register’s coverage is fresh, but the vulnerabilities themselves
14 May 2026 4 min read
Claude Code 2.1.141 Is a Control-Plane Release Wearing a Patch-Note Hoodie
claude-code

Claude Code 2.1.141 Is a Control-Plane Release Wearing a Patch-Note Hoodie

Claude Code 2.1.141 looks like a patch release until you read the nouns. Hooks. Workload identity. MCP auth states. Remote Control tokens. OTel spans. Background-agent permission modes. That is not a grab bag; it is a control plane getting stress-tested by real users. The headline feature is not
14 May 2026 4 min read
Unsloth’s Qwen3.6 MTP GGUFs Make Local Coding Agents Less Theoretical
qwen

Unsloth’s Qwen3.6 MTP GGUFs Make Local Coding Agents Less Theoretical

Unsloth’s new Qwen3.6-27B MTP GGUF release is not interesting because the internet needed another quantized model file. It is interesting because it documents the messy runtime details that decide whether local coding agents feel usable or collapse into a pile of parser errors, exhausted token budgets, and “works
13 May 2026 5 min read
Qwen Code v0.15.11 Ships the Boring Parts Coding Agents Actually Need
qwen

Qwen Code v0.15.11 Ships the Boring Parts Coding Agents Actually Need

Qwen Code v0.15.11 is the kind of release that looks skippable if you only read model cards and benchmark tables. That is exactly why it matters. The open coding-agent race is past the point where “the model can edit files” is a meaningful claim. Every serious team testing
13 May 2026 5 min read
← Newer Posts Page 100 of 136 Older Posts →
The LGTM © 2026
  • Sign up
Powered by Ghost