The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe
DGX Spark’s Nemotron 3 Super Benchmark Is Useful Because It Measures Stability, Not Just Speed
nvidia

DGX Spark’s Nemotron 3 Super Benchmark Is Useful Because It Measures Stability, Not Just Speed

The most useful number in the latest DGX Spark Nemotron benchmark is not 23.45 tokens per second. It is zero. Zero crashes. Zero out-of-memory errors. A same-evening NVIDIA Developer Forum post reports Nemotron-3-Super-120B-A12B-NVFP4 running on a single DGX Spark at 23.45 tokens/sec for a tg128 single-session test,
13 May 2026 5 min read
Nemotron 3 Nano Looks Fast on Jetson Thor — Until Concurrency Makes the Runtime Tell the Truth
nvidia

Nemotron 3 Nano Looks Fast on Jetson Thor — Until Concurrency Makes the Runtime Tell the Truth

Jetson Thor can run serious local AI models now. That is no longer the interesting part. The interesting part is that a same-day NVIDIA Developer Forum benchmark shows exactly where the glossy edge-AI story starts paying rent: concurrency, runtime support, kernel selection, and parser glue. A user tested NVIDIA’s
13 May 2026 5 min read
Microsoft Put Grok 4.3 Behind Azure's Enterprise Guardrails
xai

Microsoft Put Grok 4.3 Behind Azure's Enterprise Guardrails

Microsoft adding Grok 4.3 to Foundry looks, at first glance, like the usual model-catalog checkbox: another frontier model, another deployment option, another pricing table for procurement to squint at. That is the boring read. The useful read is that Microsoft just put xAI’s most important developer model inside
13 May 2026 5 min read
Phoenix 15.8.0 Turns Agent Sessions Into Queryable Evidence — Then Hardens the Doors Around Them
ai-frameworks

Phoenix 15.8.0 Turns Agent Sessions Into Queryable Evidence — Then Hardens the Doors Around Them

Phoenix 15.8.0 is easy to misread as another observability dashboard release. That would miss the useful part. The release turns more agent traces into session-level evidence, then patches the UI and expression-evaluation surfaces that could corrupt the same evidence layer. That combination is the story. Arize shipped Phoenix
13 May 2026 4 min read
Pydantic AI 1.95.1 Fixes the Observability Regression That Durable Agents Cannot Afford
ai-frameworks

Pydantic AI 1.95.1 Fixes the Observability Regression That Durable Agents Cannot Afford

Pydantic AI 1.95.1 is a two-line release note with a production-sized warning label: if your agent framework treats observability as an optional plugin, your durable workflows will eventually make that lie expensive. The patch shipped on May 13 with two bug fixes. First, Pydantic AI now eagerly imports
13 May 2026 4 min read
Agent Orchestrator v0.7.0 Turns Coding Agents Into a PR Factory
agentic-coding

Agent Orchestrator v0.7.0 Turns Coding Agents Into a PR Factory

The most revealing thing about Agent Orchestrator v0.7.0 is that it does not treat the coding agent as the product. It treats the pull request as the product. That is a more useful abstraction. Chat transcripts do not ship. Branches, CI runs, review comments, and merged PRs do.
13 May 2026 5 min read
oh-my-openagent’s Wakeup Fixes Show Agentic Coding Has a Harness Problem
agentic-coding

oh-my-openagent’s Wakeup Fixes Show Agentic Coding Has a Harness Problem

The most honest agentic-coding release notes are the ones that sound like incident reports. oh-my-openagent v4.1.1 does not promise a smarter model, a magical benchmark jump, or a new mascot with a suspiciously large context window. It fixes background wakeups, continuation hooks, synthetic resumes, and dangerous interactive shell
13 May 2026 5 min read
Copilot Code Review Adds Severity and Grouping Because AI Review Noise Is Now the Bottleneck
codex

Copilot Code Review Adds Severity and Grouping Because AI Review Noise Is Now the Bottleneck

GitHub’s latest Copilot code review change is not flashy. That is why it is worth paying attention to. Copilot review comments now carry High, Medium, and Low severity labels. Similar suggestions can be grouped together instead of repeated across a large pull request. Users opted into GitHub’s new
13 May 2026 6 min read
OpenAI’s Windows Sandbox Is the Codex Postmortem We Needed Before the Incident
codex

OpenAI’s Windows Sandbox Is the Codex Postmortem We Needed Before the Incident

OpenAI’s Windows sandbox writeup is the kind of security post the agent industry needs more of: detailed, awkward, and honest about the places where the first version was not enough. The headline is easy to undersell. “Codex now has a better Windows sandbox” sounds like platform plumbing, useful mostly
13 May 2026 5 min read
AgentMemory 0.9.11 Shows the Agent Plugin Format War Is Really About Portable Authority
openclaw

AgentMemory 0.9.11 Shows the Agent Plugin Format War Is Really About Portable Authority

AgentMemory 0.9.11 looks, at first glance, like a small integration release: a Codex plugin manifest, a marketplace entry, and a fix for an OpenClaw memory-slot bug. That undersells it. The release is a useful snapshot of where coding-agent infrastructure is going: memory, hooks, skills, MCP servers, and host
13 May 2026 5 min read
AnyFlow Is NVIDIA’s Argument That Video Diffusion Needs a Throttle, Not Just a Bigger Engine
nvidia

AnyFlow Is NVIDIA’s Argument That Video Diffusion Needs a Throttle, Not Just a Bigger Engine

Video-generation products do not need a sacred sampler setting. They need a throttle. That is the practical argument hiding inside NVIDIA’s new AnyFlow checkpoints on Hugging Face. AnyFlow is framed as an “any-step” video diffusion framework: a single distilled model should adapt to arbitrary inference budgets instead of being
13 May 2026 5 min read
NVIDIA’s Wan2.2 FP8/NVFP4 Checkpoints Are the Boring Part of Video AI That Actually Matters
nvidia

NVIDIA’s Wan2.2 FP8/NVFP4 Checkpoints Are the Boring Part of Video AI That Actually Matters

The least glamorous part of generative video is becoming the part that matters: not whether the demo clip looks impressive, but whether the model can be served, measured, rolled back, and kept inside a cost envelope without turning every request into a GPU bonfire. NVIDIA’s new FP8 and NVFP4
13 May 2026 5 min read
← Newer Posts Page 101 of 136 Older Posts →
The LGTM © 2026
  • Sign up
Powered by Ghost