The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe
Blackwell’s MLPerf Sweep Is Really a Software-Stack Story
nvidia

Blackwell’s MLPerf Sweep Is Really a Software-Stack Story

MLPerf results usually arrive dressed as trophy charts: fastest here, largest there, a few bars that get recycled into cloud sales decks before lunch. NVIDIA’s MLPerf Training 6.0 sweep is more useful than that if you read it as an engineering diff. The headline is Blackwell winning every
16 Jun 2026 5 min read
Copilot Usage Metrics Just Got Less Wrong — Now Admins Need to Relearn Their Baselines
azure-ai

Copilot Usage Metrics Just Got Less Wrong — Now Admins Need to Relearn Their Baselines

GitHub made Copilot usage metrics more accurate, which is good news in the same way a scale getting calibrated is good news: useful, necessary, and immediately dangerous if everyone forgets the baseline changed. The June 15 update adds server-side telemetry to Copilot reports, so enterprise dashboards will include more active
16 Jun 2026 5 min read
GitHub Code Quality GA Turns AI Review Into a Three-Part Bill
azure-ai

GitHub Code Quality GA Turns AI Review Into a Three-Part Bill

GitHub Code Quality is leaving preview, and the feature announcement is not the part engineering leaders should read twice. The pricing model is. On July 20, 2026, GitHub turns Code Quality into a paid product that combines a per-committer subscription, AI-credit consumption, and GitHub Actions minutes. That is not a
16 Jun 2026 5 min read
GitHub Models Retirement Pushes New AI Prototyping Back Toward Azure AI Foundry
azure-ai

GitHub Models Retirement Pushes New AI Prototyping Back Toward Azure AI Foundry

GitHub Models is not dead yet. It is just no longer the place Microsoft wants new AI platform work to begin. That distinction matters, because retirement notices usually sound operational while the real message is architectural: the model experimentation surface inside GitHub is being narrowed, and the serious path now
16 Jun 2026 5 min read
SpaceX Buying Cursor Is the Clearest Sign Yet That Grok Build Is Not Just a Side Quest
xai

SpaceX Buying Cursor Is the Clearest Sign Yet That Grok Build Is Not Just a Side Quest

SpaceX buying Cursor for an implied $60 billion is easy to misread as another entry in the Elon Musk mega-deal ledger. The more useful read is narrower and more important for developers: xAI is buying the workflow layer Grok Build does not yet own. The June 16 SEC filing says
16 Jun 2026 5 min read
Satellites Turns Agentic SDLC From Prompt Advice Into Reviewer-Gated State Transitions
agentic-coding

Satellites Turns Agentic SDLC From Prompt Advice Into Reviewer-Gated State Transitions

The most fragile sentence in agentic development is “the agent will follow the process.” It sounds reasonable until the task gets long, the context fills, a subagent is dispatched, or the model discovers that changing the checklist is easier than satisfying it. Prompt advice is useful. It is not a
16 Jun 2026 4 min read
CLIProxyAPI Shows the Shadow API Layer Around Coding Agents Is Becoming Real Infrastructure
agentic-coding

CLIProxyAPI Shows the Shadow API Layer Around Coding Agents Is Becoming Real Infrastructure

The official story of coding agents is polished: better models, cleaner IDE integrations, safer enterprise controls. The unofficial story is messier and probably more revealing. Developers are already building proxy layers around Gemini CLI, Antigravity, ChatGPT Codex, Claude Code, and Grok Build because the market has made three things simultaneously
16 Jun 2026 4 min read
OneHarness Makes the Coding-Agent Benchmark Problem Look Like an Integration Problem
agentic-coding

OneHarness Makes the Coding-Agent Benchmark Problem Look Like an Integration Problem

The weakest part of most “best AI coding agent” comparisons is not the model take. It is the harness. If Claude Code gets the system prompt as a real system instruction, Codex gets it as user text, OpenCode starts in a different effective working directory, and Cursor cannot resume the
16 Jun 2026 4 min read
Pixel Drop Shows Google’s Gemini Strategy: Put Creation Tools Where the Camera Roll Already Is
google-ai

Pixel Drop Shows Google’s Gemini Strategy: Put Creation Tools Where the Camera Roll Already Is

Google’s June Pixel Drop is easy to misread as a laundry list: Screen Reactions, Gemini Omni video editing, Gemini music generation, Android 17 Bubbles, Magic Cue in Snapchat, Voice Translate on Pixel 10a, AirDrop-compatible Quick Share, Ask Photos in more European markets, call-screening updates, emergency-contact automation. Product people love
16 Jun 2026 4 min read
Android 17 Turns Gemini Into a Platform Contract, Not Just an Assistant
google-ai

Android 17 Turns Gemini Into a Platform Contract, Not Just an Assistant

Android 17’s loudest features are the ones users will notice first: Bubbles, Screen Reactions, foldable improvements, tighter security, better memory behavior. Fine. The developer story is sharper: Google is turning Android apps into callable capability surfaces for Gemini and other agents. That is a much bigger shift than “Gemini
16 Jun 2026 4 min read
google-ai

DeepMind’s Planning Prototype Is Gemini Doing Bureaucracy — Exactly Where Agents Get Tested

Google DeepMind’s newest public-sector AI project is not a chatbot trying to sound clever. It is Gemini pointed at bureaucracy: PDFs, maps, consultation letters, local policy, precedents, missing fields, draft reports, and the thousand tiny frictions that turn a householder planning application into a slow queue. That is precisely
16 Jun 2026 4 min read
DeepRubric Shows Deep Research Agents Need Better Rewards Before They Need More Rollouts
ai-models

DeepRubric Shows Deep Research Agents Need Better Rewards Before They Need More Rollouts

Deep research agents do not mainly need longer traces. They need rewards that can tell the difference between grounded synthesis and beautifully cited filler. That is why DeepRubric is more useful than the average “agent benchmark goes up” paper. It treats the reward pipeline as the product, not as a
16 Jun 2026 4 min read
← Newer Posts Page 22 of 136 Older Posts →
The LGTM © 2026
  • Sign up
Powered by Ghost