AIOps / Observability / Infrastructure

News, with an operator’s perspective.

Developments in AI and the infrastructure behind it. What changed, how it works, and what it means for the people running it.

Our editorial standards

Following the story

Developments since our original reporting, with dated updates and supporting sources.

Releases, research & operational impact

Latest coverage

Inference infrastructureNews analysis

TensorRT’s multi-GPU serving results put capacity in focus

NVIDIA’s September 21 demonstration accelerates one video-generation workload across eight GPUs. A service decision still needs the same offered load and completed-work count.

3 min read

Agent evaluationNews analysis

SWE-Serve tests coding agents through live inference

NVIDIA’s new SGLang benchmark finds patches that pass local checks but fail a running server. Its result is a reason to examine what a green test actually covers.

3 min read

Local AI agentsNews analysis

Antigravity brings local models to its agent SDK

Google adds local Gemma and LiteRT workflows. Running the model on a workstation still leaves a separate decision about what its tools may read, change or send.

3 min read

AI infrastructureNews analysis

NVIDIA Topograph puts GPU placement evidence in focus

New Topograph guidance connects physical network topology to workload schedulers. The operational check is whether the scheduler’s view survives cluster change.

3 min read