Skip to content

Brief · 28 September 2026

What changed

NVIDIA announced DSX MaxLPS, a new software‑hardware stack for its AI‑factory platform that promises higher throughput and lower energy per token, and made it generally available to DGX Cloud users today.

One number

16,000times

OpenAI agents scanned the UNCTAD statistics site between April and June

source ↗

Still vapor

NVIDIA markets MaxLPS as “maximizing AI factory throughput and efficiency,” yet the blog offers no independent benchmark or per‑token energy figure, making the claim hard to verify against real‑world DGX workloads.

The most concrete shift today is NVIDIA’s DSX MaxLPS rollout. The blog frames the stack as a way to squeeze more FLOPs out of existing DGX nodes while shaving power, but it stops short of publishing any head‑to‑head throughput numbers or power‑per‑token metrics. Practitioners on the forums are already asking for a side‑by‑side comparison with the prior DSX release, noting that without hard data the promised gains remain speculative.

In parallel, Hugging Face’s Holo4 model was introduced as a “generalist computer‑use agent.” The announcement touts broader tool use and multi‑modal reasoning, but again no benchmark deltas (SWE‑bench, MMLU, etc.) are disclosed, leaving buyers without a clear performance‑to‑price ratio.

Meanwhile, a separate security story shows OpenAI agents hammering the UNCTAD site 16,000 times in a few months, underscoring how quickly agent traffic can scale when unchecked.

Operators should demand transparent throughput and energy figures from NVIDIA before committing additional DGX capacity, and watch for Holo4’s first third‑party benchmark releases to gauge whether the model’s versatility translates into real compute savings.

Composed by the MadCoolStuff editor pipeline · Groq · openai/gpt-oss-120b · 2026-09-28

Tags

What we read