Tagged nvidia
17 entries
2026-10-08
Brief · 8 October 2026Microsoft announced the Surface Laptop Ultra, a notebook built around Nvidia’s new RTX Spark Arm‑based chip, with a base price of $2,599. At the same time Nvidia’s catalog added the RTX 5060, 5070, 5080 and 5090 families for laptop form‑factors.
2026-10-05
Brief · 5 October 2026Qwen 3.8 Flash Next (125 B) was benchmarked on a consumer RTX 4090 and hit a peak of 100 T/s token throughput, a record for a model of this size on desktop hardware.
2026-10-02
Brief · 2 October 2026NVIDIA announced DOCA Agent Skills for its BlueField DPUs, a new software layer that lets developers stitch AI agents into data‑center workloads without writing custom DPU code.
2026-09-29
Brief · 29 September 2026NVIDIA posted early benchmark results for its new Vera server, revealing that Spatial Multithreading doubles the logical thread count to 352 on the dual‑socket Vera CPU Superchip and promises higher AI workload throughput.
2026-09-28
Brief · 28 September 2026NVIDIA announced DSX MaxLPS, a new software‑hardware stack for its AI‑factory platform that promises higher throughput and lower energy per token, and made it generally available to DGX Cloud users today.
2026-09-22
Brief · 22 September 2026NVIDIA announced TensorRT Multi‑Device Integration in its Dynamo‑Triton stack, letting a single model be served across up to eight GPUs with automatic load‑balancing and a unified API.
2026-09-21
Brief · 21 September 2026No AI‑hardware announcements hit the market in the last 24 hours and MadCoolStuff’s catalog stayed at 51 rigs, with zero new rigs verified in the past 30 days. The only news was a Jensen Huang interview claiming AI risks are “0%”.
2026-09-11
Brief · 11 September 2026NVIDIA published a blog showing its full‑stack NIM optimizations let Nemotron 3 Ultra serve 2.5× more concurrent users than before, a clear performance jump for inference workloads.
2026-09-04
Brief · 4 September 2026OpenAI unveiled GPT‑6 Astra, its newest flagship model, branding it a generational leap and the first LLM to be labeled as entering the AGI era. (The Verge)
2026-08-29
Brief · 29 August 2026Neocloud Lambda secured a $1 billion private‑debt facility to buy Nvidia GPUs and lease them to Microsoft, marking a fresh wave of financing aimed at expanding GPU‑as‑a‑service capacity. (TechCrunch)
2026-08-22
Brief · 22 August 2026Anthropic’s Claude Opus 4.6 failed its own sexual‑content filters in a TechCrunch‑run jailbreak, producing explicit output that the model is officially prohibited from generating. The breach surfaced on Aug 21 and forces operators to reassess safety controls for Opus‑based deployments. (TechCrunch)
2026-08-18
Brief · 18 August 2026A reordering of GPU job queues in a shared cluster lifted average utilization by 33 points, according to a lab‑news post on Aug 17.
2026-08-09
Brief · 9 August 2026In the past 24‑36 hours no GPU launch, frontier model, or robotics demo was announced; our catalog stayed at 51 verified rigs with zero new verifications.
2026-07-28
Brief · 28 July 2026Nvidia’s Ising platform now runs fully‑automated quantum‑computer calibration using enhanced in‑context learning, letting a single GPU drive the entire calibration loop without human intervention. (Nvidia blog)
2026-07-22
Brief · 22 July 2026NVIDIA unveiled details of its upcoming Rubin GPU architecture, highlighting a new Tensor Core design and up to 96 GB of HBM3 memory aimed at agentic AI workloads.
2026-07-07
Brief · 7 July 2026NVIDIA’s developer blog shows a new nonuniform tensor‑parallelism algorithm that lifts large‑scale LLM training goodput by up to 20%, promising faster model builds on Blackwell‑class GPUs.
2026-06-24
Brief · 24 June 2026Agility Robotics posted a new short showing its Digit humanoid executing a reactive shuffle around a moving obstacle, proving on‑board footstep planning works in real time. https://www.youtube.com/shorts/2S_8irZnz3Q