Tag
#nvidia
Every story tagged nvidia, newest first.

AWS's new G7 instances won its own benchmark with half the GPUs
AWS benchmarked its Blackwell-based G7 instances against G5, G6 and G6e on 30B mixture-of-experts models. The two-GPU box beat the four-GPU boxes, the cheapest configuration and the fastest one turned out to be different machines, and the same model cost 2.5 times more per token on retrieval traffic than on chat.
Ahmad J · Sep 8, 2026 · 6 min read

The US–China Chip Export Controls: Where They Restrict, and Where They Don't
Painted as a wall or a sieve, the US-China chip controls are really a colander with deliberately sized holes. This maps what the thresholds actually catch, where enforcement leaks, and why the line keeps moving, from FinFET nodes and TPP limits to the A800-to-H200 design-around loop.
Ahmad J · Aug 11, 2026 · 5 min read

CUDA Lock-In Is Real: A Precise Cost Accounting of What Switching GPU Vendors Actually Breaks
Switching away from CUDA isn't one migration: it's five separate porting problems, hidden revalidation costs, and an org chart that fights you the whole way.
Ahmad J · Aug 6, 2026 · 9 min read

The best home-server hardware for self-hosting AI in 2026
The used RTX 3090 is still the value pick for local AI in 2026, but a brutal memory shortage reshuffled every price, and 128GB unified boxes (DGX Spark, Strix Halo, Mac Studio) now run big MoE models a 24GB GPU can't hold. Here's what to actually buy.
Ahmad J · Aug 3, 2026 · 13 min read

The best cloud GPU providers for AI training in 2026
CoreWeave shipped NVIDIA's GB300 NVL72 first and runs GB200 at scale; Google's Ironwood TPU reached GA in late 2025; AWS and Azure now ship Blackwell; Lambda and Nebius rent B200 affordably. A skeptic's 2026 buyer's guide to hardware, trade-offs, and representative pricing.
Ahmad J · Aug 3, 2026 · 11 min read
More stories
Guide · aiThe best cloud hosting for running AI models in 2026Aug 2, 2026 · 11 min read
Article · chipsHBM Explained: Why High-Bandwidth Memory Is the Real Bottleneck in AI ChipsJul 29, 2026 · 12 min read
Article · chipsWhat Chiplets Are and Why Chipmakers Moved to ThemJul 28, 2026 · 10 min read
Guide · aiThe best GPUs for running large language models locally in 2026Jun 20, 2026 · 10 min read
Guide · aiThe best mini PCs for local AI inference in 2026Jun 20, 2026 · 10 min read