Ideas
NVIDIA's CUDA ecosystem delivered Day 0 support for DeepSeekV4 across B200, B300, and GB300, with vLLM/SGLang working out of the box. Despite TensorRT-LLM bugs, the overall software maturity and GB300
NVIDIA's CUDA ecosystem delivered Day 0 support for DeepSeekV4 across B200, B300, and GB300, with vLLM/SGLang working out of the box. Despite TensorRT-LLM bugs, the overall software maturity and GB300's rack-scale performance advantage ($0.156 per million output tokens) reinforce NVIDIA's inference leadership.
Risk: Huawei's Day 0 readiness and AMD's rapid 100x improvement suggest competitive pressure is rising; Nvidia's moat is not absolute.
AMD's MI355X and ROCm stack had disastrous Day 0 performance—only FP8 mode worked, distributed inference did not function, and ATOM engine ran at batch size 1 with no production customers. Although a
AMD's MI355X and ROCm stack had disastrous Day 0 performance—only FP8 mode worked, distributed inference did not function, and ATOM engine ran at batch size 1 with no production customers. Although a 100x improvement was achieved by Day 26, the article highlights that AMD SGLang and vLLM progress lags far behind CUDA, and the company is focusing on ATOM instead of the widely-used vLLM.
Risk: If AMD continues to prioritize proprietary engines over open-source ecosystem, it may cede inference market share to NVIDIA and Huawei.
CoreWeave is explicitly thanked for contributing compute to the open-source community, scrambling to provide two spare GB300 NVL72 racks for benchmarking. This validates CoreWeave's role as a key enab
CoreWeave is explicitly thanked for contributing compute to the open-source community, scrambling to provide two spare GB300 NVL72 racks for benchmarking. This validates CoreWeave's role as a key enabler of cutting-edge AI infrastructure and positions it as a go-to cloud provider for frontier inference workloads.
Risk: Heavy dependency on NVIDIA hardware supply and potential competition from other GPU cloud providers.
This newsletter, published June 09, 2026,
features Bryan Shan
discussing NVDA, AMD, CRWV.
3 trade ideas extracted by AI with direction and confidence scoring.
Speakers:
Bryan Shan
· Tickers:
NVDA,
AMD,
CRWV