[BidClub_]

PERSON DIRECTORY

Cam Quilici

Cam Quilici appears in 2 indexed conversations across SemiAnalysis. This directory brings every appearance, source, TL;DR, digest, and transcript into one searchable feed.

2 EPISODES1 SHOW
2 episodes
Language
SemiAnalysisEN · 35 min

Ep. 017 - DeepSeek V4 and Huawei Ascend NPU Performance (InferenceX) | Kimbo Chen, Cam Quilici, Bryan Shan, Jordan Nanos

Kimbo ChenCam QuiliciBryan ShanJordan Nanos

DeepSeek V4 changes inference economics by combining million-token context with roughly 100X lower KV-cache usage, while MegaMoE claims 1.5-1.73X speedups through communication-computation overlap.Huawei Ascend delivered credible release-day performance, but the larger catalyst is software iteration: AgentX will test realistic Claude Code traces, caching, and prefill-decode systems beyond synthetic chip benchmarks.

SemiAnalysisEN · 51 min

Ep. 002 - InferenceX 2.0 Release (Technical Staff) | Cam Quilici, Bryan Shan, Doug O'Laughlin, Jordan Nanos

Cam QuiliciBryan ShanDoug O'LaughlinJordan Nanos

InferenceX 2.0 shows GB200/GB300 delivering 20x DeepSeek-R1 throughput per GPU versus a fully tuned H100 at low interactivity, and 80–100x at 100 tokens/sec/user.NVLink’s 72-GPU domain, software optimization, and multi-token prediction drive the gap, while MI355’s roughly 25% TCO advantage over B200 highlights margin potential; composability, legacy fleets, and larger frontier models remain risks.