跳到正文
北京时间
原文
Rohan Paul· @rohanpaul_ai · X·· 1 天前AI 评分45
AI 导读

投资人Chamath Palihapitiya指出,在AI计算中,Prefill阶段是计算受限的,因此随着上下文增长,大规模并行GPU(如Nvidia)占据优势;而Decode阶段受内存带宽限制,因为生成每个新token都需要扫描已生成的内容。

正文

Chamath on all important “prefill” and “decode.” in AI compute.
Prefill is compute-bound; massive parallel GPUs win, so Nvidia dominates as context grows.
Decode is memory-bandwidth bound as each next token depends on scanning what’s already generated

来源:Rohan Paul · x.com