Show notes
NVIDA 技術 blog: https://developer.nvidia.com/blog/inside-nvidia-groq-3-lpx-the-low-latency-inference-accelerator-for-the-nvidia-vera-rubin-platform/ Dylan Patel 訪談: https://www.youtube.com/watch?v=mDG_Hx3BSUE ▍GTC 相關討論 01:20 近期崛起的 OpenClaw 如何融入生活流程,以及 OpenClaw、claude cowork、NemoClaw 彼此的差異。 07:08 NVIDIA 投資 Groq,以及於 GTP 大會中推出 LPU 的脈絡。 09:01 LLM 在推論時的 Prefill 與 Decode 階段分流,以及 LPU 如何優化 Token 產出的延遲問題。 ▍AI 當前的瓶頸? 28:50 SemiAnalysis 創辦人 Dylan Patel 指出了 AI 的兩大瓶頸:EUV、記憶體,以及相關的衍生稀缺資源。 ______________________  ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ ̄ Podcast 回饋表單:https://statementdog.cc/PMOa8 Podcast 業務合作聯絡信箱:[email protected]



