NIEK2-0137

① SA Source

Context Before

For LPUs, deploying a draft model or MTP layers is quite different from applying AFD. FFNs are stateless, while draft models and MTP layers require dynamic KV cache loading. Each FFN is around hundreds of megabytes, whereas draft models and MTP layers take up tens of gigabytes. To support this memory usage, LPUs can access up to 256 GB of DDR5 per Fabric Expansion Logic FPGAs on the LPX compute tray.

LPX Rack System

Evidence

This 32 tray 1U version that Nvidia has shown off at GTC is very close to Groq’s original server design before the acquisition

Context After

image

Source: SemiAnalysis Accelerator Model

② Atomic Claim

Nvidia 在 GTC 展示的 32-tray、1U 版本,與 Groq 被收購前的原始伺服器設計非常接近。

  • Epistemic Mode: ASSERTED
  • Mapping Status: COMPLETE

③ Semantic Frame

{
  "additional_nodes": [],
  "frame_type": "RELATION",
  "object": {
    "id": "02_companies/Groq",
    "label": "Groq"
  },
  "predicate": "ACQUIRES",
  "qualifiers": {
    "condition_text": null,
    "numeric_mentions": [
      "32",
      "1"
    ],
    "temporal_mentions": []
  },
  "subject": {
    "id": "02_companies/NVDA",
    "label": "Nvidia"
  }
}

④ Canonical Entity Mapping

RoleSurface LabelCanonical Target
subjectNvidiaNVDA
objectGroqGroq

⑤ Human Review

請在 Properties 逐項確認:

  • 原文 → Atomic Claim 是否忠實
  • Atomic Claim → Semantic Frame 是否忠實
  • Canonical Entity mapping 是否正確
  • Epistemic mode 是否保留原文語氣
  • 最後選擇 review_action

Review state

Markdown 內文不是正式 approval。只有 Apply bridge 寫入的 Decision Ledger event 才是正式決策。