IX2-0022

① SA Source

Context Before

Key Observations and Results to Highlight

We see competitive perf per TCO results on FP8 MI355X disagg+wideEP SGLang on AMD compared to FP8 B200 disagg+wideEP SGLang, but when compared to widely used Dynamo TRTLLM B200 FP8, TRT continues to framemog. This is amazing news that AMD SGLang Disagg prefill+wideEP for FP8 is able to match NVIDIA’s SGLang performance.

Evidence

We also see that for single node aggregated serving, AMD’s SGLang delivers better perf per TCO than NVIDIA’s SGLang for FP8

Context After

SemiAnalysis InferenceX is free open source software and reader-supported. To receive new posts and support our work consider becoming a free or paid subscriber.

Subscribed

② Atomic Claim

在 single-node aggregated serving 中,AMDSGLangFP8 下,每單位 TCO 效能優於 NVIDIASGLang

  • Epistemic Mode: ASSERTED
  • Mapping Status: COMPLETE

③ Semantic Frame

{
  "comparison_expression": "在 single-node aggregated serving 中,AMD 的 SGLang 在 FP8 下,每單位 TCO 效能優於 NVIDIA 的 SGLang。",
  "entities": [
    {
      "id": "02_companies/AMD",
      "label": "AMD"
    },
    {
      "id": "04_knowledge_base/SGLang",
      "label": "SGLang"
    },
    {
      "id": "04_knowledge_base/FP8",
      "label": "FP8"
    },
    {
      "id": "02_companies/NVDA",
      "label": "NVIDIA"
    }
  ],
  "frame_type": "COMPARISON",
  "metric": "COST",
  "operator": "OUTPERFORMS",
  "qualifiers": {
    "condition_text": null,
    "numeric_mentions": [],
    "temporal_mentions": []
  }
}

④ Canonical Entity Mapping

RoleSurface LabelCanonical Target
comparison_entity_0AMDAMD
comparison_entity_1SGLangSGLang
comparison_entity_2FP8FP8
comparison_entity_3NVIDIANVDA

⑤ Human Review

請在 Properties 逐項確認:

  • 原文 → Atomic Claim 是否忠實
  • Atomic Claim → Semantic Frame 是否忠實
  • Canonical Entity mapping 是否正確
  • Epistemic mode 是否保留原文語氣
  • 最後選擇 review_action

Review state

Markdown 內文不是正式 approval。只有 Apply bridge 寫入的 Decision Ledger event 才是正式決策。