跳到正文
Elon Musk (@elonmusk)· @elonmusk · X·· 2 小時前AI 評分56
AI 導讀

Grok 4.7 在法律領域獲得首位,並與 Harvey 合作推出 Harvey LAB-AA v1.1。

  • Grok 4.7 在 Legal Agent Benchmark (LAB) 中排名第一。
  • Harvey LAB-AA v1.1 更新評分方法,新增「hallucination check」並要求正確回答不包含重大誤述。
  • 這些更新有助於驗證模型在法律資訊處理上的準確性與可靠性。
正文

Grok 4.7 ranks first in legal matters
Today we are announcing Harvey LAB-AA v1.1 in collaboration with Harvey. This updates our scoring methodology for the Legal Agent Benchmark (LAB) to add a hallucination check and require correct responses to not include material misstatements. LAB-AA v1.1's new headline metric,

來源:Elon Musk (@elonmusk) · x.com