Elon Musk (@elonmusk)· @elonmusk · X·· 2 小時前AI 評分56
AI 導讀
Grok 4.7 在法律領域獲得首位,並與 Harvey 合作推出 Harvey LAB-AA v1.1。
- Grok 4.7 在 Legal Agent Benchmark (LAB) 中排名第一。
- Harvey LAB-AA v1.1 更新評分方法,新增「hallucination check」並要求正確回答不包含重大誤述。
- 這些更新有助於驗證模型在法律資訊處理上的準確性與可靠性。
正文
Grok 4.7 ranks first in legal matters
Today we are announcing Harvey LAB-AA v1.1 in collaboration with Harvey. This updates our scoring methodology for the Legal Agent Benchmark (LAB) to add a hallucination check and require correct responses to not include material misstatements. LAB-AA v1.1's new headline metric,
來源:Elon Musk (@elonmusk) · x.com