Artificial Analysis· @ArtificialAnlys · X·· 3 小时前AI 评分39
AI 导读
在 Hallucination-Gated All-Pass Rate 高于 0% 的模型中,有四款在得分与每任务成本之间构成了帕累托前沿:GPT-6 Luna (max)、GPT-6.1 Sol (max)、Muse Spark 1.3 (max) 和 Grok 4.7 (xhigh)。Grok 4.7 (xhigh) 以每任务约 $9.50 领先,Muse Spark 1.3 (max) 以约 $4.20 位居第二,而三款 Claude 模型每任务成本约 $18 至 $22。GPT-6 Luna (max) 最便宜,每任务约 $0.22,得分 3.3%。
正文
Among models with a Hallucination-Gated All-Pass Rate above 0%, four set the Pareto frontier for score vs. Cost per Task: GPT-6 Luna (max), GPT-6.1 Sol (max), Muse Spark 1.3 (max) and Grok 4.7 (xhigh). Grok 4.7 (xhigh) leads at ~$9.50 per task and Muse Spark 1.3 (max) comes second at ~$4.20, while the three Claude models cost ~$18 to ~$22 per task. GPT-6 Luna (max) is the cheapest at ~$0.22 per task, scoring 3.3%.
来源:Artificial Analysis · x.com