Grok 4.5 Beats Fable 5 And Opus 4.8 In Agent AI Test With 51.4% Score
Grok 4.5 gave Elon Musk’s cost-and-performance claim new support after an independent agent benchmark ranked the model first. Key Points: Grok 4.5 scored 51.4% on AutomationBench-AA, ahead of two Claude models. The model cost $0.34 per task, far below its closest Anthropic rivals. Its higher rule-vi