Crypto news

23.07.2026
16:58

Grok 4.5 crushed the competition: Next-generation AI agents storm the browser

The artificial intelligence market is entering a new phase. It's no longer just about generating text or images — now AI is learning to act independently on the internet. Developers are creating "browser assistants" that can visit websites, click buttons, fill out forms, and collect data without human intervention. And, according to my data, Grok has become the leader in this race.

A comparative analysis of five leading AI tools, conducted by Browser Use founder Gregor Zunic, revealed a clear leader. The criterion was strict: the frequency of successfully completing tasks on a website without a single error. Here, Grok, based on the Grok 4.5 model, showed an impressive result — 76.42% of tasks correctly completed.

Who's Who in the Browser Agent Ranking

For a clearer picture, here is the full breakdown. In second place with a result of 72.70% is the assistant from OpenAI on the GPT 5.5 model. Third and fourth places are occupied by two tools from Anthropic on the Opus 4.8 model — 70.75% and 67.92% respectively. Closing the list is GPT via the OpenClaw plugin with a score of 60.38%. The difference between the leader and the outsider is nearly 16 percentage points, which in the world of high-precision algorithms is a chasm.

It is important to emphasize that all assistants operated under identical conditions. The only variable was the AI model itself, which makes the decisions. It was the assistant's "brain," its ability to analyze context and choose the correct sequence of actions, that determined the outcome. The results clearly demonstrate that the choice of model is critical.

What This Means for the Market

Interestingly, even with the same developer — Anthropic — the difference in effectiveness between the two connection options was nearly three points. This suggests that the integration architecture and method of connecting to the browser are no less important than the neural network itself. We see that Grok 4.5 didn't just win; it did so by a significant margin, which calls into question the positions of competitors in the agentic AI segment.

My analysis: The market for autonomous browser agents is rapidly maturing. Grok 4.5 has proven its worth, but I wouldn't rush to count out OpenAI and Anthropic. The AI arms race is just beginning, and each new training cycle could radically change the balance of power. However, for users, this is an unequivocal plus — competition will accelerate the emergence of truly useful digital assistants.