AI
Grok Isn’t Just the AI Inside X Anymore
Grok 4.5 is faster, cheaper and far more capable than the version most people still associate with Twitter. The bigger story may be what xAI is building around it.
For a long time, Grok was easy to categorize.
It was Elon Musk’s AI chatbot inside X: occasionally useful, intentionally less filtered than its competitors, and differentiated mostly by its direct connection to the social network.
That description no longer works.
With Grok 4.5, a rapidly expanding standalone application, persistent skills, connectors, automation and increasingly serious developer tools, xAI is turning Grok into a much broader AI platform.
And the model itself has become good enough that the comparison is no longer Grok versus “real” frontier AI.
Grok is now part of that group.
Grok 4.5 closes the performance gap
The most important change is straightforward: Grok has become legitimately competitive at difficult work.
xAI introduced Grok 4.5 in July as its strongest model for coding, agentic tasks and knowledge work. The company says the model was trained across coding, science, engineering and mathematics, with a particular focus on completing real-world engineering tasks efficiently.
Independent testing generally supports the idea that this is a major step forward.
Artificial Analysis currently gives Grok 4.5 an Intelligence Index score of 54 and places it among the leading proprietary models it has tested. The firm’s initial analysis found especially strong performance in coding and agentic knowledge work.
That matters because previous generations of Grok often felt like they were chasing the leaders.
Grok 4.5 feels different.
It is no longer necessary to make excuses for the model because it has X integration, personality or access to current information. The underlying reasoning engine is increasingly capable of standing on its own.
That does not mean it universally beats OpenAI, Anthropic or Google. Benchmark leadership changes depending on the task, reasoning setting and evaluation.
But that is almost beside the point.
The important development is that choosing Grok no longer necessarily means accepting a major intelligence penalty.
Efficiency may be Grok’s biggest advantage
Raw intelligence is only one part of the AI race.
The more interesting number may be how much computation a model requires to reach an answer.
xAI says Grok 4.5 averaged 15,954 output tokens per SWE-Bench Pro task in its testing, compared with 67,020 for Claude Opus 4.8 at its maximum reasoning setting — approximately 4.2 times fewer output tokens. The company also claims roughly twice the token efficiency of comparable leading models across its broader testing.
Even allowing for the usual caution around vendor-run benchmarks, the direction is important.
The next phase of AI will not be won purely by whoever can make the smartest model at any cost.
It will increasingly be about intelligence per dollar, intelligence per token and intelligence per second.
That is particularly important as AI moves from answering occasional questions to running agents that may perform dozens or hundreds of actions.
A model that uses substantially fewer tokens to complete the same task can become dramatically cheaper at scale.
And Grok’s API pricing makes that efficiency more interesting.
For requests below its long-context pricing threshold, Grok 4.5 currently costs $2 per million input tokens and $6 per million output tokens, with cached input priced lower. It supports a 500,000-token context window.
That puts xAI in a position to compete not only on model quality but on the economics of actually deploying AI.
Grok is starting to remember how you work
Another major change is less visible on a benchmark chart.
Grok is becoming persistent.
xAI introduced Skills earlier this year, allowing users to teach Grok recurring preferences, instructions and workflows that can carry across conversations. Instead of repeatedly explaining how you want something formatted or how a recurring task should be handled, those instructions can become reusable expertise.
That is an important distinction.
Traditional chatbots are excellent at individual conversations. The more useful assistants are beginning to accumulate context about how you work.
Memory becomes especially powerful when combined with connectors.
Grok can now connect to outside services and work with email, files and calendars directly from its web, iOS and Android applications. xAI has also been expanding Grok into productivity software including Google Workspace and Microsoft Office workflows.
The result is a product that increasingly looks less like a chatbot and more like an operating layer sitting across your information.
That is the direction the entire AI industry is moving.
Grok is moving there quickly.
The standalone Grok app is becoming the real product
This is where I think many users still misunderstand Grok.
Using Grok inside X and using the full Grok platform are increasingly different experiences.
Inside X, Grok has an enormous advantage: context.
It exists directly alongside the real-time conversation happening on the network. That makes it exceptionally convenient for understanding breaking news, reactions, narratives and whatever people are discussing at that moment.
For quick questions while using X, that integration is hard to beat.
But the standalone Grok experience is increasingly where xAI’s larger ambitions become obvious.
Grok is available on the web, iOS and Android, with conversations, settings and subscriptions synced across those platforms. The standalone environment supports file uploads, voice, image and video creation, connectors and other productivity features.
And as of July 22, Grok 4.5 itself has been rolled out across web, iOS, Android and X.
So I increasingly see the two versions this way:
Grok inside X is the discovery layer.
Standalone Grok is becoming the work layer.
That combination is unusual.
ChatGPT, Claude and Gemini have enormous distribution of their own, but none of them is embedded directly into a social platform with X’s stream of continuously changing public conversation.
Grok does not merely search the web.
Its relationship with X gives it a native doorway into what people are talking about right now.
That could prove to be a much bigger advantage than it initially appeared.
Grok is also moving beyond chat
The strongest evidence that xAI sees Grok as a platform is everything being built around the model.
Grok Build has developed into a dedicated coding agent, and xAI recently open-sourced the Grok Build harness. Developers can inspect how context is assembled, how tools are called and how the agent loop operates.
Grok can also run Automations.
Users can define jobs that execute on a schedule or in response to an email trigger, allowing Grok to perform recurring research or monitor workflows without requiring a new prompt every time.
Then there are integrations with Google Workspace, Microsoft 365, cloud platforms and developer infrastructure.
Taken individually, none of these features changes the AI market.
Taken together, they show where xAI is headed.
The chatbot is becoming an agent.
The agent is becoming a platform.

The price of entry is aggressive
Grok is also relatively easy to try.
xAI currently offers a free Grok tier, while its main SuperGrok subscription is listed at $30 per month and includes access to Grok 4.5, higher limits, connectors, Expert and image/video generation.
For developers, the usage-based API pricing may be even more significant.
As AI applications become increasingly agentic, the cost of running the model repeatedly matters far more than the cost of a single chat response.
That brings us back to efficiency.
A frontier model does not have to be number one on every benchmark to become extremely important.
It needs to be good enough, fast enough and inexpensive enough that people actually build things with it.
Grok 4.5 appears to be moving decisively in that direction.
The real Grok advantage might be distribution
OpenAI has ChatGPT.
Google has Search, Android and Workspace.
Microsoft has Windows and Office.
Anthropic has built an extraordinarily strong product around Claude and has become deeply embedded in developer workflows.
xAI has something different.
It has X.
That gives Grok immediate exposure to a massive stream of news, discussion, creators, companies, investors, developers and public conversation.
At first, that integration made Grok feel like a feature of X.
The relationship may ultimately work in the opposite direction.
X could become one of the primary distribution engines for Grok.
Someone sees a discussion.
They ask Grok what is happening.
They move into deeper research.
Eventually they use Grok to analyze a file, write something, build software or automate a workflow.
That funnel begins somewhere competitors cannot easily replicate: directly inside the conversation.
Grok has grown up
I don’t think Grok is simply “the edgy AI from Twitter” anymore.
That characterization is already outdated.
Grok 4.5 has brought xAI much closer to the frontier on raw model capability. Its pricing and token efficiency make it increasingly interesting for developers. Skills and connectors are giving the assistant persistence. Automations are pushing it toward autonomous work. And the standalone Grok application is evolving into something much larger than the chatbot embedded inside X.
There are still reasons someone might prefer ChatGPT, Claude or Gemini depending on the workload.
But the decision is becoming much less obvious.
That is the biggest change.
A year ago, the interesting question was whether Grok could catch the AI leaders.
In 2026, the more interesting question may be what happens now that it has.
Continue reading
The AI trade isn’t dead → [Here]
Apple’s AI tax: Why AI infrastructure is raising the cost of computing → [Here]
ChatGPT vs. Claude: where the frontier AI race stands → [Here]