How Grok Build handles this
You are in the terminal with Grok Build, partway through a change, and want to know where the site stands. Ask it, and it calls get_score_history through the Hub and returns the stored runs, with the score, LCP, FCP, CLS, TBT, Speed Index and TTFB for mobile and desktop. If you name two sites it reads both and sets them side by side, which helps when one of them is the control you compare against.
For a fresh run it calls run_speed_test, a PageSpeed Insights run from the Hub, mobile or desktop, taking up to about two minutes. The Hub saves the result in its own report history, the Reports tab of the dashboard, which get_score_history may not read. run_benchmark is the quick option: it times the home page cached and uncached, with no Lighthouse score, and get_benchmark_history shows past benchmark runs so you can see whether a settings change coincided with a better time. Name the tool you want when you ask, so the answer is a score or a timing and never a blend.
Grok Build keeps the sign-in it receives from the Hub on your machine, in its own credentials store, so you sign in once rather than on every session. A written test plan helps here: say how many runs you want and for which device, and have it stop there. It makes the session predictable and keeps your use of the shared allowance small.
Set up Grok Build once
Already connected? Skip to the prompts. Alternatives and troubleshooting are on the Grok Build guide.
1Put your sites in xSpeed Hub
Sign in at app.xspeedcache.com with Google or email; the Hub is free and has no site cap. Then connect each WordPress site from its own dashboard: click Connect Hub in the xSpeed Cache top bar, then Connect via xSpeed Hub. Each site needs the free xSpeed Cache plugin.
2Add xSpeed Hub with grok mcp add
Run this once. Add --scope project to write the server to .grok/config.toml in the current directory instead of your user config, which is only sensible if the repo never holds a token. Grok Build handles the OAuth flow for a remote server on its own.
grok mcp add --transport http xspeedhub https://app.xspeedcache.com/xspeed/mcp
3Sign in from the TUI
Start grok and open /mcps. Select xspeedhub and press i to authenticate, or just use a Hub tool and let the browser flow start on first use. Sign in to xSpeed Hub and approve access. There is no token to paste; Grok Build stores the credentials it receives in ~/.grok/mcp_credentials.json. If the connection fails, run grok mcp doctor xspeedhub.
/mcps
Full Grok Build setup, sign-in options and FAQ
Before you send a prompt that changes something
Reads change nothing on your sites, though contact_support emails xSpeed support. Writes run once your connection allows them, so your client's approval prompt and a read-only connection (the token's Read-only everywhere switch, or a Viewer sign-in) are the gates that matter. Grok Build starts in Ask mode, which prompts for anything a rule has not already allowed, so a Hub write tool prompts until you say otherwise. You can allow the reads with a rule such as MCPTool(xspeedhub__get_*) in the permission section of your config, and leave the writes unmatched. Evaluation runs deny, then ask, then allow. Two other modes remove the prompt: Always-approve, from /always-approve, Ctrl+O or --always-approve, approves tool calls unless a deny rule or a hook stops them, and Auto lets a classifier approve tools it judges safe. If a Hub write must never run unprompted, add a deny rule. The Hub itself has no confirmation step.
What do I ask?
Three prompts written for Grok Build. More for this job are below.
Ask xSpeed Hub for the PageSpeed history of shop and list the last three mobile scores with their dates.
Use xSpeed Hub to run one desktop test on blog and tell me how it compares with the previous desktop run. Stop after that one.
With xSpeed Hub, show the benchmark history for shop and say whether response time changed after the settings change I made on Monday.
What happens, step by step
PageSpeed Insights is Google's Lighthouse audit, and its score and Core Web Vitals are the numbers most people mean by "my speed". Through xSpeed Hub an agent can read the scores a site has already stored, run a new PageSpeed test on mobile or desktop, and read or run a separate benchmark of cached against uncached response time. The three answer different questions, and the Hub tells the agent which to use for each.
01
Read before you test
get_score_history returns the stored PageSpeed Insights and GTmetrix runs for a site: the score per run, with LCP, FCP, CLS, TBT, Speed Index and TTFB, mobile and desktop. It starts nothing. The Hub tells the agent to use it whenever you ask what your score is, or whether it improved, because a new test spends a shared allowance.
02
Run a new test when you need one
run_speed_test runs a real PageSpeed Insights audit on the site from the Hub and returns the score, the Core Web Vitals and the top opportunities. strategy is mobile (the default) or desktop, and one call tests one of them. It can take up to about 2 minutes, so the agent should say it is running before it calls.
03
Test the other device
Mobile and desktop scores differ, often by a lot. For both, the agent makes two calls, one per strategy. A failure is reported with its reason: PageSpeed not configured, quota reached, a timeout, or a site the audit cannot reach.
04
Measure the cache on its own
run_benchmark times the home page with the cache bypassed and again served from cache, and returns the response times. It is a read tool, and it accepts site: "all" for every site. It does not return a Lighthouse score, so the agent must not use it to answer a PageSpeed question.
05
Compare over time
get_benchmark_history returns past benchmark runs, uncached against cached, together with the settings changes on the same timeline, which is how you answer "did that change help?". For PageSpeed scores over time, use get_score_history.
Reference
| Read stored scores | get_score_history (read): PageSpeed Insights and GTmetrix runs, starts nothing |
|---|---|
| Metrics per run | Score, LCP, FCP, CLS, TBT, Speed Index, TTFB, for mobile and desktop |
| Run a new test | run_speed_test (write, uses the shared allowance): one PageSpeed Insights run per call |
| strategy | mobile (default) or desktop |
| Duration | Up to about 2 minutes per run |
| Why a test is a write | It spends a shared PageSpeed allowance and stores a report; it changes nothing on the site |
| Read-only access | A connection token with Read-only everywhere on, or a Viewer member's sign-in: run_speed_test is refused; get_score_history, run_benchmark and get_benchmark_history still work. An OAuth sign-in gets the scopes the client asks for |
| Benchmark | run_benchmark (read): cached against uncached response time for the home page, no Lighthouse score |
| Benchmark history | get_benchmark_history (read): past runs plus settings changes on the same timeline |
| Scheduling | The Hub does not schedule PageSpeed tests itself; its daily pass only reads reports the plugin already stored |
| Where a new test is stored | In the Hub's own report history, the one the Hub dashboard shows |
Rules worth keeping
- Check the history first. Ask for stored scores before starting a test, and run a new one only when you want a fresh measurement or something changed.
- Never answer a PageSpeed or Lighthouse question with run_benchmark. It times your own cache and carries no score.
- run_speed_test changes nothing on your site, but the Hub classes it a write because it spends a shared allowance. A connection token with Read-only everywhere on, or a Viewer sign-in, refuses it. An OAuth sign-in gets the scopes the client asks for, so there your client's approval prompt is the gate.
- Lab scores move a few points between runs of the same page. The agent should compare like with like, same site, same strategy, and not report one run's difference as a trend.
- A new test does not run on a timer. If you want a regular test, the agent has to be started by something else, such as a scheduled job in your client, and a scheduled run that spends the allowance should be one you are comfortable leaving unattended.
Good to know with Grok Build
Grok Build starts in Ask mode, which prompts for any tool call no rule has allowed, so run_speed_test asks until you say otherwise. Always-approve, from /always-approve, Ctrl+O or the always-approve flag, skips the prompt for it too, and Auto lets a classifier approve what it judges safe. Allow the reads with a rule such as MCPTool(xspeedhub__get_*), leave run_speed_test unmatched, and add a deny rule if it must never run unprompted, because deny rules still apply under Always-approve. The Hub runs the call as soon as the connection allows writes, so you can also state the number of runs in each request or use a read-only connection (the connection token with Read-only everywhere on, or a Viewer sign-in).
More prompts for this job
They work in any client connected to xSpeed Hub.
Using xSpeed Hub, what is the mobile PageSpeed score for shop? Check the stored history before you run anything.
Use xSpeed Hub to run a new mobile PageSpeed test on blog and tell me the score, LCP and the top opportunities.
With xSpeed Hub, test shop on desktop and mobile and tell me which one is worse.
Using xSpeed Hub, has the PageSpeed score on docs improved over the last few runs?
Use xSpeed Hub to benchmark the cache on every site and tell me which one saves the least time.
Did turning on minification on shop help? Use xSpeed Hub to compare the benchmark history around that change.
My PageSpeed test failed on staging. Use xSpeed Hub to tell me why.
Frequently asked questions
Keep going
Track PageSpeed scores with other agents
More with Grok Build
Documentation
- How to run a speed test
- PageSpeed Insights integration
- Running speed tests in xSpeed Hub
- How to run a PageSpeed audit
- How to write prompts for xSpeed Hub
- Grok Build + xSpeed
- Every AI agent that works with xSpeed
From the blog
- Reading a PageSpeed Report in 2026: 30% of the Score Is a Metric Google Does Not Rank You On
- 45% of the WordPress Sites Our Scan Calls Clean Are Failing Google in 2026: Lab Scores Against Field Data
- The Same Page Scored 61 and 73 Twelve Minutes Apart: How to Benchmark WordPress Caching Plugins in 2026