+Field test

September 1, 2026

I Gave an AI Agent One Job: Find My Competitors. Here Is the Raw Output.

Everyone demos MCP with a toy question. I ran one real use case through the Semrush MCP against a live vinyl wrap shop and pasted the unedited results. Fifteen competitors, ranked, in one tool call. Here is what it found and where it still needs a human.

01The test

One use case, one tool call, no cherry-picking

Most MCP demos are magic tricks. Someone asks the agent something vague, it spits out something plausible, everyone claps, nobody checks. I wanted the opposite. Pick one narrow job an agent should be good at, run it against a real business, and paste the output without editing the embarrassing parts.

The job: find the competitors. Not the ones the owner already knows about. The ones stealing search traffic while nobody is looking. I pointed an AI agent at the Semrush MCP and gave it a single vinyl wrap shop, wrapsredefined.com, and one instruction. Find who competes with this domain organically and rank them by how much they actually overlap.

No follow-up prompts. No coaching. One tool call. Here is what came back.

02The receipts

What the agent found

The agent pulled the organic competitor set and ranked it by relevance, which is Semrush's measure of how much two domains fight over the same keywords. The top of the list:

jekyllandhydes.com came back as the closest competitor, sharing 48 keywords. Then fortunewraps.com and wrapdaddysdfw.com, both wrap shops, both fighting over the same terms. So far, so obvious.

Then it got useful. supreme-wraps.com sits on 554 organic keywords pulling an estimated 1,820 monthly visits. 360wraps.com pulls around 1,606. zillawraps.com holds 418 keywords for about 787 visits. Those are not names a busy shop owner rattles off. They are the ones quietly outranking him on the searches that turn into a truck in the bay.

Fifteen competitors, ranked, in the time it takes to reheat coffee. The part that matters is not the speed. It is that the agent worked from live organic data instead of a hunch. Nobody guessed. The number came from a tool call I could re-run and check.

03The plumbing

Why this one is a good first test

Competitor discovery is the perfect use case to start with, because you can falsify it in five minutes. Pull the list, open the top three domains, and see for yourself whether they actually sell what your client sells. If the agent hallucinated, you will know before it costs you anything.

The reason it works is that the agent is not remembering anything. MCP, the Model Context Protocol, lets the assistant call live data directly, the same way an app calls an API. An AI model's memory of who competes with a wrap shop is a party trick. An agent pulling the real organic profile over the Semrush MCP is doing research. The difference is the whole ballgame.

A model on its own is a guy at the end of the bar who is confident about everything and right about half of it. Give him the actual data and make him show his work, and suddenly he is worth listening to.

04The catch

Where the agent stops and I start

The list is a starting point, not a strategy. The agent ranked by keyword overlap, which is honest but dumb. It cannot tell me that jekyllandhydes runs a different service mix, or that supreme-wraps is the real threat because it is winning the commercial terms that end in a sale, not the how-to searches that end in a closed tab.

That judgment is still mine. The agent hands me a ranked list in one call. I decide which of those fifteen is worth fighting, which to ignore, and which to quietly borrow a page structure from. The tool killed the busywork. It did not kill the thinking, and anyone selling you that part is selling you nothing.

My agency, Local Howl, runs discovery passes like this before every client teardown. The agent does the pulling. A human does the deciding. That order does not reverse.

05The playbook

Run this yourself before you trust it with anything bigger

If you want to test an MCP use case and actually learn something, do not start by automating a deliverable. Start with one read-only job you can check by hand. Competitor discovery is a good one. So is a keyword gap between you and one rival.

The stack is short. A Semrush plan with real data behind it, the MCP connection turned on, and an assistant that supports it. Semrush is what I run these tests against, because the organic and competitor data is deep enough that the agent has something real to work from. Point it at a domain, ask for the competitors, and check the answer against the interface. If the numbers match, you have found a workflow you can trust. If they do not, better to learn that on a test than on a client.

I wrote a longer piece on how these agent workflows run in production once you trust them. This one is the before. Run the small test first.

06The verdict

Test it small. Then decide.

One use case. One tool call. Fifteen real competitors, ranked, that the owner had never named. That is what a grounded agent buys you, and it is genuinely useful.

It is also not a strategy, a wheel you can take your hands off, or a reason to fire your analyst. It is a faster shovel. The hole still needs someone who knows where to dig. Test the shovel on something small, watch it work, and then decide how much of the digging you hand it.

You own a number that isn't moving. Let's move it.