What is PetalBot and Why it Hammers B2B Sites


Key takeaways
|
If you've ever pulled a raw access log for a North American B2B site, you've seen PetalBot. It's often in the top five user agents by request volume, sitting right behind Googlebot and Bingbot, crawling pages your buyers will never find through the product it serves.
That's the whole story with PetalBot. Legitimate operator, polite behavior, zero return.
What is PetalBot?
PetalBot is the web crawler operated by Huawei to build the index for Petal Search, the default search engine on Huawei smartphones and tablets that ship without Google Mobile Services. It identifies itself with the user agent string PetalBot and links to a documentation page at aspiegel.com, the Huawei subsidiary that runs it. |
Huawei built Petal Search after US export restrictions cut its devices off from Google Play and Google Search. To fill that gap, Huawei needed its own web index, and PetalBot builds it by crawling the open web the same way Googlebot does.
The result is a crawler with a genuine purpose and a very narrow audience. Petal Search users are overwhelmingly Huawei device owners in China, parts of Asia, the Middle East, and pockets of Europe where Huawei phones sold well before the restrictions. If your buyers are IT directors and marketing VPs at enterprise technology companies in the US, the probability that one of them finds you through Petal Search is close to zero.
Why does PetalBot crawl so aggressively?
PetalBot is building a full web index from a relatively recent start, so it crawls broadly and frequently to catch up. It also tends to recrawl pages more often than mature search engines do. The result is a high request volume relative to the traffic Petal Search sends back, which is why it stands out in logs. |
Mature crawlers like Googlebot have decades of signals about which pages change often and which don't. They allocate crawl budget accordingly. PetalBot doesn't have that history, so it crawls more uniformly and more often. On a content-heavy B2B site with a large /learn library, that adds up to thousands of requests a month.
It's worth remembering that this is a small slice of total bot traffic. Imperva's 2025 Bad Bot Report put automated traffic at 51% of all web requests in 2024, and PetalBot is a rounding error inside that number. But it's a visible rounding error, and it's one of the easiest to remove.
Does PetalBot respect robots.txt?
Yes. Huawei documents that PetalBot honors robots.txt Disallow directives and the Crawl-delay directive. Blocking it in robots.txt is sufficient. You do not need a network-edge rule for PetalBot the way you do for Bytespider. |
This is the meaningful difference between PetalBot and Bytespider. Both are Chinese-operated crawlers that Western site owners frequently want to block. Bytespider ignores robots.txt and requires a Cloudflare WAF rule. PetalBot reads the file and stops. In my experience it's picked up robots.txt changes within a few days.
Should you block PetalBot?
For a B2B site selling to buyers in North America or Europe, yes. PetalBot feeds a search engine your buyers don't use, and it consumes server resources doing so. Blocking it costs nothing in visibility. The only reason to leave it allowed is if you sell into markets where Huawei devices have meaningful share. |
Run the same three-part test I use in the bad bots hub. Is it documented? Yes. Does it obey robots.txt? Yes. Does the product it feeds ever surface your brand to a buyer? For most B2B technology brands, no.
Two out of three makes it a judgment call rather than an obvious block, and the judgment comes down to your market. MQL Magnet's clients sell to enterprise technology buyers in the US and Western Europe. I block PetalBot. If you sell into the Gulf states or Southeast Asia, check your analytics for Petal Search referrals before you decide. On Wix, the bot blocking walkthrough shows where to look.
How do you block PetalBot?
Add a named block to robots.txt: User-agent: PetalBot followed by Disallow: /. Keep it separate from your User-agent: * group. Publish and verify at yourdomain.com/robots.txt. PetalBot will stop within a few crawl cycles. |
User-agent: PetalBot
Disallow: /
I stack PetalBot with the other SEO-spam and low-value crawlers in a single block, one user agent per line above a shared Disallow. The full list I run is in the bad bots hub, and the paste-ready Wix version is in the Wix walkthrough.
What you must not do is put PetalBot in the same block as Googlebot or Bingbot, or add it to a blanket AI crawler rule in Cloudflare. It has nothing to do with AI. The Cloudflare guide covers keeping those categories separate.
Does PetalBot affect AI visibility?
No. PetalBot feeds Petal Search only. It has no connection to Google, Bing, OpenAI, Anthropic, or Perplexity. Blocking it has no effect on your rankings in Google or Bing, on AI Overviews, or on whether ChatGPT, Claude, or Perplexity cite your content. |
In the Engine Optimization Matrix, PetalBot touches the SEO row only, and only for a search engine outside your market. The crawlers that matter for the other three rows are covered in the AI crawlers overview, and the case for keeping them is in the GPTBot and ClaudeBot posts.
If you want the cleanest signal on what's actually crawling you before you start blocking, the server log analysis guide is the place to start. Or book 30 minutes and I'll read the logs with you.
Frequently asked questions
What is PetalBot?
PetalBot is Huawei's web crawler. It builds the index for Petal Search, the search engine on Huawei devices that ship without Google services. Its user agent string is PetalBot and its documentation is hosted by Aspiegel, a Huawei subsidiary.
Who operates PetalBot?
Huawei, through its Irish subsidiary Aspiegel Limited, which runs Petal Search outside China.
Does PetalBot obey robots.txt?
Yes. Huawei documents that PetalBot honors Disallow directives and Crawl-delay. Blocking it in robots.txt is sufficient.
Why does PetalBot crawl my site so much?
Petal Search is a relatively young index, so PetalBot crawls broadly and recrawls frequently to catch up. It lacks the historical signals mature crawlers use to prioritize.
Should I block PetalBot?
If your buyers are in North America or Europe, yes. Petal Search has negligible share there. If you sell into markets with meaningful Huawei device share, check your analytics for Petal Search referrals first.
How do I block PetalBot?
Add User-agent: PetalBot and Disallow: / as a separate block in robots.txt, below your default rules. Publish and verify.
Is PetalBot malicious?
No. It's a legitimate, documented search engine crawler. It's low-value for most Western B2B sites, not malicious.
Does blocking PetalBot affect Google rankings?
No. PetalBot has no connection to Google, Bing, or any AI engine. Blocking it changes nothing about your visibility in those systems.
How is PetalBot different from Bytespider?
Both are Chinese-operated crawlers, but PetalBot obeys robots.txt and feeds a search engine. Bytespider is widely reported to ignore robots.txt and feeds ByteDance model training. Bytespider requires network-edge blocking; PetalBot doesn't.
Can I slow PetalBot down instead of blocking it?
Yes. Huawei documents support for Crawl-delay. Add Crawl-delay: 10 under the PetalBot user agent to space requests ten seconds apart.
About the author
Harold Bell is Founder and CEO of MQL Magnet, a B2B content marketing agency serving enterprise technology brands. He's the creator of the Engine Optimization Matrix (EOM), author of Partner Over Product, and a Forbes Communications Council member.



