PanguBot
Huawei's AI training crawler, collecting web content for its PanGu family of large language models.
Reference reviewed by Sona
| Operator | Huawei |
|---|---|
| Powers | PanGu large language model training |
| Purpose | Model training |
| User-agent token | PanguBot |
| Respects robots.txt | Mostly |
PanguBot gathers publicly available web content used to train Huawei's PanGu multimodal large language models, which power AI features across Huawei's cloud and device products.
PanguBot is tracked in community crawler lists with its own user-agent token; Huawei's own documentation for it is thinner than for PetalBot, so treat robots.txt compliance as expected but not guaranteed.
Allow PanguBot
Your content can be represented in Huawei's PanGu models and the AI features they power.
User-agent: PanguBot Allow: /
Block PanguBot
Keep your content out of Huawei's training data - consider server-level enforcement given the limited documentation.
User-agent: PanguBot Disallow: /
Can PanguBot read your page right now?
Test any URL and see exactly what AI crawlers receive.