Rankealo

Zhipu AI · Training crawlers

ChatGLM-Spider

Zhipu’s ChatGLM public-content spider. Treat it as a training-style crawl unless Zhipu documents a separate search UA.

Optional training opt-out
Operator
Zhipu AI
Traffic type
Training crawlers
Verification
Not independently verifiable
robots.txt
Compliance not documented

Usually means: Bots that collect public pages for future models. Blocking them opts you out of training. It does not, by itself, remove you from live AI search citations.

User-agent

robots.txt and most WAF rules match the token, not the full Mozilla string. Operators often wrap ChatGLM-Spider in extra product or version text.

ChatGLM-Spider

What ChatGLM-Spider does

ChatGLM-Spider collects public documents for Zhipu AI systems. It is not a Google ranking crawler.

How to get discovered

If you want Zhipu products to read you on demand, allow the spider or any documented user-fetch token. Otherwise Disallow and enforce at the edge if needed.

Why ChatGLM-Spider might skip you

Default unknown-bot blocks, or robots files that never mention Zhipu tokens.

robots.txt rule

This snippet opts out of training-style collection. Keep the search and user-fetch tokens from the same operator allowed if you still want citations.

# Optional training opt-out — does not remove you from AI search by itself
User-agent: ChatGLM-Spider
Disallow: /

Zhipu AI has not published a machine-readable range file for this token. Do not treat the user-agent alone as proof of identity.

JavaScript: reads raw HTML only.

ChatGLM-Spider FAQ

What is ChatGLM-Spider?

ChatGLM-Spider collects public documents for Zhipu AI systems. It is not a Google ranking crawler.

Should I allow ChatGLM-Spider in robots.txt?

Only if you want to opt out of training-style collection. Blocking ChatGLM-Spider does not, by itself, remove you from live AI search citations — those use the search and user-fetch tokens from Zhipu AI.

How do I know a request is really ChatGLM-Spider?

Zhipu AI has not published a range file for this token. Treat the user-agent as a claimed identity, and do not build a hard IP allowlist from blog posts or screenshots.

Letting ChatGLM-Spider in is the start. Getting cited is the job.

Rankealo checks whether AI crawlers can reach you, then publishes pages built to be retrieved and quoted in ChatGPT, Claude, Perplexity, and Gemini.

Reading is step one. Measuring your AI visibility is step two.

Rankealo tracks how often your brand is mentioned and cited across the major AI engines, then helps you publish the pages that close the gaps.