CyberMax
Home › Use cases › Velbrake

Keep your articles cited in AI answers without feeding training crawlers

For: publishing

Publishers want AI answers to cite and link their stories, not to have the archive copied into training sets. Velbrake splits the bots: training crawlers are refused on the server, search and assistant bots keep reading, and impostors using a crawler's name are blocked.

How publishing teams use Velbrake

  1. Install the free plugin and leave the default policy on.
  2. Enter your host's price per GB to see each bot's cost over 30 days.
  3. With Pro, protect paths such as /premium/ or the archive from every bot you choose.
  4. Publish llms.txt from Pro to point AI assistants at your best pages.

If the site sits behind a full-page cache or CDN, add matching bot rules there; cached pages never reach WordPress.

Velbrake built-in bot list: GPTBot and ClaudeBot blocked, OAI-SearchBot and PerplexityBot allowed
Real output: Velbrake's built-in bot list and default policy (version 1.0.0)

Built-in bot list and default policy, Velbrake 1.0.0

BotOperatorGroupDefault policyIdentity check
GPTBotOpenAIAI trainingBlockedOpenAI IP list
OAI-SearchBotOpenAIAI searchAllowedOpenAI IP list
ClaudeBotAnthropicAI trainingBlockedAnthropic IP list
PerplexityBotPerplexityAI searchAllowedPerplexity IP list
BytespiderByteDanceAI trainingBlockeduser agent

Source: products/velbrake/store/listing.json (built-in bot list and default policy of Velbrake 1.0.0).

FAQ

Can I block one AI company completely?

Yes; switch each of its bots (training and search) to blocked.

Does it replace robots.txt?

No, it adds to it. Keep your robots.txt rules; Velbrake also enforces them on the server for requests that reach WordPress.

Velbrake: full guide with prices and alternatives · All use cases · Store