Bot directory · Search engine crawler

DuckDuckBot

DuckDuckGo's modest crawler, supplementing the licensed indexes behind its privacy-focused search.

OperatorDuckDuckGo
CategorySearch engine crawler
Respects robots.txtYes, per operator documentation
robots.txt tokenDuckDuckBot
User-agent stringDuckDuckBot/1.0; (+http://duckduckgo.com/duckduckbot.html)

Official documentation: duckduckgo.com

What DuckDuckBot does

DuckDuckBot crawls to improve DuckDuckGo results alongside data licensed from partners. It is documented, low-volume, and honors robots.txt.

Why it visits your site

Filling gaps and freshening results in DuckDuckGo's index.

The bigger picture: search engine crawlers

Search crawlers remain the one category with a decades-old, well-understood bargain: they consume crawl budget and return ranked visibility. Everything about them is comparatively mature - published IP ranges, verification methods, granular directives, webmaster consoles. Blocking them is almost always self-harm, which is precisely why they need to be cleanly separated from AI crawlers in your data: a blanket bot block that catches Googlebot costs you your organic channel.

The complication the AI era added is that search indexes now feed AI features too - AI Overviews draw on Google's index, Copilot on Bing's. Operators have answered with control tokens that split AI use from search use of the same crawl. The practical upshot: your lever for AI concerns is usually a token, not a block on the search crawler itself.

Should you block it?

No practical reason - volume is tiny and the search traffic, while smaller than Google's, is real.

Block via robots.txt

Add these lines to the robots.txt at your site root to ask DuckDuckBot to stay away:

User-agent: DuckDuckBot
Disallow: /

Remember that robots.txt is a request, not a lock - compliance is voluntary, and enforcement happens at the server. Viz's AI Control pairs robots.txt rules with server-level 403 responses and then probes your site to verify the block is actually holding.

How to see DuckDuckBot traffic on your site

You have three windows onto it, each with a catch. Raw hosting access logs contain every DuckDuckBot request, but many managed WordPress hosts don't expose them, and when they do you are grepping text files by hand. A CDN dashboard sees the traffic at the edge, but it lives outside WordPress and speaks in totals, not in your site's terms. And JavaScript analytics - GA4 and friends - will never show it at all, because DuckDuckBot doesn't run scripts.

Viz's answer is the request log: filter to DuckDuckBot and read exactly which URLs it fetched and when, right inside wp-admin - then watch the same filter after any blocking decision to see whether the visits stopped, slowed, or kept coming.

Measure first

Is DuckDuckBot on your site right now? Find out in two minutes.