AI search crawlers

ImagesiftBot

ImagesiftBot is a bot operated by ImageSift that indexes pages so an AI answer engine can cite them. Once images and text are downloaded from a webpage, ImageSift analyzes this data from the page and stores the information in an index. It is documented as honouring robots.txt, but no IP range list or reverse-DNS convention is published, so a request carrying this user agent cannot be proven genuine.

ImagesiftBot at a glance

Reference data for the ImagesiftBot crawler
User agentImagesiftBot

Token only. The operator has not published a full user-agent string, so match on the substring rather than an exact header.

OperatorImageSift
PurposeImageSiftBot is a web crawler that scrapes the internet for publicly available images to support their suite of web intelligence products
Respects robots.txtYes, documentedsource
Verification methodNone published, user agent only

Crawls for images rather than text, so a text-only robots.txt review can miss it entirely.

https://imagesift.com/about
Crawl patternNot published
Robots.txt tokensImagesiftBot
Operator documentationhttps://imagesift.com/about

Once images and text are downloaded from a webpage, ImageSift analyzes this data from the page and stores the information in an index. Their web intelligence products use this index to enable search and retrieval of similar images.

robots.txt

Allow or block ImagesiftBot

Paste one of these into the robots.txt file at the root of your domain. Rules are per token, so a block on one crawler leaves every other bot untouched.

Block ImagesiftBot
User-agent: ImagesiftBot
Disallow: /

Blocks every path for this crawler only.

Allow ImagesiftBot
User-agent: ImagesiftBot
Allow: /

Explicit allow. Useful when a wildcard rule above it would otherwise catch this crawler.

Allow, except the pages you never want quoted
User-agent: ImagesiftBot
Allow: /
Disallow: /account/
Disallow: /checkout/
Disallow: /search

Edit the Disallow paths to match your own account, checkout and search URLs.

What blocking actually costs you

Blocking ImagesiftBot takes the site out of the index behind its answers, so it stops being cited there. If referral traffic from AI answers matters to you, block the training crawler instead and leave this one alone.

Block every crawler in the directory at once
Verification

Is that really ImagesiftBot?

A user-agent header is a string the client chooses. Scrapers copy ImagesiftBot precisely because site owners allow it. Crawls for images rather than text, so a text-only robots.txt review can miss it entirely.

Paste the IP address from your access log below. You will see whether it belongs to a hosting provider, a residential proxy pool or a Tor exit, plus the ASN that announces it, which is what tells you whether the claim holds up.

Related crawlers

Agentscan

Identify crawlers automatically instead of by hand

Agentscan checks a request against published crawler ranges, reverse DNS and hosting data, then returns a verdict your edge can act on. One call per request, no range files to keep up to date.

FAQ

ImagesiftBot questions

ImagesiftBot is a bot operated by ImageSift that indexes pages so an AI answer engine can cite them. Once images and text are downloaded from a webpage, ImageSift analyzes this data from the page and stores the information in an index. It is documented as honouring robots.txt, but no IP range list or reverse-DNS convention is published, so a request carrying this user agent cannot be proven genuine.