Buzzbot Web Robot

A summary of the Buzzbot Internet robot. Including details for the owner, description, HTTP user agent and whether this robot adheres to the robot exclusion standard.

Who owns the Buzzbot robot? Is it a good or a bad robot? And why is it visiting your website?

Shown below is a sample log file entry for the Buzzbot web robot. It’s derived from an Apache web server log file. From the log entry information about how the robot identifies itself, HTTP User Agent, and where it is hosted are given.

Server Log File

vntweb.co.uk 54.226.118.197 - - [29/Mar/2019:08:07:55 +0000] "GET /robots.txt HTTP/1.0" 200 117 "-" "Buzzbot/1.0 (Buzzbot; http://www.buzzstream.com; buzzbot@buzzstream.com)"

HTTP User Agent

Buzzbot/1.0

IP Addresses

The observed IP address was 54.226.118.197.

WHOIS DNS command gives the following information about the IP address:

NetRange:54.224.0.0 – 54.239.255.255
OrgName:Amazon Technologies Inc.
Address:410 Terry Ave N.
City:Seattle
PostalCode:98109
StateProv:WA
Country:US
Updated:2017-01-28

As can be seen from the above the observed IP address is a part of a block assigned to Amazon Technologies Inc.

Owner

BuzzStream

Country

USA

Exclusion

The user-agent string includes a reference to the website http://www.buzzstream.com.

The referenced website doesn’t provide details about the that Buzzbot and whether it supports robots exclusion text and also obeys the crawl delay.

The robots.txt exclusion standard, which was described at http://www.robotstxt.org/wc/exclusion.html#robotstxt, but is not currently available. Information is available on the same website https://www.robotstxt.org/robotstxt.html and also on the w3c website at https://www.w3.org/TR/html4/appendix/notes.html#h-B.4.1.1

If you wish to prevent the robot from indexing a website and to adjust its crawl rate.

You may wish to try including the following entry in the robots.txt file to prevent Buzzbot from visiting your site

User-agent: Buzzbot
Disallow: / 

Also to control the frequency of Buzzbot visiting your site, a minimum acceptable delay between consecutive requests, try add the following to the robots.txt file:

User-agent: Buzzbot
Crawl-Delay: 10

In this example the delay has been set to 10 seconds.

As is common with website crawlers there is a delay between changes made to the robots.txt file and the change being implemented.

Take care making changes to the robots.txt file. A misunderstanding in configuration or an error in configuration can lead to important search engines excluding your website.

Further Info

Buzzstream offers to

“Research influencers, manage your relationships, and conduct outreach that’s personalized and efficient. “

It’s services include:

  • Research influencers automatically – Buzzstream provides contacts information social profiles and site metrics.
  • Keep track of all your conversations – saving emails and tweets and allowing the setting of reminders allowing you to keep rack of conversations.
  • Let your team focus on the important stuff – a centralised database for the entire team, affording collaboration within the group.
  • Find out what works and what doesn’t – with insights into your outreach campaigns.

Buzzstream’s software is used for;

  • digital PR
  • link building
  • content promotion