5 min read
    Annie Veale

    And now you want to crawl my site too?!

    From a handful of search bots

    For years, most websites were built with a small, predictable set of crawlers in mind -- primarily Google and Bing. These bots indexed pages for ranking, followed established rules, and were relatively well-behaved. To dozens of AI crawlers, agents, scrapers, trainers Bots read, summarise, compare, and decide

    To a crowded layer of automated readers

    Today, sites are visited by dozens of bots with very different goals: • Search engines (Googlebot, Bingbot) • Social crawlers (LinkedInBot) that decide visibility and trust • Market intelligence bots (SemrushBot, AhrefsBot) mapping links, entities, and relationships • AI crawlers and agents reading content for summarisation, comparison, and training Many of these are not concerned with rankings at all.

    Have Questions About This Article?

    Ask Annie's AI assistant for personalized advice and deeper insights

    AI Assistant

    Hi! I'm Annie's AI assistant. I can help you with digital marketing dilemmas, freelance, contract and service queries as well as provide you with routes to informational resources and tools.

    The New Reality

    Modern bots don't browse websites -- they extract meaning. They read quickly, skip presentation, and rely on structure, consistency, and clarity. Websites are increasingly treated as inputs into other systems, not destinations in their own right. Why this matters When crawl control, rate limiting, or structure fails, the impact is rarely visible immediately. Instead, sites quietly lose: • Link previews • Entity recognition • Inclusion in automated summaries

    Associated Topics

    The entities, definitions, relationships and tools related to this article.

    Related questions