Using the Wayback Machine for competitor research

By Owen Fisher·Updated 24 July 2026·8 min read

The Wayback Machine, run by the Internet Archive, stores dated snapshots of public web pages going back to 1996. It is the best free way to see how a rival's positioning and pricing changed over the years. Its limits are sparse snapshots, broken styling and no alerts. Archives look backwards; Stravue watches look forwards.

What is the Wayback Machine?

The Wayback Machine is a free archive of the public web, run by the Internet Archive, a non-profit that has been crawling and storing pages since 1996. Enter a URL at web.archive.org and it shows every snapshot it holds of that page, each one dated, going back years or sometimes decades. It now holds hundreds of billions of pages.

For competitor research this is close to a time machine. Every claim a rival has made on their homepage, every price they have listed, every product name they have retired: if the crawler caught it, it is still there, with a date on it.

What it is genuinely good for

  • Positioning history. Step through a rival's homepage year by year and the repositioning story tells itself: who the headline spoke to in 2021 versus who it speaks to now, which product became the lead, when "platform" replaced "tool".
  • Pricing evolution. Where snapshots exist, old pricing pages show tier names, price points and what was once free. A rival's climb upmarket is usually visible in three snapshots.
  • Settling debates. "When did they drop the free tier?" can burn twenty minutes of a meeting on recollection. A dated snapshot ends the argument in one.
  • Context before a teardown. Knowing where a rival came from sharpens your read of where they are. A quick history pass makes a stronger competitor teardown.

How to use it well

  1. 1

    Enter the exact URL, not just the domain

    Go to web.archive.org and paste the full address of the page you care about. Deep pages are archived separately from the homepage, so example.com/pricing has its own snapshot history. Searching only the domain shows only the homepage's past.

  2. 2

    Read the capture calendar

    The results page shows a timeline by year and a calendar of capture dates. Dense clusters mean frequent snapshots; long empty stretches mean gaps. Pick dates either side of when you suspect the change happened, and work inwards.

  3. 3

    Compare snapshots across dates

    Open two snapshots in two tabs and look at the same part of each page: the headline, the tier table, the navigation. Note the last date the old version appears and the first date the new one does. That window is your answer.

  4. 4

    Screenshot what you find

    Archived pages load slowly and links inside them break, so capture what matters as an image the moment you find it, with the snapshot date visible. Evidence you can drop into a deck beats a link someone has to go and load.

  5. 5

    Check each key page separately

    Homepages get crawled most. Pricing, product and signup pages get crawled far less, and sometimes not for months. Run the same calendar-and-compare pass on each URL you care about rather than assuming homepage coverage extends to the rest.

The honest limits

The Wayback Machine is a remarkable public resource, and it was never designed for competitive monitoring. Five limits matter in practice.

  • Sparse snapshots for smaller sites. Famous sites get crawled daily. A niche rival might get a handful of snapshots a year, and none at all for the pages you care about most.
  • Broken rendering. Stylesheets, images and scripts are often missing from a snapshot, so pages load half-styled or scrambled. The page looked fine at the time; the archive of it does not.
  • Desktop only. Snapshots archive the desktop page. What a rival's site looked like on a phone in 2023 is simply not recorded.
  • No alerts. The archive never tells you a page changed. It is where you dig after you already suspect something, not how you find out.
  • Gaps at exactly the wrong moment. The crawl schedule is not yours. The week a rival tested new pricing and rolled it back can fall entirely between two snapshots, as if it never happened.

The archive shows what the crawler happened to catch. The change you care about most is often in a gap.

Archives look backwards. Watches look forwards.

The Wayback Machine answers what a page used to say. It cannot help with the change a rival ships next quarter, and by the time you go digging, the gaps are already fixed in place. The forward-looking version of the same idea is a watch: choose the pages that matter, and have them checked on a schedule from now on.

A pricing page captured now, rather than dug out later. Archived versions of the same page often arrive half-styled, or not at all.

Stravue watches do this with full captures. Each check is a clean full-page screenshot, desktop and phone, and when something real changes you get the before and the after side by side, dated. From the day a watch is set up, you own the history of that page: no gaps, no broken styling, no missing phone version. How to set one up is covered in monitoring competitor website changes, and the quarterly competitor watch shows a year of it in practice.

The two are complements, not rivals. Use the Wayback Machine for history that predates you. Start a watch for the history you will need a year from now.

Frequently asked questions

Is the Wayback Machine free to use?

Yes. The Wayback Machine is run by the Internet Archive, a non-profit, and browsing its snapshots is free with no account required. There is also a free "Save Page Now" feature that archives a page on demand, which is useful when you want today's version of a competitor page preserved with an independent date on it.

How often does the Wayback Machine capture a website?

It depends on the site's prominence. Heavily linked sites can be captured many times a day, while smaller sites may get only a few snapshots a year. Deep pages such as pricing are captured less often than homepages. Coverage is decided by crawl scheduling, not by what researchers need, so gaps of months are common on niche sites.

Can you see old versions of a competitor's pricing page?

Often, yes. Enter the exact URL of the pricing page at web.archive.org, not just the domain, because each page has its own snapshot history. Expect thinner coverage than the homepage and some snapshots with broken styling. Where snapshots exist, they are dated, which makes them strong evidence for when tiers, prices or a free plan changed.

Why do archived pages look broken?

A snapshot only renders correctly if the crawler also captured the page's stylesheets, images and scripts at around the same time, and often it did not. Missing assets leave pages half-styled or scrambled. The content is usually still readable, so treat snapshots as evidence of what a page said rather than a faithful picture of how it looked.

Can the Wayback Machine alert you when a competitor's site changes?

No. It is an archive, not a monitor, and it never notifies anyone of anything. To learn about changes as they happen, use a monitoring tool. Stravue watches check chosen competitor pages on a schedule and surface real changes with dated before and after captures on desktop and phone, which builds the gap-free history the archive cannot.

Capture your first competitor in minutes.

Paste a site and what you’re after. Stravue shoots the pages that answer it, on desktop and phone, ready to mark up into a deck.

Get started

From $39 a month. Cancel anytime.

Using the Wayback Machine for competitor research · Stravue