ScreenshotNeo

BlogHow-to

How to Archive a Webpage with Webpage Screenshot and the Wayback Machine

Save a webpage to the Wayback Machine, verify its archive, and keep a separate screenshot when you need a visual record.

By the ScreenshotNeo team4 October 20267 min read

To preserve one webpage as it appears now, submit its full URL using the Wayback Machine’s Save Page Now form. When the save succeeds, keep the archive URL it returns and open it to check the capture. If you also need a visual record, save a separate screenshot: an archived page and a screenshot serve related but different purposes. A Save Page Now screenshot control could not be confirmed in the official documentation reviewed for this guide, so do not rely on it being available.

1. Save one page with Save Page Now

  1. Open the Wayback Machine Save Page Now form.
  2. Enter the complete URL of the page, including its scheme (such as https://), path, and any query string that identifies the page you want to preserve.
  3. Submit the URL and wait for the result. A successful save provides a permanent archive URL for that page.
  4. Copy the returned URL, then open it in a new tab to inspect the capture.

Save Page Now is a request to capture one page once. It does not start a crawl of the site, save the page’s outlinks, or schedule future snapshots. Internet Archive’s help documentation explicitly describes it as a single-page save. Internet Archive: Save Pages in the Wayback Machine

2. Check what the archive actually preserved

A successful save does not guarantee a perfect replay. Check the returned archive URL and, where accuracy matters, inspect the important parts yourself:

  • Confirm the page title and main content are present.
  • Check key images, styles, and embedded resources.
  • Try important links and interactive controls, while watching whether they remain within the archived capture or resolve to live content.
  • Record the capture date shown in the archive URL or Wayback interface. Keep the exact URL when sharing or citing the snapshot.

Some images, CSS, or scripts may not have been archived. JavaScript that depends on the original server may not work in replay, and a missing archived resource can cause a replayed link or image to use live content. For evidence or historical research, verify the resources that matter rather than assuming the replay is complete. Internet Archive: Using the Wayback Machine

3. Save from a browser with the official extension

If you archive pages regularly while browsing, the Internet Archive’s official browser extension provides a Save Page Now action for the current page and features for finding older or newer captures. Its project repository lists browser distribution links for Chrome, Edge, Firefox, and Safari. Availability and compatibility depend on the current browser store listing, so check that listing before installing. Wayback Machine browser extension project

The extension is a convenient way to submit the current page; it does not change the scope of Save Page Now. The request still concerns a single page, and capture quality still depends on whether the site and its resources can be accessed and archived.

4. Keep a separate screenshot when you need a visual record

An archive is useful when you want a dated URL that can replay the page and its captured resources. A screenshot is a visual file showing the rendered page at the moment of capture. Use both when you need a replayable reference and a quick visual record, or when a citation workflow specifically calls for an image.

To take a screenshot yourself, open the live page in your browser, use its built-in screenshot or full-page capture feature, and save the resulting image separately from the Wayback URL. Browser controls differ, so consult the help for your browser if you need a full-page image rather than the visible viewport. Keep the original page URL, screenshot file, and archive URL together if you need to document where the image came from.

Do not assume that Save Page Now currently includes a screenshot option. The official help reviewed for this guide explains saving a page and its resources, but does not document a screenshot control.

5. Understand the limits before relying on a capture

  • One URL at a time: Save Page Now does not save a whole website, its directories, or all linked pages. Internet Archive notes that the method saves a single page, not the whole site.
  • No scheduled recapture: A save is a one-time request, not a subscription to future crawls.
  • Access restrictions: Password-protected or otherwise inaccessible pages may not be captured. Robots exclusions, an owner’s removal request, site crawling restrictions, or SSL configuration problems can also prevent a save.
  • Incomplete replay: Missing resources, server-dependent JavaScript, and other page behaviors can make an archived replay differ from the original.
  • Live resources in replay: A missing archived image or link may resolve to live content. Verify important resources and the capture date when authenticity matters.

For recurring preservation across many pages at an organization, Internet Archive describes Archive-It as a paid subscription service with technical and web-archivist support. It is an optional institutional route to investigate, not a requirement for an individual one-time capture.

6. Troubleshooting

Symptom Likely cause What to do
The save fails or no archive URL is returned The site may block crawling, be inaccessible, or have an SSL configuration that interferes. Check that the URL opens for you without authentication, confirm it is the intended public URL, and retry later. If the site owner restricts crawling or has requested exclusion, a save may not be possible.
The page appears, but images or styles are missing Some resources were not archived or could not be fetched during the capture. Inspect the resource-dependent parts of the page and treat the replay as incomplete. Keep a separate screenshot if you also need a visual record.
A button, menu, or other interaction does not work The feature may rely on JavaScript or server behavior that is unavailable in archived replay. Use the archive as a preserved reference, not as a guaranteed working copy of the live application. Save a screenshot of the relevant state if needed.
A link or image shows current content instead of the old version The resource may be missing from the archive and resolving to the live site. Check its URL and capture date. Do not treat live-resolved content as proof that it was present in the historical capture.
Only the submitted page is available Save Page Now captures the requested page once; it does not follow outlinks by default. Submit each needed page separately, or investigate a web-preservation service for an organizational crawl.
The page is absent from the archive It may be private, excluded by robots rules, removed at the owner’s request, or never successfully captured. Try Save Page Now only for content you can access and are permitted to preserve. Check existing captures in the Wayback Machine and retain a browser screenshot if you need a record of a page you can view.

7. Cite and share the result

Share the exact archive URL, including its capture timestamp. When citing a page, cite the original page in the normal way and add the Wayback capture URL and capture date when relevant. Internet Archive notes that there is no established MLA format specific to Wayback Machine resources, so include enough information for readers to identify the original page and the archived version. Internet Archive citation and replay guidance

Or skip the browser setup

If you need a screenshot file from code, ScreenshotNeo is a website screenshot API: one GET request takes a URL and returns an image or PDF. It handles the browser capture for you. It is not a Wayback archive, so keep using Save Page Now when you need an archived page URL and replayable snapshot.

See the ScreenshotNeo API documentation for request options. Example request:

curl -G "https://api.screenshotneo.com/v1/shot" -d access_key=YOUR_API_KEY --data-urlencode url=https://stripe.com -o shot.webp

Or call it from Python:

import requests

r = requests.get(
    "https://api.screenshotneo.com/v1/shot",
    params={"access_key": "YOUR_API_KEY", "url": "https://stripe.com"},
    timeout=90,
)
r.raise_for_status()
open("shot.webp", "wb").write(r.content)

Or call it from Node.js:

const q = new URLSearchParams({ access_key: 'YOUR_API_KEY', url: 'https://stripe.com' });
const res = await fetch(`https://api.screenshotneo.com/v1/shot?${q}`);
if (!res.ok) throw new Error(`Screenshot request failed: ${res.status}`);
const bytes = new Uint8Array(await res.arrayBuffer());
await import('node:fs/promises').then(fs => fs.writeFile('shot.webp', bytes));

ScreenshotNeo removes cookie banners, newsletter popups, and chat widgets before the shot. Bot checks, blank pages, and failed loads are never billed. Its MCP server lets AI agents take screenshots, and 1,000 screenshots a month are free with no card; paid plans start at $5 for 3,000. Create a free ScreenshotNeo account.

8. Frequently asked questions

Does an archive URL prove that every part of a page was preserved?

No. Open the capture and check the specific content and resources you need. A returned URL is not a guarantee that every asset or interaction will replay.

Can I use a Wayback capture as a replacement for a screenshot?

They serve different purposes. The archive provides a dated URL and may replay captured resources; a screenshot is a visual image file. Save both if your use case needs both.

Will the Wayback Machine keep checking the page for changes?

No. Save Page Now is a one-time save of a specific page. It does not schedule future captures.

Can I archive a page I can see only after signing in?

Do not assume a private or authenticated page can be captured. Save Page Now may not be able to access it, and you should only preserve content you are permitted to save.

Does ScreenshotNeo create a Wayback archive?

No. ScreenshotNeo returns a screenshot or PDF through its API. Use the Wayback Machine when you need its archived page URL and replay.