What it means to download a website
Downloading a website means saving the HTML, images, stylesheets, and other files from a live web address to your computer so you can view it without an internet connection. The result is a folder on your hard drive containing the website's structure and content, usually viewable by opening a single HTML file in your browser.
This is different from saving a single page. When you download an entire website, you're capturing multiple linked pages and their resources so navigation between them works offline. The depth of what you capture depends on your tool and settings — you might grab just the homepage and one level of links, or you might go several layers deep.
Key Takeaways
- Desktop tools like HTTrack, Wget, and Cyberdog let you download full websites to your computer by specifying a starting URL and depth level.
- Browser extensions like SingleFile and DownThemAll work from within your browser and are simpler for downloading individual pages or small site sections.
- Downloaded sites work best when viewed locally through your browser rather than by opening files directly, because relative links and resources load more reliably.
- Large websites with thousands of pages can take hours to download and consume significant disk space, so setting depth limits prevents runaway downloads.
- Respect the website owner's terms of service and robots.txt file, which may restrict automated downloading of their content.
Desktop tools for full-site downloads
HTTrack is the most widely used free tool for downloading entire websites. You give it a starting URL, set how many levels deep to follow links, and it mirrors the site to a folder on your computer. HTTrack runs on Windows, Mac, and Linux, and includes a graphical interface so you don't need to type commands. It respects robots.txt by default, meaning it won't download from sites that have explicitly forbidden automated access.
Wget is a command-line tool that does the same job but requires typing commands into a terminal. It's more powerful for advanced users because you can write scripts to download multiple sites or set precise rules about what to grab. Wget comes built into most Linux systems and Mac computers; Windows users need to install it separately.
Cyberdog is a simpler graphical alternative that works on Windows and Mac. It's less configurable than HTTrack but easier to learn if you're new to downloading websites. You paste a URL, choose your depth level, and start the download.
All three tools create a folder structure that mirrors the website's organization. When the download finishes, you open the main index.html file in your browser to start browsing offline.
Browser extensions for smaller downloads
SingleFile is a browser extension for Chrome and Firefox that saves an entire webpage — including all images, stylesheets, and scripts — into a single HTML file. This works well for individual pages or articles you want to keep, but it doesn't follow links to other pages automatically. You run it on each page you want to save.
DownThemAll is a Firefox extension that lets you download all files matching a pattern from a website. You can tell it to grab every image, every PDF, or every link on a page. It's useful for downloading a specific type of content rather than the whole site structure.
Extensions are faster to set up than desktop tools and don't require learning command-line syntax. The trade-off is that they're designed for smaller jobs — downloading a 500-page website with a desktop tool is practical, but doing it with an extension would mean running it hundreds of times.
How to set download depth and size limits
The most common mistake when downloading a website is letting it run too deep. If you tell HTTrack to follow links five levels down from the homepage, it might end up downloading thousands of pages you didn't intend to capture. Most sites have navigation menus that link to every page, so going deep creates exponential growth.
Before you start, decide how much of the site you actually need. If you want the main content, set depth to 2 or 3 levels. If you want everything, set a file-size limit instead — for example, "stop when you reach 1 gigabyte" — so the download doesn't run indefinitely. You can also exclude certain file types (like video) or certain URL patterns (like "/admin" or "/search") to keep the download focused.
HTTrack's interface lets you set these limits in the options panel before you start. Wget requires command-line flags like -l 3 for depth or -Q 500m for a 500-megabyte limit. Check the tool's documentation for the exact syntax.
Why downloaded sites sometimes don't work perfectly
Downloaded websites usually work well for reading content, but interactive features often break. Forms, search boxes, and buttons that send data to a server won't function because there's no server to receive the request. Videos hosted on external platforms (like YouTube embeds) won't play unless you also downloaded the video files themselves.
Links sometimes point to the wrong place if the original site used absolute URLs (like "https://example.com/page") instead of relative URLs (like "/page"). When you download, absolute URLs still point to the live internet, while relative URLs correctly point to your local files. Most download tools convert these automatically, but it's worth checking a few links after the download finishes.
JavaScript-heavy sites that load content dynamically may not work at all offline, because the downloaded HTML doesn't include the data that JavaScript would normally fetch from a server. Static sites with simple HTML and CSS download and work much more reliably.
Viewing your downloaded site locally
After the download finishes, open the main index.html file in your web browser — don't try to open it by double-clicking the file in your file manager. Using your browser ensures that relative links work correctly and resources load from the local folder instead of the internet.
If you're on Mac or Linux, you can also start a simple local web server from the downloaded folder's directory. This is more reliable for complex sites because it mimics how a real web server delivers files. In the terminal, navigate to the folder and run python3 -m http.server, then visit http://localhost:8000 in your browser.
Keep the downloaded folder intact — don't move individual files around or rename the folder structure, or links will break. If you want to move it to a different computer, copy the entire folder as-is.
Legal and ethical considerations
Before downloading a website, check its terms of service. Some sites explicitly forbid automated downloading or scraping. The robots.txt file in a site's root directory (like "example.com/robots.txt") tells automated tools what they're allowed to download — most download tools respect this by default, but you should verify your tool does.
Downloading a site for personal offline reading is generally considered fair use. Downloading a site to republish it, sell it, or present it as your own is not. If the site contains copyrighted material, downloading doesn't change your legal obligations around how you can use that material.
Large automated downloads can strain a website's server, especially if you set the tool to download very quickly. Most download tools include delays between requests to be respectful. If you're downloading a small personal blog, the impact is negligible. If you're downloading a large commercial site, consider doing it during off-peak hours.
Frequently Asked Questions
Can I download a website that requires a login?
Most download tools can't handle login screens automatically. HTTrack has an option to log in first by opening your browser, completing the login, and then starting the download from an authenticated session. Wget can pass cookies if you provide them. For sites behind paywalls or strict authentication, downloading usually isn't practical.
How much disk space does a typical website download take?
A small site with mostly text might be 10 to 50 megabytes. A medium site with images could be 500 megabytes to 2 gigabytes. A large site with video or thousands of pages can easily exceed 10 gigabytes. Set a size limit before you start so you know what to expect.
Will my downloaded site still work if I move it to a different folder?
Yes, as long as you move the entire folder together. The internal links are relative, so they'll still point to the right files. If you move individual files out of the folder or rename the folder structure, links will break.
What's the difference between downloading a website and taking a screenshot?
A screenshot captures only what's visible on your screen at one moment. Downloading captures the entire HTML structure and all linked pages, so you can navigate and search through the full site offline. Screenshots are useful for preserving the visual appearance of a single page; downloads are useful for keeping a working copy of a whole site.
Can I download a website that uses a lot of JavaScript?
Download tools grab the HTML and static files, but they don't execute JavaScript the way a browser does. If a site loads most of its content dynamically through JavaScript, the downloaded version will be mostly empty. Static sites with simple HTML and CSS download much more successfully.