---
Type: desktop-application
ID: com.httrack.WebHTTrack
Package: webhttrack
ProjectLicense: GPL-3.0-or-later
Name:
C: WebHTTrack Website Copier
Summary:
C: Copy websites to your computer for offline browsing
Description:
C: |-
<p>WebHTTrack is the web interface to HTTrack, an offline browser utility. It downloads a website from
the Internet to a local directory, fetching the HTML, images, and other files and rebuilding the site's
link structure so you can browse it offline.</p>
<p>A step-by-step web interface guides you through choosing the addresses to mirror and the options to
apply. Mirrors can be updated in place and interrupted downloads resumed.</p>
<p>Typical uses include:</p>
<ul>
<li>Keeping an offline copy of a website for reading without a connection</li>
<li>Archiving or preserving sites and capturing them for later reference</li>
<li>Updating an existing local mirror without downloading it again</li>
</ul>
Developer:
id: com.httrack
name:
C: Xavier Roche
Categories:
- Network
Keywords:
C:
- offline browser
- website copier
- mirror
- crawl
- archiving
Url:
homepage: https://www.httrack.com/
bugtracker: https://github.com/xroche/httrack/issues
Icon:
cached:
- name: webhttrack_httrack.jxl
width: 48
height: 48
- name: webhttrack_httrack.jxl
width: 64
height: 64
- name: webhttrack_httrack.jxl
width: 128
height: 128
remote:
- url: com/httrack/WebHTTrack/e98feee878e625288dc2eb9c4773e336/icons/128x128/webhttrack_httrack.jxl
width: 128
height: 128
stock: httrack
Launchable:
desktop-id:
- WebHTTrack.desktop
Screenshots:
- default: true
caption:
C: Choosing the addresses and options for a new mirror
thumbnails:
- url: com/httrack/WebHTTrack/e98feee878e625288dc2eb9c4773e336/screenshots/image-1_752x359@1.jxl
width: 752
height: 359
- url: com/httrack/WebHTTrack/e98feee878e625288dc2eb9c4773e336/screenshots/image-1_624x297@1.jxl
width: 624
height: 297
- url: com/httrack/WebHTTrack/e98feee878e625288dc2eb9c4773e336/screenshots/image-1_224x106@1.jxl
width: 224
height: 106
source-image:
url: com/httrack/WebHTTrack/e98feee878e625288dc2eb9c4773e336/screenshots/image-1_orig.jxl
width: 1024
height: 489
Releases:
- version: "3.50.5"
type: stable
unix-timestamp: 1790899200
description:
C: |-
<ul>
<li>Pages are no longer saved outside the mirror folder, whatever name a plugin or a cache file gives
them</li>
<li>Stopping a mirror is more reliable: Ctrl+C no longer hangs, and the time limit is kept even when
a host name lookup is stuck</li>
<li>A mirror no longer gains empty folders for pages that failed, and the next run removes temporary
files that a stopped run left</li>
<li>The size limit holds when several FTP downloads run at once</li>
<li>Stylesheets that use image-url() no longer make httrack download files the browser never loads</li>
</ul>
- version: "3.50.4"
type: stable
unix-timestamp: 1790208000
description:
C: |-
<ul>
<li>Mirrors of sites built on a JavaScript framework keep their modules, where the wrong address was
fetched and its error page was stored instead</li>
<li>A server asking the crawler to wait and come back is obeyed now, where the page used to be given
up at once. The Flow control page caps how long that wait may be</li>
<li>The log says when a site's bot protection refused the crawl, and when links above the starting
page were left out of the mirror</li>
<li>A Windows build no longer leaves a shortened message unfinished, which could stop the program</li>
<li>Connections can use Multipath TCP on Linux and macOS, so a mirror carries on when one network path
goes away</li>
</ul>
- version: "3.50.3"
type: stable
unix-timestamp: 1789689600
description:
C: |-
<ul>
<li>The capture proxy that takes a URL from your browser listens on this machine only, where any host
on your network could reach it before</li>
<li>A page saved under a name the crawled site chose can no longer run commands of its own through
the option that runs a program for each saved file</li>
<li>A site that writes a rule for HTTrack and then a rule for every crawler is read the way the standard
says, so pages it allows are no longer skipped</li>
<li>A crawl no longer stops when the file that reports its progress to the interface cannot be opened,
or when a damaged cache is repaired</li>
<li>More fixes to how the engine bounds what a hostile page or server sends it</li>
</ul>
- version: "3.50.2"
type: stable
unix-timestamp: 1788998400
description:
C: |-
<ul>
<li>Cookies exported from a browser are read correctly, so a mirror of a site that needs a login keeps
its session</li>
<li>Stopping a mirror keeps what a later resume needs, where Abort and a second Cancel used to throw
it away</li>
<li>Pages and scripts that a browser refused to run in the mirror are fixed, sites built on a JavaScript
framework included</li>
<li>Many fixes to how the engine bounds what a hostile page or server sends it</li>
</ul>
ContentRating:
oars-1.1: {}