flameshot by flameshot-org

Powerful yet simple to use screenshot software :desktop_computer: :camera_flash:

created at May 10, 2017, 7:44 p.m.

C++

206 -1

23,279 +51

1,499 +1

GitHub
badger by dgraph-io

Fast key-value DB in Go.

created at Jan. 26, 2017, 5:09 a.m.

Go

240 +0

13,427 +23

1,149 +2

GitHub
ArchiveBox by ArchiveBox

🗃 Open source self-hosted web archiving. Takes URLs/browser history/bookmarks/Pocket/Pinboard/etc., saves HTML, JS, PDFs, media, and more...

created at May 5, 2017, 8:50 a.m.

Python

172 +0

19,905 +54

1,083 +6

GitHub
SingleFile by gildas-lormeau

Web Extension for saving a faithful copy of a complete web page in a single HTML file

created at Sept. 12, 2010, 11:50 p.m.

JavaScript

114 +0

13,911 +141

921 +6

GitHub
xdotool by jordansissel

fake keyboard/mouse input, window management, and more

created at Feb. 16, 2011, 2:41 a.m.

C

56 +0

3,047 +10

311 +1

GitHub
chrome-remote-interface by cyrus-and

Chrome Debugging Protocol interface for Node.js

created at April 17, 2013, 6 p.m.

JavaScript

81 +0

4,192 +3

300 +1

GitHub
monolith by Y2Z

⬛️ CLI tool for saving complete web pages as a single HTML file

created at Feb. 20, 2017, 7:47 a.m.

Rust

62 +0

9,991 +29

284 +0

GitHub
twarc by DocNow

A command line tool (and Python library) for archiving Twitter JSON

created at Jan. 14, 2013, 2:35 p.m.

Python

35 +0

1,354 +0

253 +0

GitHub
internetarchive by jjjake

A Python and Command-Line Interface to Archive.org

created at Aug. 15, 2012, 7:18 p.m.

Python

51 +0

1,526 +8

209 +0

GitHub
pywb by webrecorder

Core Python Web Archiving Toolkit for replay and recording of web archives

created at Dec. 9, 2013, 3:30 a.m.

JavaScript

61 +1

1,309 +6

207 +1

GitHub
wikiteam by WikiTeam

Tools for downloading and preserving wikis. We archive wikis, from Wikipedia to tiniest wikis. As of 2023, WikiTeam has preserved more than 350,000 wikis.

created at June 25, 2014, 10:18 a.m.

Python

40 +0

692 +2

144 +0

GitHub
DownloadNet by dosyago

💾 DownloadNet - All content you browse online available offline. Search through the full-text of all pages in your browser history. ⭐️ Star to support our work!

created at Dec. 20, 2019, 9:47 a.m.

JavaScript

42 +0

3,654 +4

137 +0

GitHub
grab-site by ArchiveTeam

The archivist's web crawler: WARC output, dashboard for all crawls, dynamic ignore patterns

created at Feb. 5, 2015, 5:01 a.m.

Python

40 +0

1,270 +6

125 +3

GitHub
brozzler by internetarchive

brozzler - distributed browser-based web crawler

created at July 13, 2015, 11:48 p.m.

Python

36 +0

630 +0

93 +0

GitHub
wpull by ArchiveTeam

Wget-compatible web downloader and crawler.

created at Dec. 7, 2013, 1:03 p.m.

HTML

23 +0

536 +1

77 +1

GitHub
browsertrix-crawler by webrecorder

Run a high-fidelity browser-based crawler in a single Docker container

created at Nov. 2, 2020, 4:37 a.m.

TypeScript

23 +0

551 +4

69 +1

GitHub
wayback by wabarc

An archiving tool with an IM-style interface that prioritizes privacy and accessibility, integrated with various archival services including Internet Archive, archive.today, IPFS, Telegraph, and file systems.

created at June 13, 2020, 10:08 a.m.

Go

9 -2

1,658 +8

59 -1

GitHub
warcprox by internetarchive

WARC writing MITM HTTP/S proxy

created at Oct. 25, 2013, 11:27 p.m.

Python

33 +1

363 +1

55 +0

GitHub
warcio by webrecorder

Streaming WARC/ARC library for fast web archive IO

created at March 6, 2017, 6:17 p.m.

Python

22 +0

346 +1

54 +0

GitHub
auto-archiver by bellingcat

Automatically archive links to videos, images, and social media content from Google Sheets (and more).

created at Jan. 15, 2021, 10:30 a.m.

Python

19 +0

474 +4

53 +0

GitHub