ArchiveBot, an IRC bot for archiving websites
☆421Apr 17, 2026Updated 4 months ago
Alternatives and similar repositories for ArchiveBot
Users that are interested in ArchiveBot are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Wget-compatible web downloader and crawler.☆613Apr 29, 2024Updated 2 years ago
- Boot scripts for the ArchiveTeam Warrior 2☆26Jul 5, 2025Updated last year
- The archivist's web crawler: WARC output, dashboard for all crawls, dynamic ignore patterns☆1,609May 23, 2025Updated last year
- Grabbing all news.☆60Dec 23, 2019Updated 6 years ago
- Making a reusable toolkit for writing seesaw scripts☆78Jul 17, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆15Nov 5, 2018Updated 7 years ago
- The Seesaw pipeline grab script for the URLTeam (terroroftinytown) project☆28Jul 17, 2025Updated last year
- URLTeam's second generation of URL shortener archiving tools☆81Mar 12, 2026Updated 5 months ago
- Grabbing everything from reddit.☆62Feb 16, 2024Updated 2 years ago
- A Dockerfile for the ArchiveTeam Warrior☆453Jul 22, 2026Updated last month
- NOTE: This project is no longer being actively developed.. Check out https://replayweb.page / https://github.com/webrecorder/replayweb.pa…☆203Jan 22, 2025Updated last year
- Bash scripts which interact with Internet Archive Wayback Machine's Save Page Now☆147Apr 3, 2025Updated last year
- Web archiving using Google Chrome☆45Dec 30, 2019Updated 6 years ago
- A configurable, reusable tracker with dashboard☆36Dec 15, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The OpenWayback Development☆526Jan 3, 2024Updated 2 years ago
- wpull fork with fixes and faster parsing using html5-parser; used by grab-site; should go away when wpull is similarly improved☆31Sep 20, 2025Updated 11 months ago
- We back up a lot of stuff from around the web; now it's time to back up the Internet Archive, just in case.☆93Jul 13, 2020Updated 6 years ago
- Use yt-dlp to download video/metadata and upload to the Internet Archive.☆515Aug 12, 2026Updated 3 weeks ago
- A collection of tools for archiving and analysing the internet.☆81Jul 6, 2022Updated 4 years ago
- Archiving URLs (outlinks) from a variety of sources.☆25Aug 14, 2026Updated 2 weeks ago
- Sources for urls-grab.☆15Aug 25, 2026Updated last week
- Archiving all metadata from YouTube (everything except videos themselves due to size)☆35Jul 17, 2026Updated last month
- Archiving GitHub☆11Aug 5, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Photobucket image and album extractor. Improved version of PB_Shovel by Daxda, to scrape urls.☆11Mar 2, 2019Updated 7 years ago
- WARC writing MITM HTTP/S proxy☆464Jun 17, 2026Updated 2 months ago
- Scrape https://unlistedvideos.com/☆15Jul 18, 2021Updated 5 years ago
- Python library for reading and writing warc files☆249Mar 7, 2022Updated 4 years ago
- A Python and Command-Line Interface to Archive.org☆1,900Aug 25, 2026Updated last week
- A Tool To Push Web Resources Into Web Archives☆434Jan 23, 2024Updated 2 years ago
- A simple Python wrapper for the archive.is capturing service☆218Feb 11, 2025Updated last year
- An Awesome List for getting started with web archiving☆2,633Aug 17, 2026Updated 2 weeks ago
- Core Python Web Archiving Toolkit for replay and recording of web archives☆1,697Aug 26, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Wget-AT is a modern Wget with Lua hooks, Zstandard (+dictionary) WARC compression and URL-agnostic deduplication.☆138Mar 19, 2026Updated 5 months ago
- This Chrome extension save all web pages you viewed to the Wayback Machine☆14Mar 7, 2021Updated 5 years ago
- A simple Python wrapper and command-line interface for archive.org’s "Save Page Now" capturing service☆196Jun 17, 2026Updated 2 months ago
- Bash script to force the first page of items on the Internet Archive to be the default.☆15Jun 28, 2019Updated 7 years ago
- An easy-to-use and highly customizable crawler that enables you to create your own little Web archives (WARC/CDX)☆26Oct 9, 2017Updated 8 years ago
- Warrior virtual machine appliance (version 4)☆65Aug 25, 2026Updated last week
- Run a high-fidelity browser-based web archiving crawler in a single Docker container☆1,125Updated this week