Wayback Machine API interface & a command-line tool
☆602Feb 26, 2024Updated 2 years ago
Alternatives and similar repositories for waybackpy
Users that are interested in waybackpy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Python API to the Internet Archive Wayback Machine☆91Aug 3, 2026Updated 2 weeks ago
- A Tool To Push Web Resources Into Web Archives☆434Jan 23, 2024Updated 2 years ago
- A Python script to submit web pages to the Wayback Machine for archiving.☆88Jul 29, 2026Updated 3 weeks ago
- A simple Python wrapper and command-line interface for archive.org’s "Save Page Now" capturing service☆196Jun 17, 2026Updated 2 months ago
- Core Python Web Archiving Toolkit for replay and recording of web archives☆1,688Apr 10, 2026Updated 4 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Official Python package for ArchiveBox, the self-hosted internet archiving solution.☆12Oct 5, 2024Updated last year
- A CherryTree template for People OSINT. I was inspired by James Hall's CTF template and I used the lessons taught to me by Joe Gray to cr…☆11Aug 16, 2020Updated 6 years ago
- A toolkit makes it easier to archive webpages to IPFS☆12Jul 31, 2023Updated 3 years ago
- A Python and Command-Line Interface to Archive.org☆1,896Updated this week
- An Awesome List for getting started with web archiving☆2,620Updated this week
- Tools for helping you work with web platform archive downloads.☆18Mar 27, 2020Updated 6 years ago
- Command line tool for digging into WARC files☆50Aug 5, 2026Updated 2 weeks ago
- Download the entire Wayback Machine archive for a given URL.☆3,227Apr 21, 2025Updated last year
- Web Archiving Course☆24Mar 4, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Cheat sheets for using Kali Linux☆16Feb 17, 2022Updated 4 years ago
- Backend, IA-specific tools for crawling and processing the scholarly web. Content ends up in https://fatcat.wiki☆28Jul 31, 2024Updated 2 years ago
- A PDF classifier ensemble with REST API service☆23Mar 5, 2021Updated 5 years ago
- OSINT tool to download archived PDF files from archive.org for a given website.☆61Jun 20, 2020Updated 6 years ago
- Automated behaviors that run in browser to interact with complex sites automatically. Used by ArchiveWeb.page and Browsertrix Crawler.☆58Updated this week
- Bash scripts which interact with Internet Archive Wayback Machine's Save Page Now☆147Apr 3, 2025Updated last year
- ☆18Mar 31, 2025Updated last year
- A command line utility for listing and searching snapshots in web archives☆20Jun 4, 2026Updated 2 months ago
- Specifications developed and maintained by the Webrecorder community.☆143Oct 16, 2025Updated 10 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- The archivist's web crawler: WARC output, dashboard for all crawls, dynamic ignore patterns☆1,605May 23, 2025Updated last year
- Template for new OSINT command-line tools☆81Updated this week
- A command-line utility and Scrapy middleware for scraping time series data from Archive.org's Wayback Machine.☆480Feb 23, 2024Updated 2 years ago
- De-clutter a list of URLs☆391Mar 8, 2026Updated 5 months ago
- Web archiving using Google Chrome☆45Dec 30, 2019Updated 6 years ago
- Get URLs from the Wayback Machine. Able to handle large outputs.☆35Sep 15, 2023Updated 2 years ago
- ☁️ Curated Cloud OSINT resources — dorks, tools, and techniques for AWS, Azure, GCP, Oracle Cloud, and other major providers reconnaissan…☆135Apr 8, 2026Updated 4 months ago
- Extract Reddit user data☆21Apr 23, 2022Updated 4 years ago
- 🗃 Open source self-hosted web archiving. Takes URLs/browser history/bookmarks/Pocket/Pinboard/etc., saves HTML, JS, PDFs, media, and mor…☆28,130Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Go tool that gets the newest PRs from projectdiscovery/nuclei-templates.☆55Jun 13, 2023Updated 3 years ago
- Scripts for Internet Archive☆14Mar 26, 2025Updated last year
- ☆11Mar 27, 2024Updated 2 years ago
- IA's public Wayback Machine (moved from SourceForge)☆858Mar 1, 2024Updated 2 years ago
- Streaming WARC/ARC library for fast web archive IO☆468Jun 10, 2026Updated 2 months ago
- An archiving tool with an IM-style interface that prioritizes privacy and accessibility, integrated with various archival services includ…☆2,225Updated this week
- Download digitized books from Internet Archive and view with IIIF, locally and offline.☆39Apr 19, 2024Updated 2 years ago