Micron Document
► Archiving
unofficial mirror · upstream 80abcc80ebe8 · 2026-09-23 14:21 UTC



▷ Archive Services

• ↪️ 4chan Archives (https://www.reddit.com/r/FREEMEDIAHECKYEAH/wiki/social-media#wiki_.25B7_4chan_archives)
• ⭐ Internet Archive (https://archive.org/) - Internet Archive / Tools (https://www.reddit.com/r/FREEMEDIAHECKYEAH/wiki/storage/#wiki_internet_archive_tools)
• ⭐ Wayback Machine (https://web.archive.org/) - Archive Web Pages
• ⭐ Wayback Machine Tools - Downloader (https://github.com/jsvine/waybackpack) / Browser Extension (https://github.com/internetarchive/wayback-machine-webextension), 2 (https://vegetableman.github.io/vandal/) / Script (https://github.com/overcast07/wayback-machine-spn-scripts) / Auto Load (https://gitlab.com/gkrishnaks/WaybackEverywhere-Firefox)
• ⭐ Web Archives (https://github.com/dessant/web-archives) or Resurrect Pages Fork (https://github.com/Albirew/resurrect-pages-isup-edition) - Browser Extensions
• ⭐ CachedView (https://cachedview.nl/) or Quick Cache (https://cybdetective.com/quickcacheandarhivesearch.html) - Aggregate Cache Results
• Archive.today (https://archive.is/) / .li (https://archive.li/) / .ph (https://archive.ph/) / .vn (https://archive.vn/) / .fo (https://archive.fo/) / .md (https://archive.md/) - Archive Web Pages / Paywall Bypass
• Ghost Archive (https://ghostarchive.org/) - Archive Web Pages
• WebArchive.io (https://www.webarchive.io/) - Archive Web Pages
• ArchiveTeam (https://wiki.archiveteam.org/index.php/Main_Page) - Archiving Project / Wiki / Full Site Archive
• Perma.cc (https://perma.cc/) - Create Permalinks


▷ Web Archiving Tools

• 🌐 Awesome Web Archiving (https://github.com/iipc/awesome-web-archiving) - Web Archiving Tools
• 🌐 Data Hoarding (https://datahoarding.org/resources.html) - Data Hoarding Resources
• 🌐 Webrecorder (https://webrecorder.net/) - Open-Source Archiving Tools
• ↪️ Twitter Archiving (https://www.reddit.com/r/FREEMEDIAHECKYEAH/wiki/social-media/#wiki_.25B7_twitter.2Fx_archiving)
• ↪️ YouTube Archiving (https://www.reddit.com/r/FREEMEDIAHECKYEAH/wiki/social-media/#wiki_.25B7_youtube_archiving)
• ⭐ ArchiveBox (https://archivebox.io) - Self-Hosted Web Archiving / GitHub (https://github.com/archivebox/archivebox)
• ⭐ HTTrack (https://www.httrack.com/) - Website Downloader / GitHub (https://github.com/xroche/httrack)
• ⭐ Kiwix (https://get.kiwix.org/en/solutions/applications/download-options/) / Library (https://browse.library.kiwix.org/) / Zim Reader (https://zimit.kiwix.org/) / Wiki DL Guide (https://practicalbetterments.com/download-all-of-wikipedia-on-your-phone/) / Subreddit (https://www.reddit.com/r/Kiwix/) / GitHub (https://github.com/kiwix/) or DownloadNet (dn) (https://github.com/DO-SAY-GO/dn) - Offline Website Readers
• ⭐ datahoarder-website-to-markdown (https://github.com/passthesh3ll/datahoarder-website-to-markdown) - Index to Markdown Tool
• MarkdownDown (https://markdowndown.vercel.app/) - Download Web Pages as Markdown Files
• Tubeup (https://github.com/bibanon/tubeup) - Multi-VOD Service to IA Uploader
• Irchiver (https://irchiver.com/) - Automatic Web Browser Screenshots
• Monolith (https://github.com/Y2Z/monolith) or Single File (https://addons.mozilla.org/en-US/firefox/addon/single-file) - Save Webpages as HTML
• WAIL (https://machawk1.github.io/wail/) - GUI for Archiving Tools / GitHub (https://github.com/machawk1/wail)
• ReplayWeb (https://replayweb.page/) or OldWeb (https://oldweb.today/) - View Web Archive Files
• ArchiveWeb.page (https://archiveweb.page/) - Browser Extension
• WikiTeam (https://github.com/WikiTeam/wikiteam) - Archive Wikis
• Wayback (https://github.com/wabarc/wayback) - Web Archiving Tool
• Wget2 (https://gitlab.com/gnuwget/wget2) / Commands (https://www.whatismybrowser.com/developers/tools/wget-wizard/), SuckIT (https://github.com/skallwar/suckit), Cyotek WebCopy (https://www.cyotek.com/cyotek-webcopy), Website Downloader (https://github.com/AhmadIbrahiim/Website-downloader) or PageRip (https://webpagerip.com/) - Website Downloaders
• Archivematica (https://www.archivematica.org/) - Digital Preservation System
• wallabag (https://wallabag.org/) - Save Articles
• CopySite (https://xdan.ru/copysite/) - Copy Websites
• Scoop (https://github.com/harvard-lil/scoop) - Capture Engine


▷ Web Scraping / Crawling

• 🌐 Awesome Web Scraping (https://github.com/lorien/awesome-web-scraping) or Web Scraping FYI (https://webscraping.fyi/) - Web Scraping Tools / Resources
• ⁠FlareSolverr (https://github.com/FlareSolverr/FlareSolverr), Byparr (https://github.com/ThePhaseless/Byparr) or ⁠Trawl (https://github.com/germondai/trawl) - Challenge-Solving Proxies
• SpiderSuite (https://spidersuite.io/) - Advanced Web Crawler / GitHub (https://github.com/spidersuite/SpiderSuite)
• Heritrix (https://heritrix.readthedocs.io/) - Internet Archive's Web Crawler / GitHub (https://github.com/internetarchive/heritrix3)
• 80legs (https://80legs.com/) - Cloud-Based
• Crawly (https://crawly.diffbot.com/) - Online Scraper
• ⁠Scrapling (https://github.com/D4Vinci/Scrapling) - Web Scraper
• web.scraper.workers.dev (https://web.scraper.workers.dev/) - Web Scraper
• Waymore (https://github.com/xnl-h4ck3r/waymore/) - Web Scraper
• grab-site (https://github.com/ArchiveTeam/grab-site) - ArchiveTeam Web Crawler
• Instant Data Scraper (https://chromewebstore.google.com/detail/instant-data-scraper/ofaokhiedipichpaobibbnahnkdoiiah) - Browser Extension
• brozzler (https://github.com/internetarchive/brozzler) - Web Crawler
• Crawl4AI (https://github.com/unclecode/crawl4ai) - LLM-Friendly Scraper / Crawler