souvenir
Tag cloud
Picture wall
Daily
RSS Feed
  • RSS Feed
  • Daily Feed
Filters

Links per page

  • 20 links
  • 50 links
  • 100 links

Filters

Untagged links
10 results tagged archive  ✕   ✕
Paperless-ngx https://docs.paperless-ngx.com/
19/07/2026 cluster icon
  • Open Paperless : Scan, index, and archive all of your paper documents. Open Paperless is a re-think of the user interface and user experience for Mayan EDMS. The goal ...
  • Internet Archive: Wayback Machine : See the internet's history
  • jina : Jina is a neural search framework that allows to build deep learning search applications in minutes. It provides scalable indexing, querying, understa...
  • Tesseract.js : Tesseract.js is a pure Javascript port of the popular Tesseract OCR engine. This library supports over 60 languages, automatic text orientation and sc...
  • WhatWeb : WhatWeb identifies websites. Its goal is to answer the question, “What is that Website?”. WhatWeb recognises web technologies including content manage...

A community-supported supercharged document management system: scan, index and archive all your documents

archive document scan ocr search
Anthropophony https://www.anthropophony.org/
31/07/2021 cluster icon
  • CDP WASM Suite : A unique sound transformation toolkit, compiled to run anywhere. The Composer's Desktop Project — hundreds of esoteric spectral, granular and waveset ...
  • Wavacity : A free and open source audio editor for the web derived from Audacity
  • Freesound : Freesound is a collaborative database of Creative Commons Licensed sounds. Browse, download and share sounds.
  • librosa : librosa is a python package for music and audio analysis. It provides the building blocks necessary to create music information retrieval systems.
  • Free to use sounds : Royalty Free Sound Effects & Sound Libraries

Archive of an era's sounds

sound archive audio free
Colly http://go-colly.org/
24/05/2021 cluster icon
  • Revel : A high-productivity web framework for the Go language.
  • Scrapy : Scrapy is a fast high-level screen scraping and web crawling framework, used to crawl websites and extract structured data from their pages. It can be...
  • Goutte : Goutte is a screen scraping and web crawling library for PHP. Goutte provides a nice API to crawl websites and extract data from the HTML/XML response...
  • Beego : An open source framework to build and develop your applications in the Go way
  • Go kit : Go is a great general-purpose language, but microservices require a certain amount of specialized support. RPC safety, system observability, infrastru...
thumbnail

Colly is a Go framework that provides a clean interface to write any kind of crawler/scraper/spider

With Colly you can easily extract structured data from websites, which can be used for a wide range of applications, like data mining, data processing or archiving.

go scraper crawler framework datamining archive
db https://github.com/infostreams/db
25/12/2019 cluster icon
  • migra : migra is a schema comparison tool for PostgreSQL. It's a command line tool, and Python library. Find differences in database schemas as easily as runn...
  • Dolt : Dolt is a relational database, i.e. it has tables, and you can execute SQL queries against those tables. It also has version control primitives that o...
  • waybackpack : Waybackpack is a command-line tool that lets you download the entire Wayback Machine archive for a given URL.
  • VersionPress : VersionPress is a free and open source version control plugin for WordPress built on Git. You can: Undo changes Create staging sites Merge databases ...
  • pgBackRest : pgBackRest aims to be a simple, reliable backup and restore system that can seamlessly scale up to the largest databases and workloads.
thumbnail

With DB you can very easily save, restore, and archive snapshots of your database from the command line. It supports connecting to different database servers (for example a local development server and a staging or production server) and allows you to load a database dump from one environment into another environment.

database cli version archive backup
Open Paperless https://github.com/zhoubear/open-paperless
29/12/2017 cluster icon
  • Paperless-ngx : A community-supported supercharged document management system: scan, index and archive all your documents
  • MediaGoblin : MediaGoblin is a free software media publishing platform that anyone can run. You can think of it as a decentralized alternative to Flickr, YouTube, S...
  • Glances : Glances is a cross-platform curses-based system monitoring tool written in Python.
  • Data Science Toolkit : A collection of the best open data sets and open-source tools for data science, wrapped in an easy-to-use REST/JSON API with command line, Python and ...
  • TextBlob : TextBlob is a Python (2 and 3) library for processing textual data. It provides a simple API for diving into common natural language processing (NLP) ...
thumbnail

Scan, index, and archive all of your paper documents. Open Paperless is a re-think of the user interface and user experience for Mayan EDMS. The goal is to reduce the complexity and make it more suitable for home users.

python document management scan archive freesoftware
waybackpack https://github.com/jsvine/waybackpack
19/05/2016 cluster icon
  • Seashells : Seashells lets you pipe output from command-line programs to the web in real-time, even without installing any new software on your machine. You can u...
  • PURL : PURLs are persistent URLs, they provide permanent addresses for resources on the web.
  • Internet Archive: Wayback Machine : See the internet's history
  • Phantomas : PhantomJS-based web performance metrics collector and monitoring tool
  • Nativefier : Nativefier is a command line tool that allows you to easily create a desktop application for any web site with succinct and minimal configuration. App...
thumbnail

Waybackpack is a command-line tool that lets you download the entire Wayback Machine archive for a given URL.

archive web cli
oldweb.today http://oldweb.today/
02/12/2015 cluster icon
  • RequireJS : RequireJS is a JavaScript file and module loader. It is optimized for in-browser use, but it can be used in other JavaScript environments, like Rhino ...
  • ImmortalDB : ImmortalDB is the best way to store persistent key-value data in the browser. Data saved to ImmortalDB is redundantly stored in Cookies, IndexedDB, Lo...
  • The HTML5 Shiv : The HTML5 Shiv enables use of HTML5 sectioning elements in legacy Internet Explorer and provides basic HTML5 styling for Internet Explorer 6-9, Safari...
  • Paperless-ngx : A community-supported supercharged document management system: scan, index and archive all your documents
  • Selenium : Selenium automates browsers. That's it! What you do with that power is entirely up to you. Primarily, it is for automating web applications for testin...

Browse old web pages the old way with virtual browsers in the browser.

retro browser navigation archive
PURL http://purl.org/docs/index.html
20/08/2010 cluster icon
  • FreezePage : Take snapshots of any Web page
  • waybackpack : Waybackpack is a command-line tool that lets you download the entire Wayback Machine archive for a given URL.
  • Internet Archive: Wayback Machine : See the internet's history
  • IFTT : Put the internet to work for you.
  • Web Developer : The Web Developer extension adds various web developer tools to a browser. The extension is available for Chrome, Firefox and Opera, and will run on a...

PURLs are persistent URLs, they provide permanent addresses for resources on the web.

web persistent archive
FreezePage http://www.freezepage.com
16/08/2010 cluster icon
  • Hoppscotch : A free, online, open source API request builder
  • Web Developer : The Web Developer extension adds various web developer tools to a browser. The extension is available for Chrome, Firefox and Opera, and will run on a...
  • Crosswalk : Include the Crosswalk Project web runtime with your hybrid Android or Cordova / PhoneGap app, and users will consistently see it through a predictable...
  • Skipfish : Skipfish is an active web application security reconnaissance tool. It prepares an interactive sitemap for the targeted site by carrying out a recursi...
  • BackstopJS : BackstopJS automates visual regression testing of your responsive web UI by comparing DOM screenshots over time.

Take snapshots of any Web page

web tool archive
Internet Archive: Wayback Machine https://web.archive.org/
14/09/2007 cluster icon
  • waybackpack : Waybackpack is a command-line tool that lets you download the entire Wayback Machine archive for a given URL.
  • PURL : PURLs are persistent URLs, they provide permanent addresses for resources on the web.
  • FreezePage : Take snapshots of any Web page
  • History of the browser user-agent string : History of the browser user-agent string
  • Topsy : Search and Analyze the Social Web.

See the internet's history

search archive web history
1678 links
Shaarli - The personal, minimalist, super-fast, database free, bookmarking service by the Shaarli community - Theme by kalvn