RSS 8 projects tagged "Deduplication"

No download Website Updated 28 Aug 2013 WD Arkeia Network Backup

Screenshot
Pop 388.72
Vit 22.66

Arkeia Network Backup is designed for organizations that require fast, easy-to-use, and affordable data protection. It backs up critical data to disk, tape, and cloud storage. Arkeia protects all major virtual platforms including VMware, Hyper-V, XenServer, and more than 200 physical platforms including Windows, Mac OS X, Linux, Netware, most UNIX flavors, and BSDs. The company’s source-side Progressive Deduplication technology helps users realize better performance at a lower cost by reducing data volumes. Arkeia’s deduplication is crucial to accelerating replication of on-premise backups to private or public clouds.

Download Website Updated 22 Aug 2010 lessfs

Screenshot
Pop 170.14
Vit 4.78

Lessfs is a high performance inline data deduplicating file system for Linux. Lessfs complies to the POSIX standard and is very useful for backup purposes as well as providing storage for virtual machine images. Although lessfs is a file system that is implemented in user space with FUSE, it offers decent performance. Lessfs is capable of handling data rates up to 350MB/sec. It supports filesystem encryption.

No download No website Updated 16 Feb 2014 Duke

Screenshot
Pop 235.99
Vit 11.94

Duke is a fast and flexible record linkage engine. It does not use the traditional blocking (sort by key) approach, but instead relies on Lucene. This makes it high-performance (able to process 1,000,000 records in ~10 minutes). Duke can be run from the command line, but also has an API allowing incremental linking applications to be built easily. It supports reading data from CSV, JDBC, SPARQL, and NTriples, and also supports a number of string comparators and string normalizers.

No download Website Updated 26 Sep 2011 BlackHole

Screenshot
Pop 93.28
Vit 1.43

BlackHole is an data de-duplicating network block device that also supports mirroring, snapshots, and support for multiple LUNs using the same data store. It is filesystem agnostic and has been tested with ext2/3/4, NTFS, ReiserFS, and the Oracle Cluster File System (OCFS2). It supports encryption, compression, and multiple storage backends. The hashing scheme used is user configurable. The program exports an NBD device which can be mounted in Linux and GNU/Hurd.

Download No website Updated 02 Mar 2014 Pcompress

Screenshot
Pop 365.31
Vit 15.97

Pcompress is an archiver that can do compression/decompression and deduplication in parallel by splitting input data into chunks. It has a modular structure and includes support for multiple algorithms like LZMA, Bzip2, PPMD, LZ4, etc., with KECCAK/BLAKE2/SHA-256/512 chunk checksums. SSE optimizations for the bundled LZMA are included. It also implements chunk-level Content-Aware Deduplication and Delta Compression features based on a Polynomial Fingerprinting scheme. It has low metadata overhead and overlaps I/O and compression to achieve maximum parallelism. It has AES encryption capability and uses Scrypt from Tarsnap to generate per-session unique keys from passwords. It can work in pipe mode, reading from stdin and writing to stdout. It also provides some adaptive compression modes in which a suitable algorithm is chosen per chunk based on heuristics.

No download Website Updated 07 Apr 2014 Attic

Screenshot
Pop 373.85
Vit 12.56

Attic is a deduplicating backup program. The main goal of attic is to provide an efficient and secure way to back up data. The data deduplication technique used makes Attic suitable for daily backups since only actual changes are stored. Main features: space efficient storage, optional data encryption, and off-site backups.

Download No website Updated 22 Mar 2014 Fileaxy

Screenshot
Pop 77.79
Vit 13.15

Fileaxy is a file de-duplication, organization, and bulk previewing tool which utilizes a new user interface for local file management.

Download Website Updated 10 Apr 2014 Skylable SX

Screenshot
Pop 74.96
Vit 1.03

Skylable SX is a reliable, powerful, fully distributed cluster solution for your data storage needs. It can aggregate the disk space available on multiple servers and merge it into a single storage system. The cluster makes sure that your data is always replicated over multiple nodes (the exact number of copies is defined by the sysadmin) and synchronized. It has built-in support for deduplication, client-side encryption, on-the-fly compression, and much more.

Screenshot

Project Spotlight

quadtree

A Thread-safe quad tree C library.

Screenshot

Project Spotlight

dos2unix

Utilities for converting text files from DOS/Mac format to Unix format and vice versa.