restic

mirror of https://github.com/octoleo/restic.git synced 2024-11-22 12:55:18 +00:00

Author	SHA1	Message	Date
Michael Eischer	8ef2968f28	repository: remove unused index.ListPack	2022-07-02 18:39:12 +02:00
Michael Eischer	e4f20dea61	repository: inline index.encode	2022-07-02 18:39:12 +02:00
Michael Eischer	fe5a8e137a	repository: remove unused index.Store	2022-07-02 18:39:12 +02:00
Michael Eischer	628ae799ca	repository: make flushPacks private	2022-07-02 18:39:12 +02:00
Michael Eischer	ed8aa15376	repository: add Save method to MasterIndex interface	2022-07-02 18:38:56 +02:00
Michael Eischer	a77d5c4d11	repository: index saving belongs into the MasterIndex	2022-07-02 18:38:56 +02:00
MichaelEischer	2c893fe43c	Merge pull request #3798 from greatroar/errors all: Move away from pkg/errors, easy cases	2022-06-17 19:01:40 +02:00
greatroar	f92ecf13c9	all: Move away from pkg/errors, easy cases github.com/pkg/errors is no longer getting updates, because Go 1.13 went with the more flexible errors.{As,Is} function. Use those instead: errors from pkg/errors already support the Unwrap interface used by 1.13 error handling. Also: * check for io.EOF with a straight ==. That value should not be wrapped, and the chunker (whose error is checked in the cases changed) does not wrap it. * Give custom Error methods pointer receivers, so there's no ambiguity when type-switching since the value type will no longer implement error. * Make restic.ErrAlreadyLocked private, and rename it to alreadyLockedError to match the stdlib convention that error type names end in Error. * Same with rest.ErrIsNotExist => rest.notExistError. * Make s3.Backend.IsAccessDenied a private function.	2022-06-14 08:36:38 +02:00
Jayson Wang	f144920ed5	fix handling of maxKeys in SearchKey	2022-06-12 14:19:06 +02:00
greatroar	c9557b2822	internal/repository: Fix LoadBlob + fuzz test When given a buf that is big enough for a compressed blob but not its decompressed contents, the copy at the end of LoadBlob would skip the last part of the contents. Fixes #3783.	2022-06-06 17:02:28 +02:00
MichaelEischer	b2a2e5f727	Merge pull request #3753 from greatroar/indexmap-alloc repository: Re-tune indexmap allocation strategy	2022-05-14 15:44:08 +02:00
greatroar	5141228e0c	repository: Re-tune indexmap allocation strategy `fd05037e1a` changed the allocation batch size from 256 to 128 under the assumption that an indexEntry is 60 bytes on amd64, but it's 64: structs are padded out to a multiple of 8 for alignment reasons. That means we'd waste no space in malloc even without the batch allocation, at least on 64-bit machines. While that strategy cuts the overallocation down dramatically for many small indexes, it also seems to slow allocation down (Go 1.18, Linux, amd64, -benchtime=2s): name old time/op new time/op delta DecodeIndex-8 4.67s ± 5% 4.60s ± 1% ~ (p=0.953 n=10+5) DecodeIndexParallel-8 4.67s ± 3% 4.60s ± 1% ~ (p=0.953 n=10+5) IndexHasUnknown-8 37.8ns ± 8% 36.5ns ±14% ~ (p=0.841 n=5+5) IndexHasKnown-8 38.5ns ±12% 37.7ns ±10% ~ (p=0.968 n=5+5) IndexAlloc-8 615ms ±18% 607ms ± 1% ~ (p=1.000 n=10+5) IndexAllocParallel-8 245ms ±11% 285ms ± 6% +16.40% (p=0.001 n=10+5) MasterIndexAlloc-8 286ms ± 9% 275ms ± 2% ~ (p=1.000 n=10+5) LoadIndex/v1-8 27.0ms ± 4% 26.8ms ± 1% ~ (p=0.690 n=5+5) LoadIndex/v2-8 22.4ms ± 1% 22.8ms ± 2% +1.48% (p=0.016 n=5+5) name old alloc/op new alloc/op delta IndexAlloc-8 446MB ± 0% 446MB ± 0% -0.00% (p=0.000 n=8+4) IndexAllocParallel-8 446MB ± 0% 446MB ± 0% -0.00% (p=0.008 n=8+5) MasterIndexAlloc-8 213MB ± 0% 159MB ± 0% -25.47% (p=0.000 n=10+5) name old allocs/op new allocs/op delta IndexAlloc-8 913k ± 0% 2632k ± 0% +188.19% (p=0.008 n=5+5) IndexAllocParallel-8 913k ± 0% 2632k ± 0% +188.21% (p=0.008 n=5+5) MasterIndexAlloc-8 318k ± 0% 1172k ± 0% +267.86% (p=0.008 n=5+5) Instead, this patch sets a batch size of 4, which means no space is wasted by malloc on 64-bit and very little on 32-bit. It still gets very close to the savings from not allocating in batches, without requiring special code for bits.UintSize==64. Benchmark results, again for Linux/amd64: name old time/op new time/op delta DecodeIndex-8 4.67s ± 5% 4.83s ± 9% ~ (p=0.315 n=10+10) DecodeIndexParallel-8 4.67s ± 3% 4.68s ± 4% ~ (p=0.315 n=10+10) IndexHasUnknown-8 37.8ns ± 8% 44.5ns ±19% ~ (p=0.095 n=5+5) IndexHasKnown-8 38.5ns ±12% 36.9ns ± 8% ~ (p=0.690 n=5+5) IndexAlloc-8 615ms ±18% 628ms ±18% ~ (p=0.218 n=10+10) IndexAllocParallel-8 245ms ±11% 262ms ± 9% +7.02% (p=0.043 n=10+10) MasterIndexAlloc-8 286ms ± 9% 287ms ±13% ~ (p=1.000 n=10+10) LoadIndex/v1-8 27.0ms ± 4% 26.8ms ± 0% ~ (p=1.000 n=5+5) LoadIndex/v2-8 22.4ms ± 1% 22.5ms ± 0% ~ (p=0.056 n=5+5) name old alloc/op new alloc/op delta IndexAlloc-8 446MB ± 0% 446MB ± 0% ~ (p=1.000 n=8+10) IndexAllocParallel-8 446MB ± 0% 446MB ± 0% -0.00% (p=0.000 n=8+8) MasterIndexAlloc-8 213MB ± 0% 160MB ± 0% -25.02% (p=0.000 n=10+9) name old allocs/op new allocs/op delta IndexAlloc-8 913k ± 0% 1333k ± 0% +45.94% (p=0.000 n=8+10) IndexAllocParallel-8 913k ± 0% 1333k ± 0% +45.94% (p=0.000 n=8+8) MasterIndexAlloc-8 318k ± 0% 525k ± 0% +64.99% (p=0.000 n=10+10) The allocation method indexmap.newEntry has also been rewritten in a form that is a few instructions shorter.	2022-05-11 21:22:14 +02:00
MichaelEischer	df554e5f69	Merge pull request #3748 from greatroar/runworkers repository: Remove RunWorkers, report ctx.Err()	2022-05-11 19:38:46 +02:00
greatroar	2e0f1f5113	repository: Remove RunWorkers, report ctx.Err() This removes RunWorkers, which had become mere overhead by successive refactors. It also ensures that each former user of that function returns any context error that occurs, so failure to complete an operation is always reported as an error.	2022-05-10 22:26:00 +02:00
Michael Eischer	ae7e51382a	Fix error on temp file deletion on windows Apparently it can take a moment between closing a tempfile marked as DELETE_ON_CLOSE and it actually being deleted. During that time the file is inaccessible. Thus just skip deleting the temp file on windows.	2022-05-09 22:43:26 +02:00
Michael Eischer	cf5cb673fb	repository: Use existing method to collect pack ids	2022-04-30 19:14:21 +02:00
Michael Eischer	b335cb6285	repository: Refactor index IDs collection	2022-04-30 19:14:21 +02:00
Michael Eischer	4b01b06f2f	repository: Test compressed blobs in StreamPack	2022-04-30 11:34:10 +02:00
Michael Eischer	ec2b25565a	repository: test uncompressedLength field and index example	2022-04-30 11:34:10 +02:00
Michael Eischer	9ffb8920f1	repository: run blackbox tests using old and new repo version	2022-04-30 11:34:10 +02:00
Michael Eischer	abe5935693	repository: unify repository version-specific initialization Mark the master index as compressed also when initializing a new repository. This is only relevant for testing.	2022-04-30 11:34:10 +02:00
Alexander Neumann	8776031f96	Leave allocating slices to the decompress code	2022-04-30 11:34:10 +02:00
Alexander Neumann	5eb05a0afe	Configure zstd encoder/decoder	2022-04-30 11:34:10 +02:00
Michael Eischer	2f36e044db	Cleanup pack header check	2022-04-30 11:34:10 +02:00
Alexander Neumann	8b11b86383	Add option global --compression	2022-04-30 11:34:10 +02:00
Michael Eischer	7132df529e	repository: Increase index size for repo version 2 A compressed index is only about one third the size of an uncompressed one. Thus increase the number of entries in an index to avoid cluttering the repository with small indexes.	2022-04-30 11:34:10 +02:00
Michael Eischer	66f9048bce	repository: Alloc zstd encoder/decoder on demand	2022-04-30 11:34:10 +02:00
Michael Eischer	fd05037e1a	repository: recalibrate index batch allocation size	2022-04-30 11:34:10 +02:00
Michael Eischer	6fb408d90e	repository: implement pack compression	2022-04-30 11:34:10 +02:00
Michael Eischer	362ab06023	init: Add flag to specify created repository version	2022-04-30 10:07:42 +02:00
Michael Eischer	4b957e7373	repository: Implement index/snapshot/lock compression The config file is not compressed as it should remain readable by older restic versions such that these can return a proper error. As the old format for unpacked data does not include a version header, make use of a trick: The old data is always encoded as JSON. Thus it can only start with '{' or '['. For any other value the first byte indicates a versioned format. The version is set to 2 for now. Then the zstd compressed data follows.	2022-04-30 10:07:42 +02:00
Michael Eischer	e597b99b55	repository: Reduce repack workers to prevent deadlock As repack streams packs these occupy one backend connection. Uploading a new pack also requires a backend connection. To prevent a deadlock during repack when reaching the backend connections limit, simply limit the repackWorker count to always leave one connection for uploading.	2022-04-23 11:28:18 +02:00
Alexander Neumann	a059ef90f8	Merge pull request #3702 from MichaelEischer/extend-config-error Print used key name if config fails to load	2022-04-10 20:25:24 +02:00
Michael Eischer	c2aabb2686	Print used key name if config fails to load	2022-04-09 22:38:18 +02:00
Alexander Neumann	04e054465a	Merge pull request #3475 from MichaelEischer/local-sftp-conn-limit Limit concurrent operations for local / sftp backend	2022-04-09 21:33:00 +02:00
Michael Eischer	7b9ae91e04	copy: Load snapshots before indexes	2022-04-09 12:27:25 +02:00
Michael Eischer	cd783358d3	local: Limit concurrent backend operations Use a limit of 2 similar to the filereader concurrency in the archiver.	2022-04-09 12:21:38 +02:00
Michael Eischer	6408686973	repository: Simplify Blob equality check	2022-03-28 22:09:49 +02:00
Michael Eischer	243698680a	crypto: Use helpers for size calculations	2022-03-28 22:09:49 +02:00
Michael Eischer	f78bd14e28	repository: Remove pack implementation details from MasterIndex	2022-03-28 22:09:49 +02:00
Michael Eischer	dc3d77dacc	repository: make saveAndEncrypt private	2022-03-28 22:09:49 +02:00
Michael Eischer	6877e7edbb	repository: Rename LoadAndDecrypt to LoadUnpacked The method is the complement for SaveUnpacked and not for SaveAndEncrypt. The latter assembles blobs into pack files.	2022-03-28 22:09:49 +02:00
Michael Eischer	537b4c310a	copy: Implement by reusing repack The repack operation copies all selected blobs from a set of pack files into new pack files. For prune the source and destination repositories are identical. To implement copy, just use a different source and destination repository.	2022-03-26 20:47:15 +01:00
Alexander Neumann	e682f7c0d6	Add tests for StreamPack	2022-03-21 21:15:03 +01:00
Michael Eischer	bba8ba7a5b	repository: cancel streampack context after error	2022-02-12 20:18:25 +01:00
Michael Eischer	47554a3428	repository: Fix error handling in repack When storing a blob fails, this is a fatal error which must not be retried.	2022-02-12 20:18:25 +01:00
Michael Eischer	930a00ad54	checker: reuse bufio reader	2022-02-12 20:18:25 +01:00
Michael Eischer	34ebafb8b6	repository: don't crash if blob size is too short	2022-02-12 20:18:25 +01:00
Michael Eischer	becebf5d88	repository: remove unused DownloadAndHash	2022-02-12 20:18:25 +01:00
Michael Eischer	f1e58e7c7f	checker: rewrite ReadData to stream packs	2022-02-12 20:18:25 +01:00
Michael Eischer	f40abd92fa	restorer: convert to use StreamPack	2022-02-12 20:18:25 +01:00
Michael Eischer	f00f690658	repository: stream packs during repacking	2022-02-12 20:18:25 +01:00
Michael Eischer	c4a2bfcb39	repository: Add StreamPacks function The function supports efficiently loading a specified list of blobs from a single pack in a streaming fashion. That is there's no need for temporary files independent of the pack size.	2022-02-12 20:18:25 +01:00
Michael Eischer	153e2ba859	repository: Implement lisiting blobs per pack file	2022-02-12 20:18:24 +01:00
greatroar	8d2996eaaa	Replace siphash by hash/maphash In Go 1.17.1, maphash has become quite a bit faster than siphash, so we can drop one third-party dependency. maphash is just an interface to the standard Go map's hash function, which we already trust for other use cases. Benchmark results on linux/amd64, -benchtime=3s: name old time/op new time/op delta IndexHasUnknown-8 50.6ns ±10% 41.0ns ±19% -18.92% (p=0.000 n=9+10) IndexHasKnown-8 52.6ns ±12% 41.5ns ±12% -21.13% (p=0.000 n=9+10) IndexMapHash-8 3.64µs ± 1% 2.00µs ± 0% -45.09% (p=0.000 n=10+9) IndexAlloc-8 700ms ± 1% 601ms ± 6% -14.18% (p=0.000 n=8+10) IndexAllocParallel-8 205ms ± 5% 192ms ± 8% -6.18% (p=0.043 n=10+10) MasterIndexAlloc-8 319ms ± 1% 279ms ± 5% -12.58% (p=0.000 n=10+10) MasterIndexLookupSingleIndex-8 156ns ± 8% 147ns ± 6% -5.46% (p=0.023 n=10+10) MasterIndexLookupMultipleIndex-8 150ns ± 7% 142ns ± 8% -5.69% (p=0.007 n=10+10) MasterIndexLookupSingleIndexUnknown-8 74.4ns ± 6% 72.0ns ± 9% ~ (p=0.175 n=10+9) MasterIndexLookupMultipleIndexUnknown-8 67.4ns ± 9% 65.5ns ± 7% ~ (p=0.340 n=9+9) MasterIndexLookupParallel/known,indices=25-8 461ns ± 2% 445ns ± 2% -3.49% (p=0.000 n=10+10) MasterIndexLookupParallel/unknown,indices=25-8 408ns ±11% 378ns ± 5% -7.22% (p=0.035 n=10+9) MasterIndexLookupParallel/known,indices=50-8 479ns ± 1% 437ns ± 4% -8.82% (p=0.000 n=10+10) MasterIndexLookupParallel/unknown,indices=50-8 406ns ± 8% 343ns ±15% -15.44% (p=0.001 n=10+10) MasterIndexLookupParallel/known,indices=100-8 480ns ± 1% 455ns ± 5% -5.15% (p=0.000 n=8+10) MasterIndexLookupParallel/unknown,indices=100-8 391ns ±18% 382ns ± 8% ~ (p=0.315 n=10+10) MasterIndexLookupBlobSize-8 71.0ns ± 8% 57.2ns ±11% -19.36% (p=0.000 n=9+10) PackerManager-8 254ms ± 1% 254ms ± 1% ~ (p=0.285 n=15+15) name old speed new speed delta IndexMapHash-8 1.12GB/s ± 1% 2.05GB/s ± 0% +82.13% (p=0.000 n=10+9) PackerManager-8 208MB/s ± 1% 207MB/s ± 1% ~ (p=0.281 n=15+15) name old alloc/op new alloc/op delta IndexMapHash-8 0.00B 0.00B ~ (all equal) IndexAlloc-8 400MB ± 0% 400MB ± 0% ~ (p=1.000 n=9+10) IndexAllocParallel-8 401MB ± 0% 401MB ± 0% +0.00% (p=0.000 n=10+10) MasterIndexAlloc-8 258MB ± 0% 262MB ± 0% +1.42% (p=0.000 n=9+10) PackerManager-8 73.1kB ± 0% 73.1kB ± 0% ~ (p=0.382 n=13+13) name old allocs/op new allocs/op delta IndexMapHash-8 0.00 0.00 ~ (all equal) IndexAlloc-8 907k ± 0% 907k ± 0% -0.00% (p=0.000 n=10+10) IndexAllocParallel-8 907k ± 0% 907k ± 0% +0.00% (p=0.009 n=10+10) MasterIndexAlloc-8 327k ± 0% 317k ± 0% -3.06% (p=0.000 n=10+10) PackerManager-8 744 ± 0% 744 ± 0% ~ (all equal)	2021-09-19 16:05:18 +02:00
Alexander Weiss	81876d5c1b	Simplify cache logic	2021-09-03 21:01:00 +02:00
Michael Eischer	9aa2eff384	Add plumbing to calculate backend specific file hash for upload This enables the backends to request the calculation of a backend-specific hash. For the currently supported backends this will always be MD5. The hash calculation happens as early as possible, for pack files this is during assembly of the pack file. That way the hash would even capture corruptions of the temporary pack file on disk.	2021-08-04 22:17:46 +02:00
Ryan Hitchman	77bf148460	backup: add --dry-run/-n flag to show what would happen. This can be used to check how large a backup is or validate exclusions. It does not actually write any data to the underlying backend. This is implemented as a simple overlay backend that accepts writes without forwarding them, passes through reads, and generally does the minimal necessary to pretend that progress is actually happening. Fixes #1542 Example usage: $ restic -vv --dry-run . \| grep add new /changelog/unreleased/issue-1542, saved in 0.000s (350 B added) modified /cmd/restic/cmd_backup.go, saved in 0.000s (16.543 KiB added) modified /cmd/restic/global.go, saved in 0.000s (0 B added) new /internal/backend/dry/dry_backend_test.go, saved in 0.000s (3.866 KiB added) new /internal/backend/dry/dry_backend.go, saved in 0.000s (3.744 KiB added) modified /internal/backend/test/tests.go, saved in 0.000s (0 B added) modified /internal/repository/repository.go, saved in 0.000s (20.707 KiB added) modified /internal/ui/backup.go, saved in 0.000s (9.110 KiB added) modified /internal/ui/jsonstatus/status.go, saved in 0.001s (11.055 KiB added) modified /restic, saved in 0.131s (25.542 MiB added) Would add to the repo: 25.892 MiB	2021-08-04 21:19:29 +02:00
Alexander Neumann	aef3658a5f	Address review comments	2021-01-30 20:02:37 +01:00
Alexander Neumann	16313bfcc9	errcheck: Add error check for MergeFinalIndexes()	2021-01-30 20:02:37 +01:00
Alexander Neumann	75f53955ee	errcheck: Add error checks Most added checks are straight forward.	2021-01-30 20:02:37 +01:00
Alexander Neumann	bdfedf1f5b	Merge pull request #3173 from MichaelEischer/unify-index-loading Unify index loading	2021-01-28 13:50:42 +01:00
Michael Eischer	e2b0072441	check: add progress bar to the tree structure check	2021-01-28 11:10:50 +01:00
Michael Eischer	de99207046	repository: tweak comment for packs method	2020-12-22 23:01:58 +01:00
Alexander Weiss	68b74e359e	Count packs directly in RebuildIndexFiles	2020-12-22 23:01:58 +01:00
Michael Eischer	b9f5d3fe13	repository: Add test for ForAllIndexes	2020-12-22 22:36:18 +01:00
Michael Eischer	a12c5f1d37	repository: move otherwise unused LoadIndex to tests	2020-12-22 22:36:18 +01:00
Michael Eischer	24474a36f4	repository: deduplicate index loading implementation	2020-12-22 22:36:18 +01:00
Michael Eischer	96904f8972	check: extract parallel index loading	2020-12-22 22:36:18 +01:00
Alexander Neumann	36c5d39c2c	Fix issues reported by semgrep	2020-12-11 09:41:59 +01:00
Alexander Neumann	7facc8ccc1	Merge pull request #2505 from aawsome/fix-repo-configfile Fix repo configfile	2020-12-07 07:52:37 +01:00
Alexander Weiss	aa7a5f19c2	Use BlobHandle in index methods	2020-11-22 20:41:12 +01:00
Alexander Weiss	e3013271a6	Harmonize naming	2020-11-22 20:41:12 +01:00
Alexander Weiss	92bd448691	Make BlobHandle substruct of Blob	2020-11-22 20:41:10 +01:00
Alexander Weiss	ce5d630681	Add MasterIndex.PackSize()	2020-11-21 22:13:54 +01:00
Alexander Weiss	c3ddde9e7d	Return hdrSize in ListPack	2020-11-21 22:13:54 +01:00
Michael Eischer	1f43cac12d	check: Only track data blobs when unused blobs should be reported This improves the memory usage of check a lot as it now only has to track tree blobs when run using the default parameters.	2020-11-15 18:43:07 +01:00
Alexander Neumann	3c0c0c132b	Merge pull request #3006 from aawsome/new-rebuild-index Reimplement rebuild-index and remove /internal/index	2020-11-15 17:48:43 +01:00
greatroar	ab2b7d7f9a	Decrease allocation rate in internal/pack internal/repository benchmark results: name old time/op new time/op delta PackerManager-8 179ms ± 1% 181ms ± 1% +0.78% (p=0.009 n=10+10) name old speed new speed delta PackerManager-8 294MB/s ± 1% 292MB/s ± 1% -0.77% (p=0.009 n=10+10) name old alloc/op new alloc/op delta PackerManager-8 91.3kB ± 0% 72.2kB ± 0% -20.92% (p=0.000 n=9+7) name old allocs/op new allocs/op delta PackerManager-8 1.38k ± 0% 0.76k ± 0% -45.20% (p=0.000 n=10+7)	2020-11-15 16:51:47 +01:00
Alexander Weiss	30b6a0878a	Reimplement rebuild-index	2020-11-15 07:05:09 +01:00
Alexander Weiss	187c8fb259	Parallelize MasterIndex.Save()	2020-11-15 07:05:09 +01:00
Alexander Weiss	1ec628ddf5	Add extraObsolete to MasterIndex.Save	2020-11-15 07:05:09 +01:00
Alexander Weiss	5898cb341f	Use CreateIndexFromPacks() in test	2020-11-15 07:05:05 +01:00
Alexander Weiss	43732bb885	Add CreateIndexFromPacks()	2020-11-15 07:04:51 +01:00
greatroar	21b787a4d1	Stop Counters where they're constructed and started	2020-11-09 13:03:31 +01:00
greatroar	ddca699cd2	Replace restic.Progress with new progress.Counter This fixes two race conditions while cleaning up the code.	2020-11-09 12:12:35 +01:00
Alexander Weiss	fd33030556	Use in-memory index to rebuild index in prune	2020-11-06 20:23:30 +01:00
Alexander Weiss	38cc4393f6	Add Masterindex.Save(); Add Index.Packs()	2020-11-06 20:23:30 +01:00
Alexander Weiss	aaf1c44362	Fix #3062	2020-11-05 17:05:42 +01:00
Alexander Neumann	ae5302c7a8	Add comment that keepBlobs is modified	2020-11-05 10:33:38 +01:00
Alexander Neumann	866a52ad4e	Remove unneeded seek The file returned from DownloadAndHash() is already seeked to the start of the file.	2020-11-05 10:31:49 +01:00
Alexander Neumann	a4507610a0	Fix typo	2020-11-05 10:31:49 +01:00
Alexander Neumann	7def2d8ea7	Use a non-constant seed	2020-11-05 10:31:49 +01:00
Alexander Neumann	ee0112ab3b	Clarify message about expected error	2020-11-05 10:31:49 +01:00
Michael Eischer	b373f164fe	prune: Parallelize repack command	2020-11-05 10:31:49 +01:00
Alexander Neumann	3ff37215df	Merge pull request #2935 from MichaelEischer/upgrade-minio Upgrade minio SDK to version 7	2020-11-02 09:09:10 +01:00
Sergio Rubio	e708628cfd	Remove unused function Not currently used, and it'd need to be added to the MasterIndex interface first.	2020-10-28 13:24:49 +01:00
Alexander Weiss	b44ecde8b0	Fix setting of ID in DecodeIndex	2020-10-17 09:12:58 +02:00
MichaelEischer	4ba237bb93	Merge pull request #3019 from greatroar/refactor-decodeindex Refactor index decoding	2020-10-15 23:22:33 +02:00
greatroar	b27375f5ce	defer close(ch) outside repository.RunWorkers	2020-10-14 15:50:16 +02:00
greatroar	720e0ee0c7	if cond { return true }; return false => return cond	2020-10-13 20:56:43 +02:00
greatroar	27db3ec262	Refactor index decoding Decoding old-format indices no longer requires loading and decrypting twice.	2020-10-13 20:47:50 +02:00
Alexander Neumann	56883817d8	Merge pull request #2990 from MichaelEischer/fix-goreport-warnings Fix some goreport warnings	2020-10-12 20:44:56 +02:00
Michael Eischer	e638b46a13	Embed context into ReaderAt The io.Reader interface does not support contexts, such that it is necessary to embed the context into the backendReaderAt struct. This has the problem that a reader might suddenly stop working when it's contained context is canceled. However, this is now problem here as the reader instances never escape the calling function.	2020-10-09 22:39:07 +02:00
Michael Eischer	d6cfe857b7	pass proper context into MasterIndex.RebuildIndex	2020-10-09 22:39:07 +02:00
Michael Eischer	a449450021	init: pass proper context to master key generation This is no change in behavior as a canceled context did later on cause the config file creation to fail. Therefore this change just lets the repository initialization fail a bit earlier.	2020-10-09 22:37:56 +02:00
Michael Eischer	c458e114d4	pass context to Find / FindSnapshot This allows proper interruption of restic while it searches for snapshots or key files.	2020-10-09 22:37:56 +02:00
Michael Eischer	45e9a55c62	Wire context into backend layout detection	2020-10-09 22:37:24 +02:00
Michael Eischer	eba5dd831f	Fix typos reported by misspell	2020-10-06 14:55:13 +02:00
Michael Eischer	f003410402	init: Add `--copy-chunker-params` option This allows creating multiple repositories with identical chunker parameters which is required for working deduplication when copying snapshots between different repositories.	2020-09-19 16:53:05 +02:00
greatroar	23fcbb275a	Remove restic.Cache interface It was used in one code path, which then asserted its concrete type as *cache.Cache. Privatised some of the interface methods.	2020-09-18 10:48:13 +02:00
Michael Eischer	c46edcd9d6	error strings should not end with punctuation	2020-09-05 10:07:17 +02:00
Michael Eischer	d0329cf3eb	Adjust comments to match name of exported methods	2020-09-05 10:07:16 +02:00
Michael Eischer	4784540f04	repository: Simplify worker group code	2020-09-05 10:07:16 +02:00
Michael Eischer	d9a80e07b9	repository: Simplify index age calculation	2020-09-05 10:07:16 +02:00
Michael Eischer	2f8335554c	Remove a few unused variables	2020-09-05 10:06:23 +02:00
Michael Eischer	f4b9544ab2	prune: Add test that repack aborts on wrong blob	2020-08-16 11:34:01 +02:00
Michael Eischer	367449dede	prune: Reduce memory allocations while repacking The slicing operator `slice[low:high]` default to 0 for the lower bound and len(slice) for the upper bound when either or both are not specified. Fix the code to use `cap(slice)` to check for the slice capacity.	2020-08-16 11:34:01 +02:00
Michael Eischer	7042bafea5	prune: Abort repacking when a pack contains a wrong blob If a blob in a pack file can be decrypted successfully but contains data that results in a different hash than stated in the header pack, then abort repacking. As both the pack header and the blob are cryptographically verified this either means than a malicious entity tampered with the backup or indicates hardware problems on the client. prune should fail with an error in both cases.	2020-08-16 11:34:01 +02:00
aawsome	0fed6a8dfc	Use "pack file" instead of "data file" (#2885 ) - changed variable names, especially changed DataFile into PackFile - changed in some comments - always use "pack file" in docu	2020-08-16 11:16:38 +02:00
MichaelEischer	eca0f0ad24	Merge pull request #2863 from aawsome/index-no-duplicates Don't save exact duplicates in merged index	2020-08-08 18:24:14 +02:00
Alexander Weiss	b112533812	Don't save exact duplicates when merging indexes	2020-08-05 06:32:02 +02:00
Alexander Weiss	5e63294355	Add benchmark MasterIndexAlloc	2020-08-05 06:32:02 +02:00
Michael Eischer	05116e4787	prune: Cleanup progress bar handling while repacking	2020-08-03 19:32:46 +02:00
Alexander Neumann	2580eef2aa	Merge pull request #2318 from classmarkets/2175-named-keys Allow specifying user and host when adding keys	2020-08-01 13:06:31 +02:00
Michael Eischer	c847aace35	Rename Index interface to MasterIndex The interface is now only implemented by repository.MasterIndex.	2020-07-25 21:19:46 +02:00
Alexander Weiss	9d1fb94c6c	make Lookup() return all blobs + simplify syntax	2020-07-25 21:18:34 +02:00
greatroar	309598c237	Simplify sortCachedPacksFirst test in internal/repository The test now uses the fact that the sort is stable. It's not guaranteed to be, but the test is cleaner and more exhaustive. sortCachedPacksFirst no longer needs a return value.	2020-07-25 12:12:59 +02:00
greatroar	03d23e6faa	Speed up blob sorting in internal/repository name old time/op new time/op delta SortCachedPacksFirst-8 208µs ± 3% 186µs ± 3% -10.74% (p=0.000 n=10+8) name old alloc/op new alloc/op delta SortCachedPacksFirst-8 213kB ± 0% 139kB ± 0% -34.62% (p=0.000 n=10+10) name old allocs/op new allocs/op delta SortCachedPacksFirst-8 1.03k ± 0% 1.03k ± 0% -0.19% (p=0.000 n=10+10)	2020-07-25 12:12:59 +02:00
greatroar	b10acd2af7	Test and benchmark blob sorting in internal/repository	2020-07-25 12:12:58 +02:00
Alexander Weiss	a666a6d576	Add tests and merge indexes in index benchmarks	2020-07-22 21:54:02 +02:00
Alexander Weiss	e388d962a5	Merge final indexes together for faster index access	2020-07-22 21:54:02 +02:00
Alexander Weiss	3b7a3711e6	Add more realistic index benchmarks + reduce test size of BenchmarkMasterIndexLookupParallel	2020-07-21 07:18:20 +02:00
greatroar	7bda28f31f	Chaining hash table for repository.Index These are faster to construct but slower to access. The allocation rate is halved, the peak memory usage almost halved compared to standard map. Benchmark results on linux/amd64, -benchtime=3s -count=20: name old time/op new time/op delta PackerManager-8 178ms ± 0% 178ms ± 0% ~ (p=0.231 n=20+20) DecodeIndex-8 4.54s ± 0% 4.30s ± 0% -5.20% (p=0.000 n=18+17) DecodeIndexParallel-8 4.54s ± 0% 4.30s ± 0% -5.22% (p=0.000 n=19+18) IndexHasUnknown-8 44.4ns ± 5% 50.5ns ±11% +13.82% (p=0.000 n=19+17) IndexHasKnown-8 48.3ns ± 0% 51.5ns ±12% +6.68% (p=0.001 n=16+20) IndexAlloc-8 758ms ± 1% 616ms ± 1% -18.69% (p=0.000 n=19+19) IndexAllocParallel-8 234ms ± 3% 204ms ± 2% -12.60% (p=0.000 n=20+18) MasterIndexLookupSingleIndex-8 122ns ± 0% 145ns ± 9% +18.44% (p=0.000 n=14+20) MasterIndexLookupMultipleIndex-8 369ns ± 2% 429ns ± 8% +16.27% (p=0.000 n=20+20) MasterIndexLookupSingleIndexUnknown-8 68.4ns ± 5% 74.9ns ±13% +9.47% (p=0.000 n=20+20) MasterIndexLookupMultipleIndexUnknown-8 315ns ± 3% 369ns ±11% +17.14% (p=0.000 n=20+20) MasterIndexLookupParallel/known,indices=5-8 743ns ± 1% 816ns ± 2% +9.87% (p=0.000 n=17+17) MasterIndexLookupParallel/unknown,indices=5-8 238ns ± 1% 260ns ± 2% +9.14% (p=0.000 n=19+20) MasterIndexLookupParallel/known,indices=10-8 1.01µs ± 3% 1.11µs ± 2% +9.79% (p=0.000 n=19+20) MasterIndexLookupParallel/unknown,indices=10-8 222ns ± 0% 269ns ± 2% +20.83% (p=0.000 n=16+20) MasterIndexLookupParallel/known,indices=20-8 1.06µs ± 2% 1.19µs ± 2% +12.95% (p=0.000 n=19+18) MasterIndexLookupParallel/unknown,indices=20-8 413ns ± 1% 530ns ± 1% +28.19% (p=0.000 n=18+20) SaveAndEncrypt-8 30.2ms ± 1% 30.4ms ± 0% +0.71% (p=0.000 n=19+19) LoadTree-8 540µs ± 1% 576µs ± 1% +6.73% (p=0.000 n=20+20) LoadBlob-8 5.64ms ± 0% 5.64ms ± 0% ~ (p=0.883 n=18+17) LoadAndDecrypt-8 5.93ms ± 0% 5.95ms ± 1% ~ (p=0.247 n=20+19) LoadIndex-8 25.1ms ± 0% 24.5ms ± 1% -2.54% (p=0.000 n=18+17) name old speed new speed delta PackerManager-8 296MB/s ± 0% 296MB/s ± 0% ~ (p=0.229 n=20+20) SaveAndEncrypt-8 139MB/s ± 1% 138MB/s ± 0% -0.71% (p=0.000 n=19+19) LoadBlob-8 177MB/s ± 0% 177MB/s ± 0% ~ (p=0.890 n=18+17) LoadAndDecrypt-8 169MB/s ± 0% 168MB/s ± 1% ~ (p=0.227 n=20+19) name old alloc/op new alloc/op delta PackerManager-8 91.8kB ± 0% 91.8kB ± 0% ~ (p=0.772 n=12+19) IndexAlloc-8 786MB ± 0% 400MB ± 0% -49.04% (p=0.000 n=20+18) IndexAllocParallel-8 786MB ± 0% 401MB ± 0% -49.04% (p=0.000 n=19+15) SaveAndEncrypt-8 21.0MB ± 0% 21.0MB ± 0% +0.00% (p=0.000 n=19+19) name old allocs/op new allocs/op delta PackerManager-8 1.41k ± 0% 1.41k ± 0% ~ (all equal) IndexAlloc-8 977k ± 0% 907k ± 0% -7.18% (p=0.000 n=20+20) IndexAllocParallel-8 977k ± 0% 907k ± 0% -7.17% (p=0.000 n=19+15) SaveAndEncrypt-8 73.0 ± 0% 73.0 ± 0% ~ (all equal)	2020-07-19 13:58:22 +02:00
greatroar	255ba83c4b	Parallel index benchmarks + benchmark optimizations createRandomIndex was using the global RNG, which locks on every call It was also using twice as many random numbers as necessary and doing a float division in every iteration of the inner loop. BenchmarkDecodeIndex was using too short an input, especially for a parallel version. (It may now be using one that is a bit large.) Results on linux/amd64, -benchtime=3s -count=20: name old time/op new time/op delta PackerManager-8 178ms ± 0% 178ms ± 0% ~ (p=0.165 n=20+20) DecodeIndex-8 13.6µs ± 2% 4539886.8µs ± 0% +33293901.38% (p=0.000 n=20+18) IndexHasUnknown-8 44.4ns ± 7% 44.4ns ± 5% ~ (p=0.873 n=20+19) IndexHasKnown-8 49.2ns ± 3% 48.3ns ± 0% -1.86% (p=0.000 n=20+16) IndexAlloc-8 802ms ± 1% 758ms ± 1% -5.51% (p=0.000 n=20+19) MasterIndexLookupSingleIndex-8 124ns ± 1% 122ns ± 0% -1.41% (p=0.000 n=20+14) MasterIndexLookupMultipleIndex-8 373ns ± 2% 369ns ± 2% -1.13% (p=0.001 n=20+20) MasterIndexLookupSingleIndexUnknown-8 67.8ns ± 3% 68.4ns ± 5% ~ (p=0.753 n=20+20) MasterIndexLookupMultipleIndexUnknown-8 316ns ± 3% 315ns ± 3% ~ (p=0.846 n=20+20) SaveAndEncrypt-8 30.5ms ± 1% 30.2ms ± 1% -1.09% (p=0.000 n=19+19) LoadTree-8 527µs ± 1% 540µs ± 1% +2.37% (p=0.000 n=19+20) LoadBlob-8 5.65ms ± 0% 5.64ms ± 0% -0.21% (p=0.000 n=19+18) LoadAndDecrypt-8 7.07ms ± 2% 5.93ms ± 0% -16.15% (p=0.000 n=19+20) LoadIndex-8 32.1ms ± 2% 25.1ms ± 0% -21.64% (p=0.000 n=20+18) name old speed new speed delta PackerManager-8 296MB/s ± 0% 296MB/s ± 0% ~ (p=0.159 n=20+20) SaveAndEncrypt-8 138MB/s ± 1% 139MB/s ± 1% +1.10% (p=0.000 n=19+19) LoadBlob-8 177MB/s ± 0% 177MB/s ± 0% +0.21% (p=0.000 n=19+18) LoadAndDecrypt-8 141MB/s ± 2% 169MB/s ± 0% +19.24% (p=0.000 n=19+20) name old alloc/op new alloc/op delta PackerManager-8 91.8kB ± 0% 91.8kB ± 0% ~ (p=0.826 n=19+12) IndexAlloc-8 786MB ± 0% 786MB ± 0% +0.01% (p=0.000 n=20+20) SaveAndEncrypt-8 21.0MB ± 0% 21.0MB ± 0% -0.00% (p=0.012 n=20+19) name old allocs/op new allocs/op delta PackerManager-8 1.41k ± 0% 1.41k ± 0% ~ (all equal) IndexAlloc-8 977k ± 0% 977k ± 0% +0.01% (p=0.022 n=20+20) SaveAndEncrypt-8 73.0 ± 0% 73.0 ± 0% ~ (all equal)	2020-07-19 13:58:05 +02:00
greatroar	02bec13ef2	Fix repository_test.BenchmarkSaveAndEncrypt The benchmark was actually testing the speed of index lookups. name old time/op new time/op delta SaveAndEncrypt-8 101ns ± 2% 31505824ns ± 1% +31311591.31% (p=0.000 n=10+10) name old speed new speed delta SaveAndEncrypt-8 41.7TB/s ± 2% 0.0TB/s ± 1% -100.00% (p=0.000 n=10+10) name old alloc/op new alloc/op delta SaveAndEncrypt-8 1.00B ± 0% 20989508.40B ± 0% +2098950740.00% (p=0.000 n=10+10) name old allocs/op new allocs/op delta SaveAndEncrypt-8 0.00 123.00 ± 0% +Inf% (p=0.000 n=10+9) (The actual speed is ca. 131MiB/s.)	2020-07-05 17:41:42 +02:00
Alexander Weiss	1361341c58	don't save duplicate packIDs when using internal/repository/Index.Store	2020-06-14 07:56:24 +02:00
Alexander Weiss	ce4a2f4ca6	save packIDs and duplicates separately A side remark to the definition of Index.blob: Another possibility would have been to use: blob map[restic.BlobHandle]indexEntry This would have led to the following sizes: key: 32 + 1 = 33 bytes value: 8 bytes indexEntry: 8 + 4 + 4 = 16 bytes each packID: 32 bytes To save N index entries, we would therefore have needed: N OF * (33 + 8) bytes + N * 16 + N * 32 bytes / BP = N * 82 bytes More precicely, using a pointer instead of a direct entry is the better memory choice if: OF * 8 bytes + entrysize < OF * entrysize <=> entrysize > 8 bytes * OF/(OF-1) Under the assumption of OF=1.5, this means using pointers would have been the better choice if sizeof(indexEntry) > 24 bytes.	2020-06-14 07:56:21 +02:00
Alexander Weiss	cf979e2b81	make offset and length uint32	2020-06-14 07:50:19 +02:00
Michael Eischer	d92e2c5769	simplify index code	2020-06-14 07:50:19 +02:00
Alexander Weiss	7419844885	add changelog, benchmark, memory calculation	2020-06-14 07:50:15 +02:00
Alexander Weiss	d3c59d18e5	Fix inconsistency of saving/loading config file Fix saving/loading config file: Always set ID to a zero ID.	2020-06-13 16:30:23 +02:00
MichaelEischer	dd7b4f54f5	Merge pull request #2709 from greatroar/minio-sha256 Use Minio's optimized SHA-256	2020-06-12 23:32:58 +02:00
MichaelEischer	735a8074d5	Merge pull request #2773 from aawsome/index-uploads+knownblobs Fix non-intuitive repo behavior	2020-06-12 22:41:04 +02:00
Alexander Weiss	70347e95d5	disable index uploads for prune command + modifications of changelog	2020-06-12 09:24:38 +02:00
Alexander Weiss	91906911b0	Fix non-intuitive repository behavior - The SaveBlob method now checks for duplicates. - Moves handling of pending blobs to MasterIndex. -> also cleans up pending index entries when they are saved in the index -> when using SaveBlob no need to care about index any longer - Always check for full index and save it when storing packs. -> removes the need of an index uploader -> also removes the verbose "uploaded intermediate index" messages - The Flush method now also saves the index - Fix race condition when checking and saving full/non-finalized indexes	2020-06-11 13:05:23 +02:00
MichaelEischer	6856d1e422	Merge pull request #2749 from aawsome/fix-fullindex Change condition for full index	2020-06-10 20:40:19 +02:00
Alexander Weiss	8c1261ff02	changed condition for full index	2020-06-07 22:00:49 +02:00
greatroar	f97a680887	Fix repository benchmarks BenchmarkLoad{AndDecrypt,Blob} were spending between 38% and 50% of their time measuring SHA-256 performance in their checks.	2020-04-28 07:57:29 +02:00
greatroar	42a3db05b0	Use Minio's optimized SHA-256 internal/repository benchmarks on an Intel i7-3770k: name old speed new speed delta PackerManager-8 209MB/s ± 1% 291MB/s ± 1% +38.94% (p=0.008 n=5+5) SaveAndEncrypt-8 112MB/s ± 1% 135MB/s ± 1% +20.25% (p=0.008 n=5+5)	2020-04-28 07:57:18 +02:00

1 2 3 4 5 ...

330 Commits