restic

mirror of https://github.com/octoleo/restic.git synced 2024-12-02 18:08:28 +00:00

Author	SHA1	Message	Date
Michael Eischer	fbcbd5318c	repository: extract LoadTree/SaveTree The repository has no real idea what a Tree is. So these methods never belonged there.	2022-07-17 13:11:28 +02:00
Lorenz Bausch	d6e3c7f28e	Wording: change repo to repository	2022-07-08 20:05:35 +02:00
Michael Eischer	6f53ecc1ae	adapt workers based on whether an operation is CPU or IO-bound Use runtime.GOMAXPROCS(0) as worker count for CPU-bound tasks, repo.Connections() for IO-bound task and a combination if a task can be both. Streaming packs is treated as IO-bound as adding more worker cannot provide a speedup. Typical IO-bound tasks are download / uploading / deleting files. Decoding / Encoding / Verifying are usually CPU-bound. Several tasks are a combination of both, e.g. for combined download and decode functions. In the latter case add both limits together. As the backends have their own concurrency limits restic still won't download more than repo.Connections() files in parallel, but the additional workers can decode already downloaded data in parallel.	2022-07-03 12:19:26 +02:00
Michael Eischer	753e56ee29	repository: Limit to a single pending pack file Use only a single not completed pack file to keep the number of open and active pack files low. The main change here is to defer hashing the pack file to the upload step. This prevents the pack assembly step to become a bottleneck as the only task is now to write data to the temporary pack file. The tests are cleaned up to no longer reimplement packer manager functions.	2022-07-02 22:42:34 +02:00
Michael Eischer	120ccc8754	repository: Rework blob saving to use an async pack uploader Previously, SaveAndEncrypt would assemble blobs into packs and either return immediately if the pack is not yet full or upload the pack file otherwise. The upload will block the current goroutine until it finishes. Now, the upload is done using separate goroutines. This requires changes to the error handling. As uploads are no longer tied to a SaveAndEncrypt call, failed uploads are signaled using an errgroup. To count the uploaded amount of data, the pack header overhead is no longer returned by `packer.Finalize` but rather by `packer.HeaderOverhead`. This helper method is necessary to continue returning the pack header overhead directly to the responsible call to `repository.SaveBlob`. Without the method this would not be possible, as packs are finalized asynchronously.	2022-07-02 22:42:34 +02:00
Michael Eischer	a6e9e08034	Account for pack header overhead at each entry This will miss the pack header crypto overhead and the length field, which only amount to a few bytes per pack file.	2022-07-02 18:55:58 +02:00
Alexander Neumann	99634c0936	Return real size from SaveBlob	2022-07-02 18:55:12 +02:00
Michael Eischer	ec7c9ce88b	drop unused repository.Loader interface	2022-07-02 18:39:59 +02:00
Michael Eischer	e68c3a4e62	repository: simplify CreateIndexFromPacks	2022-07-02 18:39:59 +02:00
Michael Eischer	bf81bf0795	repository: Properly set id for finalized index As MergeFinalIndex and index uploads can occur concurrently, it is necessary for MergeFinalIndex to check whether the IDs for an index were already set before merging it. Otherwise, we'd loose the ID of an index which is set _after_ uploading it.	2022-07-02 18:39:59 +02:00
Michael Eischer	628ae799ca	repository: make flushPacks private	2022-07-02 18:39:12 +02:00
Michael Eischer	a77d5c4d11	repository: index saving belongs into the MasterIndex	2022-07-02 18:38:56 +02:00
greatroar	c9557b2822	internal/repository: Fix LoadBlob + fuzz test When given a buf that is big enough for a compressed blob but not its decompressed contents, the copy at the end of LoadBlob would skip the last part of the contents. Fixes #3783.	2022-06-06 17:02:28 +02:00
greatroar	2e0f1f5113	repository: Remove RunWorkers, report ctx.Err() This removes RunWorkers, which had become mere overhead by successive refactors. It also ensures that each former user of that function returns any context error that occurs, so failure to complete an operation is always reported as an error.	2022-05-10 22:26:00 +02:00
Michael Eischer	cf5cb673fb	repository: Use existing method to collect pack ids	2022-04-30 19:14:21 +02:00
Michael Eischer	b335cb6285	repository: Refactor index IDs collection	2022-04-30 19:14:21 +02:00
Michael Eischer	abe5935693	repository: unify repository version-specific initialization Mark the master index as compressed also when initializing a new repository. This is only relevant for testing.	2022-04-30 11:34:10 +02:00
Alexander Neumann	8776031f96	Leave allocating slices to the decompress code	2022-04-30 11:34:10 +02:00
Alexander Neumann	5eb05a0afe	Configure zstd encoder/decoder	2022-04-30 11:34:10 +02:00
Michael Eischer	2f36e044db	Cleanup pack header check	2022-04-30 11:34:10 +02:00
Alexander Neumann	8b11b86383	Add option global --compression	2022-04-30 11:34:10 +02:00
Michael Eischer	7132df529e	repository: Increase index size for repo version 2 A compressed index is only about one third the size of an uncompressed one. Thus increase the number of entries in an index to avoid cluttering the repository with small indexes.	2022-04-30 11:34:10 +02:00
Michael Eischer	66f9048bce	repository: Alloc zstd encoder/decoder on demand	2022-04-30 11:34:10 +02:00
Michael Eischer	6fb408d90e	repository: implement pack compression	2022-04-30 11:34:10 +02:00
Michael Eischer	362ab06023	init: Add flag to specify created repository version	2022-04-30 10:07:42 +02:00
Michael Eischer	4b957e7373	repository: Implement index/snapshot/lock compression The config file is not compressed as it should remain readable by older restic versions such that these can return a proper error. As the old format for unpacked data does not include a version header, make use of a trick: The old data is always encoded as JSON. Thus it can only start with '{' or '['. For any other value the first byte indicates a versioned format. The version is set to 2 for now. Then the zstd compressed data follows.	2022-04-30 10:07:42 +02:00
Michael Eischer	c2aabb2686	Print used key name if config fails to load	2022-04-09 22:38:18 +02:00
Michael Eischer	dc3d77dacc	repository: make saveAndEncrypt private	2022-03-28 22:09:49 +02:00
Michael Eischer	6877e7edbb	repository: Rename LoadAndDecrypt to LoadUnpacked The method is the complement for SaveUnpacked and not for SaveAndEncrypt. The latter assembles blobs into pack files.	2022-03-28 22:09:49 +02:00
Michael Eischer	bba8ba7a5b	repository: cancel streampack context after error	2022-02-12 20:18:25 +01:00
Michael Eischer	47554a3428	repository: Fix error handling in repack When storing a blob fails, this is a fatal error which must not be retried.	2022-02-12 20:18:25 +01:00
Michael Eischer	930a00ad54	checker: reuse bufio reader	2022-02-12 20:18:25 +01:00
Michael Eischer	34ebafb8b6	repository: don't crash if blob size is too short	2022-02-12 20:18:25 +01:00
Michael Eischer	becebf5d88	repository: remove unused DownloadAndHash	2022-02-12 20:18:25 +01:00
Michael Eischer	f1e58e7c7f	checker: rewrite ReadData to stream packs	2022-02-12 20:18:25 +01:00
Michael Eischer	f40abd92fa	restorer: convert to use StreamPack	2022-02-12 20:18:25 +01:00
Michael Eischer	c4a2bfcb39	repository: Add StreamPacks function The function supports efficiently loading a specified list of blobs from a single pack in a streaming fashion. That is there's no need for temporary files independent of the pack size.	2022-02-12 20:18:25 +01:00
Alexander Weiss	81876d5c1b	Simplify cache logic	2021-09-03 21:01:00 +02:00
Michael Eischer	9aa2eff384	Add plumbing to calculate backend specific file hash for upload This enables the backends to request the calculation of a backend-specific hash. For the currently supported backends this will always be MD5. The hash calculation happens as early as possible, for pack files this is during assembly of the pack file. That way the hash would even capture corruptions of the temporary pack file on disk.	2021-08-04 22:17:46 +02:00
Ryan Hitchman	77bf148460	backup: add --dry-run/-n flag to show what would happen. This can be used to check how large a backup is or validate exclusions. It does not actually write any data to the underlying backend. This is implemented as a simple overlay backend that accepts writes without forwarding them, passes through reads, and generally does the minimal necessary to pretend that progress is actually happening. Fixes #1542 Example usage: $ restic -vv --dry-run . \| grep add new /changelog/unreleased/issue-1542, saved in 0.000s (350 B added) modified /cmd/restic/cmd_backup.go, saved in 0.000s (16.543 KiB added) modified /cmd/restic/global.go, saved in 0.000s (0 B added) new /internal/backend/dry/dry_backend_test.go, saved in 0.000s (3.866 KiB added) new /internal/backend/dry/dry_backend.go, saved in 0.000s (3.744 KiB added) modified /internal/backend/test/tests.go, saved in 0.000s (0 B added) modified /internal/repository/repository.go, saved in 0.000s (20.707 KiB added) modified /internal/ui/backup.go, saved in 0.000s (9.110 KiB added) modified /internal/ui/jsonstatus/status.go, saved in 0.001s (11.055 KiB added) modified /restic, saved in 0.131s (25.542 MiB added) Would add to the repo: 25.892 MiB	2021-08-04 21:19:29 +02:00
Alexander Neumann	aef3658a5f	Address review comments	2021-01-30 20:02:37 +01:00
Alexander Neumann	16313bfcc9	errcheck: Add error check for MergeFinalIndexes()	2021-01-30 20:02:37 +01:00
Alexander Neumann	75f53955ee	errcheck: Add error checks Most added checks are straight forward.	2021-01-30 20:02:37 +01:00
Michael Eischer	a12c5f1d37	repository: move otherwise unused LoadIndex to tests	2020-12-22 22:36:18 +01:00
Michael Eischer	24474a36f4	repository: deduplicate index loading implementation	2020-12-22 22:36:18 +01:00
Alexander Neumann	36c5d39c2c	Fix issues reported by semgrep	2020-12-11 09:41:59 +01:00
Alexander Neumann	7facc8ccc1	Merge pull request #2505 from aawsome/fix-repo-configfile Fix repo configfile	2020-12-07 07:52:37 +01:00
Alexander Weiss	aa7a5f19c2	Use BlobHandle in index methods	2020-11-22 20:41:12 +01:00
Alexander Weiss	c3ddde9e7d	Return hdrSize in ListPack	2020-11-21 22:13:54 +01:00
Alexander Weiss	43732bb885	Add CreateIndexFromPacks()	2020-11-15 07:04:51 +01:00
Alexander Weiss	fd33030556	Use in-memory index to rebuild index in prune	2020-11-06 20:23:30 +01:00
Alexander Neumann	a4507610a0	Fix typo	2020-11-05 10:31:49 +01:00
Alexander Weiss	b44ecde8b0	Fix setting of ID in DecodeIndex	2020-10-17 09:12:58 +02:00
MichaelEischer	4ba237bb93	Merge pull request #3019 from greatroar/refactor-decodeindex Refactor index decoding	2020-10-15 23:22:33 +02:00
greatroar	b27375f5ce	defer close(ch) outside repository.RunWorkers	2020-10-14 15:50:16 +02:00
greatroar	27db3ec262	Refactor index decoding Decoding old-format indices no longer requires loading and decrypting twice.	2020-10-13 20:47:50 +02:00
Alexander Neumann	56883817d8	Merge pull request #2990 from MichaelEischer/fix-goreport-warnings Fix some goreport warnings	2020-10-12 20:44:56 +02:00
Michael Eischer	e638b46a13	Embed context into ReaderAt The io.Reader interface does not support contexts, such that it is necessary to embed the context into the backendReaderAt struct. This has the problem that a reader might suddenly stop working when it's contained context is canceled. However, this is now problem here as the reader instances never escape the calling function.	2020-10-09 22:39:07 +02:00
Michael Eischer	a449450021	init: pass proper context to master key generation This is no change in behavior as a canceled context did later on cause the config file creation to fail. Therefore this change just lets the repository initialization fail a bit earlier.	2020-10-09 22:37:56 +02:00
Michael Eischer	c458e114d4	pass context to Find / FindSnapshot This allows proper interruption of restic while it searches for snapshots or key files.	2020-10-09 22:37:56 +02:00
Michael Eischer	eba5dd831f	Fix typos reported by misspell	2020-10-06 14:55:13 +02:00
Michael Eischer	f003410402	init: Add `--copy-chunker-params` option This allows creating multiple repositories with identical chunker parameters which is required for working deduplication when copying snapshots between different repositories.	2020-09-19 16:53:05 +02:00
greatroar	23fcbb275a	Remove restic.Cache interface It was used in one code path, which then asserted its concrete type as *cache.Cache. Privatised some of the interface methods.	2020-09-18 10:48:13 +02:00
Michael Eischer	d0329cf3eb	Adjust comments to match name of exported methods	2020-09-05 10:07:16 +02:00
Michael Eischer	4784540f04	repository: Simplify worker group code	2020-09-05 10:07:16 +02:00
aawsome	0fed6a8dfc	Use "pack file" instead of "data file" (#2885 ) - changed variable names, especially changed DataFile into PackFile - changed in some comments - always use "pack file" in docu	2020-08-16 11:16:38 +02:00
Michael Eischer	c847aace35	Rename Index interface to MasterIndex The interface is now only implemented by repository.MasterIndex.	2020-07-25 21:19:46 +02:00
Alexander Weiss	9d1fb94c6c	make Lookup() return all blobs + simplify syntax	2020-07-25 21:18:34 +02:00
greatroar	309598c237	Simplify sortCachedPacksFirst test in internal/repository The test now uses the fact that the sort is stable. It's not guaranteed to be, but the test is cleaner and more exhaustive. sortCachedPacksFirst no longer needs a return value.	2020-07-25 12:12:59 +02:00
greatroar	03d23e6faa	Speed up blob sorting in internal/repository name old time/op new time/op delta SortCachedPacksFirst-8 208µs ± 3% 186µs ± 3% -10.74% (p=0.000 n=10+8) name old alloc/op new alloc/op delta SortCachedPacksFirst-8 213kB ± 0% 139kB ± 0% -34.62% (p=0.000 n=10+10) name old allocs/op new allocs/op delta SortCachedPacksFirst-8 1.03k ± 0% 1.03k ± 0% -0.19% (p=0.000 n=10+10)	2020-07-25 12:12:59 +02:00
greatroar	b10acd2af7	Test and benchmark blob sorting in internal/repository	2020-07-25 12:12:58 +02:00
Alexander Weiss	e388d962a5	Merge final indexes together for faster index access	2020-07-22 21:54:02 +02:00
Alexander Weiss	d3c59d18e5	Fix inconsistency of saving/loading config file Fix saving/loading config file: Always set ID to a zero ID.	2020-06-13 16:30:23 +02:00
MichaelEischer	dd7b4f54f5	Merge pull request #2709 from greatroar/minio-sha256 Use Minio's optimized SHA-256	2020-06-12 23:32:58 +02:00
Alexander Weiss	70347e95d5	disable index uploads for prune command + modifications of changelog	2020-06-12 09:24:38 +02:00
Alexander Weiss	91906911b0	Fix non-intuitive repository behavior - The SaveBlob method now checks for duplicates. - Moves handling of pending blobs to MasterIndex. -> also cleans up pending index entries when they are saved in the index -> when using SaveBlob no need to care about index any longer - Always check for full index and save it when storing packs. -> removes the need of an index uploader -> also removes the verbose "uploaded intermediate index" messages - The Flush method now also saves the index - Fix race condition when checking and saving full/non-finalized indexes	2020-06-11 13:05:23 +02:00
greatroar	42a3db05b0	Use Minio's optimized SHA-256 internal/repository benchmarks on an Intel i7-3770k: name old speed new speed delta PackerManager-8 209MB/s ± 1% 291MB/s ± 1% +38.94% (p=0.008 n=5+5) SaveAndEncrypt-8 112MB/s ± 1% 135MB/s ± 1% +20.25% (p=0.008 n=5+5)	2020-04-28 07:57:18 +02:00
greatroar	e7d7b85d59	Merge Repository.{LoadBlob,loadBlob} Pushing the allocation logic down into the former loadBlob body means that fewer allocations have to be performed: name old time/op new time/op delta LoadTree-8 478µs ± 1% 481µs ± 2% ~ (p=0.315 n=9+10) LoadBlob-8 11.6ms ± 1% 11.6ms ± 2% ~ (p=0.393 n=10+10) LoadAndDecrypt-8 13.3ms ± 3% 13.3ms ± 3% ~ (p=0.905 n=10+9) LoadIndex-8 33.6ms ± 2% 33.2ms ± 1% -1.15% (p=0.028 n=10+9) name old alloc/op new alloc/op delta LoadTree-8 41.2kB ± 0% 41.1kB ± 0% -0.23% (p=0.000 n=10+10) LoadBlob-8 2.28kB ± 0% 2.18kB ± 0% -4.21% (p=0.000 n=10+10) LoadAndDecrypt-8 2.10MB ± 0% 2.10MB ± 0% ~ (all equal) LoadIndex-8 5.22MB ± 0% 5.22MB ± 0% ~ (p=0.631 n=10+10) name old allocs/op new allocs/op delta LoadTree-8 652 ± 0% 651 ± 0% -0.15% (p=0.000 n=10+10) LoadBlob-8 24.0 ± 0% 23.0 ± 0% -4.17% (p=0.000 n=10+10) LoadAndDecrypt-8 30.0 ± 0% 30.0 ± 0% ~ (all equal) LoadIndex-8 30.2k ± 0% 30.2k ± 0% ~ (p=0.610 n=10+10) name old speed new speed delta LoadBlob-8 86.4MB/s ± 1% 85.9MB/s ± 2% ~ (p=0.393 n=10+10) LoadAndDecrypt-8 75.4MB/s ± 3% 75.4MB/s ± 3% ~ (p=0.858 n=10+9)	2020-04-23 10:04:20 +02:00
greatroar	be5a0ff59f	Centralize buffer allocation and size checking in Repository.LoadBlob Benchmark results for internal/repository: name old time/op new time/op delta LoadTree-8 479µs ± 2% 478µs ± 1% ~ (p=0.780 n=10+9) LoadBlob-8 11.6ms ± 2% 11.6ms ± 1% ~ (p=0.631 n=10+10) LoadAndDecrypt-8 13.2ms ± 2% 13.3ms ± 3% ~ (p=0.631 n=10+10) name old alloc/op new alloc/op delta LoadTree-8 41.2kB ± 0% 41.2kB ± 0% ~ (all equal) LoadBlob-8 2.28kB ± 0% 2.28kB ± 0% ~ (all equal) LoadAndDecrypt-8 2.10MB ± 0% 2.10MB ± 0% ~ (all equal) name old allocs/op new allocs/op delta LoadTree-8 652 ± 0% 652 ± 0% ~ (all equal) LoadBlob-8 24.0 ± 0% 24.0 ± 0% ~ (all equal) LoadAndDecrypt-8 30.0 ± 0% 30.0 ± 0% ~ (all equal) name old speed new speed delta LoadBlob-8 86.2MB/s ± 2% 86.4MB/s ± 1% ~ (p=0.594 n=10+10) LoadAndDecrypt-8 75.7MB/s ± 2% 75.4MB/s ± 3% ~ (p=0.617 n=10+10)	2020-04-23 10:04:20 +02:00
Michael Eischer	b46cc6d57e	repository: Don't sort one element pack lists When loading a blob, restic first looks up pack files containing the blob. To avoid unnecessary work an already cached pack file is preferred. However, if there is only a single pack file to choose from (which is the normal case) sorting the one-element list won't change anything. Therefore avoid the unnecessary cache check in that case.	2020-03-07 10:26:06 +01:00
greatroar	8526cc6647	Remove sync.Pool from internal/repository The pool was used improperly, causing more allocations to be performed than without it. name old time/op new time/op delta SaveAndEncrypt-8 36.8ms ± 2% 36.9ms ± 2% ~ (p=0.218 n=10+10) name old speed new speed delta SaveAndEncrypt-8 114MB/s ± 2% 114MB/s ± 2% ~ (p=0.218 n=10+10) name old alloc/op new alloc/op delta SaveAndEncrypt-8 21.1MB ± 0% 21.0MB ± 0% -0.44% (p=0.000 n=10+10) name old allocs/op new allocs/op delta SaveAndEncrypt-8 79.0 ± 0% 77.0 ± 0% -2.53% (p=0.000 n=10+10)	2020-02-29 17:54:46 +01:00
Alexander Bruyako	da48b925ff	remove unnecessary error return I was running "golangci-lint" and found this two warnings internal/checker/checker.go:135:18: (Checker).LoadIndex$3 - result 0 (error) is always nil (unparam) final := func() error { ^ internal/repository/repository.go:457:18: (Repository).LoadIndex$3 - result 0 (error) is always nil (unparam) final := func() error { ^ It turns out that these functions are used only in "RunWorkers(...)", which is used only two times in whole project right after this "final" functions. And because these "final" functions always return "nil", I've descided, that it would be better to remove requriments for "final" func to return error to avoid magick "return nil" at their end.	2020-01-27 18:28:21 +03:00
Alexander Neumann	88716794e3	Check errors returned by LoadIndex() Bug was reported in the forum here: https://forum.restic.net/t/check-rebuild-index-prune/1848/13	2019-06-30 21:34:53 +02:00
Alexander Neumann	66efa425bf	Reuse buffer in worker functions	2019-04-13 13:38:39 +02:00
Alexander Neumann	d51e9d1b98	Add []byte to repo.LoadAndDecrypt and utils.LoadAll This commit changes the signatures for repository.LoadAndDecrypt and utils.LoadAll to allow passing in a []byte as the buffer to use. This buffer is enlarged as needed, and returned back to the caller for further use. In later commits, this allows reducing allocations by reusing a buffer for multiple calls, e.g. in a worker function.	2019-04-13 13:38:39 +02:00
Alexander Neumann	e046428c94	Replace FilesInParallel with an errgroup.Group	2019-04-13 13:38:39 +02:00
Chris Howie	1688713400	Add key hinting (#2097 )	2018-11-25 09:13:18 -05:00
Alexander Neumann	bfa18ee8ec	DownloadAndHash: Check error returned by Load()	2018-10-28 21:28:56 +01:00
Alexander Neumann	a56b8fad87	repository: Improve buffer pooling	2018-04-22 11:37:05 +02:00
Alexander Neumann	e68a7fea8a	check: Allow filling the cache during check Closes #1665	2018-04-01 13:59:27 +02:00
Alexander Neumann	2e7ec717c1	repository: Move cache preparation into function	2018-04-01 13:59:27 +02:00
Alexander Neumann	b3e1089cf9	Return error message for config decryption failure See #1663	2018-03-09 21:05:35 +01:00
Alexander Neumann	99f7fd74e3	backend: Improve Save() As mentioned in issue [#1560](https://github.com/restic/restic/pull/1560#issuecomment-364689346) this changes the signature for `backend.Save()`. It now takes a parameter of interface type `RewindReader`, so that the backend implementations or our `RetryBackend` middleware can reset the reader to the beginning and then retry an upload operation. The `RewindReader` interface also provides a `Length()` method, which is used in the backend to get the size of the data to be saved. This removes several ugly hacks we had to do to pull the size back out of the `io.Reader` passed to `Save()` before. In the `s3` and `rest` backend this is actively used.	2018-03-03 15:49:44 +01:00
Alexander Neumann	c323f73bf9	Ignore files in the repo with invalid names Closes #1641	2018-02-26 20:53:38 +01:00
Igor Fedorenko	ab040d8811	Introduced repository.DownloadAndHash helper Signed-off-by: Igor Fedorenko <igor@ifedorenko.com>	2018-02-16 21:13:11 -05:00
Alexander Neumann	ff3de66ddf	Merge pull request #1582 from restic/optimize-debug-log Optimize debug logs	2018-01-26 21:57:18 +01:00
Alexander Neumann	6eb2d76435	index: Lower parallel load to 4	2018-01-26 21:10:38 +01:00
Alexander Neumann	663c57ab4d	debug: Remove manual Str() call Log()	2018-01-25 20:49:41 +01:00
Alexander Neumann	9c55e8d69c	Merge pull request #1549 from MJDSys/more_index_lookup_avoids More optimizations to avoid calling Index.Lookup()	2018-01-24 20:53:30 +01:00
Igor Fedorenko	0084e42cb6	Optimize Repository.ListPack() Use pack file size returned by Backend.List() to avoid extra per-pack Backend.Stat() requests Signed-off-by: Igor Fedorenko <igor@ifedorenko.com>	2018-01-23 22:39:51 -05:00

1 2 3 4

167 Commits