restic

mirror of https://github.com/octoleo/restic.git synced 2024-12-26 04:17:29 +00:00

Author	SHA1	Message	Date
Michael Eischer	6328b7e1f5	replace "too small" with "too short" in error messages	2024-05-18 19:59:26 +02:00
Michael Eischer	ffe5439149	Merge pull request #4605 from MichaelEischer/better-restorer-error-handling Rework repository.StreamPacks & better restorer error handling	2024-05-01 16:37:41 +02:00
Michael Eischer	940a3159b5	let index.Each() and pack.Size() return error on canceled context This forces a caller to actually check that the function did complete.	2024-04-22 22:39:32 +02:00
Michael Eischer	31624aeffd	Improve command shutdown on context cancellation	2024-04-22 22:31:38 +02:00
Michael Eischer	b15d867414	Merge pull request #4763 from MichaelEischer/refactor-prune Refactor repair index / prune into the repository package	2024-04-22 22:24:53 +02:00
Michael Eischer	20d8eed400	repository: streamPack: separate requests for gap larger than 1MB With most cloud providers, traffic is much more expensive than API calls. Thus slightly bias streamPack towards a bit more API calls in exchange for slightly less traffic.	2024-04-22 21:21:23 +02:00
Michael Eischer	cf700d8794	repository: streamPack: reuse zstd decoder	2024-04-22 21:21:23 +02:00
Michael Eischer	666a0b0bdb	repository: streamPack: replace streaming with chunked download Due to the interface of streamPack, we cannot guarantee that operations progress fast enough that the underlying connections remains open. This introduces partial failures which massively complicate the error handling. Switch to a simpler approach that retrieves the pack in chunks of 32MB. If a blob is larger than this limit, then it is downloaded separately. To avoid multiple copies in memory, an auxiliary interface `discardReader` is introduced that allows directly accessing the downloaded byte slices, while still supporting the streaming used by the `check` command.	2024-04-22 21:21:23 +02:00
Michael Eischer	621012dac0	repository: Add blob loading fallback to LoadBlobsFromPack Try to retrieve individual blobs via LoadBlob if streaming did not work.	2024-04-21 21:35:55 +02:00
Michael Eischer	10355c3fb6	repository: Better error message if blob is larger than 4GB	2024-04-19 22:00:35 +02:00
Michael Eischer	09587e6c08	repository: duplicate a few blobs in prune tests	2024-04-14 13:57:19 +02:00
Michael Eischer	defd7ae729	prune/repair index: reset in-memory index after command The current in-memory index becomes stale after prune or repair index have run. Thus, just drop the in-memory index altogether once these commands have finished.	2024-04-14 13:46:24 +02:00
Michael Eischer	038586dc9d	repository: add minimal test for prune	2024-04-14 13:45:17 +02:00
Michael Eischer	d8622c86eb	prune: clean up internal interface	2024-04-14 13:45:15 +02:00
Michael Eischer	8d507c1372	repository: add basic test for RepairIndex	2024-04-14 13:45:15 +02:00
Michael Eischer	310db03c0e	repair index: improve log output if index cannot be deleted The operation will always fail with an error if an index cannot be deleted. Thus, this change is purely cosmetic.	2024-04-14 13:45:13 +02:00
Michael Eischer	7d1b9cde34	repository: use normal Init method in tests	2024-04-14 13:45:11 +02:00
Michael Eischer	b25fc2c89d	repository: remove redundant flushes from tests	2024-04-14 13:45:10 +02:00
Michael Eischer	c65459cd8a	repository: speed up tests	2024-04-14 13:45:10 +02:00
Michael Eischer	4c9a10ca37	repair packs: deduplicate index rebuild	2024-04-14 13:45:02 +02:00
Michael Eischer	85e4021619	prune: move additional option checks to repository	2024-04-14 13:44:58 +02:00
Michael Eischer	fc3b548625	prune: move logic into repository package	2024-04-10 21:30:52 +02:00
Michael Eischer	866ddf5698	repair index: refactor code into repository package	2024-04-10 21:30:52 +02:00
Michael Eischer	5e98f1e2eb	repository: fix test setup race conditions	2024-03-28 23:17:02 +01:00
Michael Eischer	dc441c57a7	repository: unify repository initialization in tests Tests should use a helper from internal/repository/testing.go to construct a Repository object.	2024-03-28 23:17:02 +01:00
Michael Eischer	3ba1fa3cee	repository: remove a few global variables	2024-03-28 23:17:02 +01:00
Michael Eischer	044e8bf821	repository: parallelize lock tests	2024-03-28 23:17:02 +01:00
Michael Eischer	e8df50fa3c	repository: remove global list of locks	2024-03-28 22:46:33 +01:00
Michael Eischer	cbb5f89252	lock: move code to repository package	2024-03-28 22:46:33 +01:00
Michael Eischer	4c3218ef9f	repository: include packID in StreamPack for decrypt/decompress errors	2024-02-17 19:38:01 +01:00
Michael Eischer	18b0bbbf42	repository: use fmt.Errorf in StreamPacks	2024-02-17 19:37:32 +01:00
Alexander Neumann	c0514dd8ba	Fix linter errors (except for tests)	2024-02-10 22:58:10 +01:00
Michael Eischer	5957417b1f	Apply changelog entry / documentation improvements from review	2024-02-04 18:55:41 +01:00
Michael Eischer	86b38a0b17	rename `--no-verify-pack` to `--no-extra-verify`	2024-02-04 17:01:05 +01:00
Michael Eischer	c97a271e89	repository: ask users to report corrupted data while saving blobs	2024-02-04 15:31:42 +01:00
Michael Eischer	193140525c	repository: test verification of blobs/unpacked data	2024-02-04 15:31:42 +01:00
Michael Eischer	2dbb18128c	repository: Allow skipping verification for tests Some tests have to explicitly create pack files with blobs that don't match their ID. For those blobs the builtin verification of the repository must be disabled.	2024-02-03 18:22:47 +01:00
Michael Eischer	30a84e9003	backup: verify unpacked files before upload	2024-02-03 18:22:47 +01:00
Michael Eischer	c01a0c6da7	backup: verify blobs before upload This only covers the blobs themselves, the pack header is not verified so far. Unpacked files are also not covered by the integrity check.	2024-02-03 18:22:47 +01:00
Michael Eischer	16e3f79e8b	repository: make repo.Options configurable for test repos	2024-02-03 18:22:47 +01:00
Michael Eischer	bb92b487f7	repository: fix repack test	2024-02-03 18:22:47 +01:00
Michael Eischer	bfb56b78e1	replace some usages of restic.Repository with more specific interface This should eventually make it easier to test the code.	2024-01-27 13:02:02 +01:00
Michael Eischer	3424088274	Merge pull request #4644 from MichaelEischer/refactor-repair-packs Refactor and test `repair packs`	2024-01-27 13:00:51 +01:00
Michael Eischer	f0e1ad2285	fix linter warning	2024-01-27 12:51:45 +01:00
Michael Eischer	fd579421dd	repository: deduplicate test	2024-01-27 12:51:45 +01:00
Michael Eischer	42c9318b9c	repair pack: add tests	2024-01-27 12:51:45 +01:00
Michael Eischer	764b0bacd6	repair pack: add support for truncated files	2024-01-27 12:51:45 +01:00
Michael Eischer	7c351bc53c	repair pack: reenable auto index updates The method is not available on the restic.Repository interface that is used for testing. Drop the call as a small amount of additional index writes is not a problem.	2024-01-27 12:51:45 +01:00
Michael Eischer	feeab84204	repair pack: extract the repair logic into the repository package Currently, the cmd/restic package contains a significant amount of code that modifies repository internals. This code should in the mid-term move into the repository package.	2024-01-27 12:51:45 +01:00
Michael Eischer	cb50832d50	index: let MasterIndex.Save also delete obsolete indexes	2024-01-27 12:51:08 +01:00
Michael Eischer	c13bf0b607	repository: Introduce RemoveKey function This replaces directly removing keys via the backend.	2024-01-27 12:42:58 +01:00
Michael Eischer	2c310a526e	repository: Replace StreamPack function with LoadBlobsFromPack method LoadBlobsFromPack is now part of the repository struct. This ensures that users of that method don't have to deal will internals of the repository implementation. The filerestorer tests now also contain far fewer pack file implementation details.	2024-01-19 21:40:43 +01:00
Michael Eischer	6b7b5c89e9	repository: prepare StreamPack refactor	2024-01-19 21:40:43 +01:00
Michael Eischer	fb422497af	repository: split StreamPack implementation Move the actual decoding of the pack data into a separate iterator.	2024-01-19 21:39:55 +01:00
Michael Eischer	77b1c52673	repository: test that StreamPack only delivers blobs once	2024-01-07 10:54:53 +01:00
Michael Eischer	fe5c337ca2	repository: StreamPack delivers blobs at most once If an error occurred while streaming a pack file, this could result in passing some of the blobs multiple times to the callback function. This significantly complicates using StreamPack correctly and is unnecessary. Retries do not change the content of a blob and thus only deliver the same result over and over again.	2024-01-07 10:54:49 +01:00
Andrea Gelmini	241916d55b	Fix typos	2023-12-06 13:11:55 +01:00
Michael Eischer	45962c2847	Merge pull request #4499 from MichaelEischer/modular-backend-code Split backend code from restic package	2023-10-27 20:19:20 +02:00
Leo R. Lundgren	aafb806a8c	doc: Correct two typos	2023-10-27 18:56:32 +02:00
Michael Eischer	c7b770eb1f	convert MemorizeList to be repository based Ideally, code that uses a repository shouldn't directly interact with the underlying backend. Thus, move MemorizeList one layer up.	2023-10-25 23:01:35 +02:00
Michael Eischer	1b8a67fe76	move Backend interface to backend package	2023-10-25 23:00:18 +02:00
Michael Eischer	b6d79bdf6f	restic: decouple restic.Handle	2023-10-25 22:54:07 +02:00
Michael Eischer	cb9cbe55d9	repository: store oversized blobs in separate pack files Store oversized blobs in separate pack files as the blobs is large enough to warrant its own pack file. This simplifies the garbage collection of such blobs and keeps the cache smaller, as oversize (tree) blobs only have to be downloaded if they are actually used.	2023-10-17 22:52:16 +02:00
Michael Eischer	3fd0ad7448	repository: list index files only once	2023-10-01 19:53:26 +02:00
arjunajesh	ed65a7dbca	implement progress bar for index loading	2023-10-01 19:52:59 +02:00
Michael Eischer	191c47d30e	Merge pull request #4353 from MichaelEischer/tune-gc Tune Go garbage collector	2023-06-16 23:24:39 +02:00
Michael Eischer	eef0ee7a85	repository: trigger GC after loading the index Loading the index requires some scratch space, thus make sure that this memory does not factor into the targeted gc memory usage limit.	2023-06-02 21:56:14 +02:00
Michael Eischer	ffca602315	repository: Fix panic in benchmarkLoadIndex	2023-05-28 23:55:47 +02:00
Michael Eischer	d1a5ec7839	Rename unused testing parameter to _ The parameter is an additional marker that the test helper must only be used for tests.	2023-05-18 21:17:53 +02:00
Michael Eischer	1514593f22	Remove unused context or testing parameters	2023-05-18 21:17:53 +02:00
Michael Eischer	e01baeabba	Use either test or rtest to refer to internal test helpers A single test file should not use both names.	2023-05-18 21:15:45 +02:00
Michael Eischer	5773b86d02	repository: Push all usage of errors.Fatal out of the package As the `Fatal` error type only includes a string, it becomes impossible to inspect the contained error. This is for a example a problem for the fuse implementation, which must be able to detect context.Canceled errors. Co-authored-by: greatroar <61184462+greatroar@users.noreply.github.com>	2023-05-18 17:27:41 +02:00
greatroar	d129baba7a	repository: Reuse buffers in Repository.LoadUnpacked This method had a buffer argument, but that was nil at all call sites. That's removed, and instead LoadUnpacked now reuses whatever it allocates inside its retry loop.	2023-01-30 22:01:01 +01:00
Michael Eischer	1adf28a2b5	repository: properly return invalid data error in LoadUnpacked The retry backend does not return the original error, if its execution is interrupted by canceling the context. Thus, we have to manually ensure that the invalid data error gets returned. Additionally, use the retry backend for some of the repository tests, as this is the configuration which will be used by restic.	2023-01-14 17:57:02 +01:00
Michael Eischer	6d9675c323	repository: cleanup error message on invalid data The retry printed the filename twice: ``` Load(<lock/04804cba82>, 0, 0) returned error, retrying after 720.254544ms: load(<lock/04804cba82>): invalid data returned ``` now the warning has changed to ``` Load(<lock/04804cba82>, 0, 0) returned error, retrying after 720.254544ms: invalid data returned ```	2023-01-14 17:57:02 +01:00
Michael Eischer	90fb6f70b4	Merge pull request #4089 from greatroar/errors Clean up error handling further	2022-12-24 10:41:56 +01:00
greatroar	b150dd0235	all: Replace some errors.Wrap calls by errors.WithStack Mostly changed the ones that repeat the name of a system call, which is already contained in os.PathError.Op. internal/fs.Reader had to be changed to actually return such errors.	2022-12-17 09:41:07 +01:00
greatroar	c0b5ec55ab	repository: Remove empty cleanup functions in tests TestRepository and its variants always returned no-op cleanup functions. If they ever do need to do cleanup, using testing.T.Cleanup is easier than passing these functions around.	2022-12-11 11:06:25 +01:00
Michael Eischer	40ac678252	backend: remove Test method The Test method was only used in exactly one place, namely when trying to create a new repository it was used to check whether a config file already exists. Use a combination of Stat() and IsNotExist() instead.	2022-12-03 11:28:10 +01:00
Michael Eischer	ff7ef5007e	Replace most usages of ioutil with the underlying function The ioutil functions are deprecated since Go 1.17 and only wrap another library function. Thus directly call the underlying function. This commit only mechanically replaces the function calls.	2022-12-02 19:36:43 +01:00
Michael Eischer	a1eb923876	remove no longer necessary conditional compiles	2022-11-27 13:18:44 +01:00
Alexander Neumann	8dd95b710e	Merge pull request #3992 from MichaelEischer/err-on-invalid-compression Return error if RESTIC_COMPRESSION env variable is invalid	2022-11-04 19:41:34 +01:00
greatroar	137f0bc944	repository: Fix benchmarkSaveAndEncrypt	2022-10-29 23:09:17 +02:00
Michael Eischer	01f0db4e56	return error if RESTIC_COMPRESSION env variable is invalid	2022-10-29 22:03:39 +02:00
Michael Eischer	c4fc5c97f9	prune: Use a single CountedBlobSet to track blobs The set covers necessary, existing and duplicate blobs. This removes the duplicate sets used to track whether all necessary blobs also exist. This reduces the memory usage of prune by about 20-30%.	2022-10-22 18:45:12 +02:00
Michael Eischer	8d62a7adb4	identify keys by ID and not name	2022-10-15 16:07:43 +02:00
Michael Eischer	02634dce7a	restic: change Find to return ids That way consumers no longer have to manually convert the returned name to an id.	2022-10-15 16:06:54 +02:00
Michael Eischer	2e3f1c08c5	repository: split index into a separate package	2022-10-08 21:15:34 +02:00
Michael Eischer	5760ba6989	Merge pull request #3949 from MichaelEischer/simplify-mixedpacks repository: remove IsMixedPack and add replacement for checker	2022-10-08 21:14:14 +02:00
Michael Eischer	4bb5240720	repository: remove unused PrefixLength	2022-10-03 12:15:53 +02:00
Michael Eischer	999fe29976	repository: hide prepareCache	2022-10-03 12:15:53 +02:00
Michael Eischer	ddcf549eba	repository: remove IsMixedPack and add replacement for checker Repositories with mixed packs are probably quite rare by now. When loading data blobs from a mixed pack file, this will no longer trigger caching that file. However, usually tree blobs are accessed first such that this shouldn't make much of a difference. The checker gets a simpler replacement.	2022-10-03 12:03:59 +02:00
Michael Eischer	5c6b6edefe	retry index, lock and snapshot loading on hash mismatch	2022-09-25 11:35:35 +02:00
Michael Eischer	78d2312ee9	Merge pull request #3854 from MichaelEischer/sparsefiles restore: Add support for sparse files	2022-09-24 22:04:02 +02:00
Michael Eischer	c147422ba5	repository: special case SaveBlob for all zero chunks Sparse files contain large regions containing only zero bytes. Checking that a blob only contains zeros is possible with over 100GB/s for modern x86 CPUs. Calculating sha256 hashes is only possible with 500MB/s (or 2GB/s using hardware acceleration). Thus we can speed up the hash calculation for all zero blobs (which always have length chunker.MinSize) by checking for zero bytes and then using the precomputed hash. The all zeros check is only performed for blobs with the minimal chunk size, and thus should add no overhead most of the time. For chunks which are not all zero but have the minimal chunks size, the overhead will be below 2% based on the above performance numbers. This allows reading sparse sections of files as fast as the kernel can return data to us. On my system using BTRFS this resulted in about 4GB/s.	2022-09-24 21:39:39 +02:00
Michael Eischer	1ebd57247a	repository: optimize MasterIndex.Each Sending data through a channel at very high frequency is extremely inefficient. Thus use simple callbacks instead of channels. > name old time/op new time/op delta > MasterIndexEach-16 6.68s ±24% 0.96s ± 2% -85.64% (p=0.008 n=5+5)	2022-09-24 12:21:59 +02:00
Michael Eischer	825b95e313	repository: add benchmark for MasterIndex.Each	2022-09-24 12:21:59 +02:00
Michael Eischer	7682149c9d	repository: cleanup copy connection count check	2022-08-28 11:40:56 +02:00
Michael Eischer	b03277ead5	repository: don't hang when copying using a single connection	2022-08-28 11:40:31 +02:00
MichaelEischer	bee15dd555	Merge pull request #3879 from MichaelEischer/mem-optimize Some random (minor) memory-allocation optimizations	2022-08-26 20:33:02 +02:00

1 2 3 4 5 ...

412 Commits