restic

mirror of https://github.com/octoleo/restic.git synced 2024-12-01 17:23:57 +00:00

Author	SHA1	Message	Date
Michael Eischer	623556bab6	b2: Increase list size to maximum Just request as many files as possible in one call to reduce the number of network roundtrips.	2022-08-21 11:20:03 +02:00
Michael Eischer	de0162ea76	backend/retry: Overwrite failed uploads instead of deleting them For backends which are able to atomically replace files, we just can overwrite the old copy, if it is necessary to retry an upload. This has the benefit of issuing one operation less and might be beneficial if a backend storage, due to bugs or similar, could mix up the order of the upload and delete calls.	2022-08-21 11:14:53 +02:00
Michael Eischer	fc506f8538	b2: Repeat deleting until all file versions are removed When hard deleting the latest file version on B2, this uncovers earlier versions. If an upload required retries, multiple version might exist for a file. Thus to reliably delete a file, we have to remove all versions of it.	2022-08-21 11:11:00 +02:00
Michael Eischer	cc4728d287	repository: Do not report ignored packs in EachByPack Ignored packs were reported as an empty pack by EachByPack. The most immediate effect of this is that the progress bar for rebuilding the index reports processing more packs than actually exist.	2022-08-21 10:38:40 +02:00
Michael Eischer	7a992fc794	repository: Reduce buffer reallocations in ForAllIndexes Previously the buffer was grown incrementally inside `repo.LoadUnpacked`. But we can do better as we already know how large the index will be. Allocate a bit more memory to increase the chance that the buffer can be reused in the future.	2022-08-19 21:13:40 +02:00
Michael Eischer	77b1980d8e	repository: MasterIndex.Packs: reduce allocations	2022-08-19 21:10:43 +02:00
Michael Eischer	6ff9517e45	repository: MasterIndex.ListPacks / Index.EachByPack allow earlier GC Allow earlier garbage collection of some of the intermediate data structures.	2022-08-19 21:06:33 +02:00
Michael Eischer	ce902aac67	cache: Just try to open cache entry without calling stat first Instead of first checking whether a file is in the repository cache and then opening it, we just can open the file. This saves one stat call. If the file is in the cache, everything is fine and otherwise the code follows its normal fallback path.	2022-08-19 20:59:06 +02:00
MichaelEischer	0d9ac78437	Merge pull request #3873 from MichaelEischer/gofmt-comments gofmt comments	2022-08-19 19:54:30 +02:00
MichaelEischer	7e96a5af62	Merge pull request #3872 from MichaelEischer/fuse-fix mount: Only remember successful snapshot refreshes	2022-08-19 19:21:29 +02:00
Michael Eischer	f414db987d	gofmt all files Apparently the rules for comment formatting have changed with go 1.19.	2022-08-19 19:12:26 +02:00
Michael Eischer	522406b4f0	mount: Only remember successful snapshot refreshes If the context provided by the fuse library is canceled before the index was loaded this could lead to missing snapshots.	2022-08-19 19:07:07 +02:00
Michael Eischer	af50fe9ac0	mount: Map slashes in tags to underscores Suggested-by: greatroar <>	2022-08-19 18:17:57 +02:00
Michael Eischer	2ea6c82cf6	comment cleanup gofmt reformatted the comment	2022-08-18 20:15:38 +02:00
Michael Eischer	bb27f7408c	forget: Fail test if duration parsing error is missing	2022-08-18 20:14:09 +02:00
Leo R. Lundgren	6f517858e8	forget: Error when invalid unit is given in duration policy	2022-08-10 13:37:26 +02:00
MichaelEischer	9ad3ad5972	Merge pull request #3850 from lbausch/go1.19 Update tests to Go 1.19	2022-08-07 14:56:17 +02:00
MichaelEischer	2930a102de	Merge pull request #3731 from metalsp0rk/feature/min-packsize-flag Feature: min packsize flag	2022-08-07 14:54:45 +02:00
Michael Eischer	f3fdc66b32	restic: Use stable sorting in snapshot policy sort.Sort is not guaranteed to be stable. Go 1.19 has changed the sorting algorithm which resulted in changes of the sort order. When comparing snapshots with identical timestamp but different paths and tags lists, there is not meaningful order among them. So just keep their order stable.	2022-08-07 14:10:40 +02:00
Michael Eischer	0b7291b8b2	mount: Fix parent inode used by snapshots dir	2022-08-07 13:03:32 +02:00
greatroar	cfa80e2c6b	mount: remove unused inode field from root node	2022-08-07 13:03:26 +02:00
MichaelEischer	74ae76036f	Merge pull request #2913 from aawsome/mount-snapshot-slashes mount: Make snapshots dir structure customizable	2022-08-07 12:27:59 +02:00
Michael Eischer	caa17988a3	fuse: Redesign snapshot dirstruct Cleanly separate the directory presentation and the snapshot directory structure. SnapshotsDir now translates the dirStruct into a format usable by the fuse library and contains only minimal special case rules. All decisions have moved into SnapshotsDirStructure which now creates a fully preassembled tree data structure.	2022-08-07 12:13:06 +02:00
Michael Eischer	1ed775e3a8	debug: support roundtripper logging also for release builds Different from debug builds do not use the eofDetectRoundTripper if logging is disabled.	2022-08-05 23:49:39 +02:00
Michael Eischer	38becfc436	debug: enable debug support for release builds	2022-08-05 23:49:39 +02:00
Michael Eischer	82c268c917	Remove unused hooks mechanism	2022-08-05 23:49:39 +02:00
Michael Eischer	7266f07c87	repository: StreamPack in parts if there are too large gaps For large pack sizes we might be only interested in the first and last blob of a pack file. Thus stream a pack file in multiple parts if the gaps between requested blobs grow too large.	2022-08-05 23:48:36 +02:00
Michael Eischer	7f3b2be1e8	s3: Disable multipart uploads below 200MB	2022-08-05 23:48:36 +02:00
Michael Eischer	1b076cda97	rename option to --pack-size	2022-08-05 23:47:43 +02:00
Kyle Brennan	1e3f05c3f1	repository: prevent header overfill	2022-08-05 23:47:12 +02:00
Michael Eischer	0a6fa602c8	add option for setting min pack size	2022-08-05 23:47:12 +02:00
Michael Eischer	2db7733ee3	fuse: remove unused MetaDir	2022-08-05 23:46:46 +02:00
Michael Eischer	f678f7cb04	fuse: cleanup test	2022-08-05 23:46:46 +02:00
Alexander Weiss	1751afae26	Make snapshots dirs in mount command customizable	2022-08-05 23:46:46 +02:00
Alexander Weiss	57f4003f2f	Generalize fuse snapshot dirs implemetation + allow "/" in tags and snapshot template	2022-08-05 23:46:46 +02:00
Alexander Weiss	696c18e031	Add possibility to set snapshot ID (used in test)	2022-08-05 23:46:46 +02:00
MichaelEischer	04a8ee80fb	Merge pull request #3829 from MichaelEischer/prune-refactor Split prune into slightly small functions	2022-08-05 23:29:52 +02:00
greatroar	ad6ac680af	internal/restic: Handle EINVAL for xattr on Solaris Also make the errors a bit less verbose by not prepending the operation, since pkg/xattr already does that. Old errors looked like Listxattr: xattr.list /myfiles/.zfs/snapshot: invalid argument	2022-08-01 12:45:17 +02:00
Michael Eischer	73053674d9	repository: Test fallback to existing blobs	2022-07-30 17:37:07 +02:00
Michael Eischer	623770eebb	repository: try to recover from invalid blob while repacking If a blob that should be kept is invalid, Repack will now try to request the blob using LoadBlob. Only return an error if that fails.	2022-07-30 17:37:07 +02:00
greatroar	2bdc40e612	Speed up restic init over slow SFTP links pkg/sftp.Client.MkdirAll(d) does a Stat to determine if d exists and is a directory, then a recursive call to create the parent, so the calls for data/?? each take three round trips. Doing a Mkdir first should eliminate two round trips for 255/256 data directories as well as all but one of the top-level directories. Also, we can do all of the calls concurrently. This may reintroduce some of the Stat calls when multiple goroutines try to create the same parent, but at the default number of connections, that should not be much of a problem.	2022-07-30 13:09:08 +02:00
greatroar	23ebec717c	Remove stale comments from backend/sftp The preExec and postExec functions were removed in `0bdb131521` from 2018.	2022-07-30 13:07:25 +02:00
Michael Eischer	4a10ebed15	archiver: reduce memory usage for large files FutureBlob now uses a Take() method as a more memory-efficient way to retrieve the futures result. In addition, futures are now collected while saving the file. As only a limited number of blobs can be queued for uploading, for a large file nearly all FutureBlobs already have their result ready, such that the FutureBlob object just consumes memory.	2022-07-23 14:45:07 +02:00
Michael Eischer	b817681a11	archiver: Incrementally serialize tree nodes That way it is not necessary to keep both the Nodes forming a Tree and the serialized JSON version in memory.	2022-07-23 14:45:07 +02:00
Michael Eischer	c206a101a3	archiver: unify FutureTree/File into futureNode There is no real difference between the FutureTree and FutureFile structs. However, differentiating both increases the size of the FutureNode struct. The FutureNode struct is now only 16 bytes large on 64bit platforms. That way is has a very low overhead if the corresponding file/directory was not processed yet. There is a special case for nodes that were reused from the parent snapshot, as a go channel seems to have 96 bytes overhead which would result in a memory usage regression.	2022-07-23 14:45:07 +02:00
Michael Eischer	32f4997733	archiver: remove unused fileInfo from progress callback	2022-07-23 14:16:23 +02:00
Michael Eischer	dcb00fd2d1	archiver: cleanup Saver interface	2022-07-23 14:16:23 +02:00
Michael Eischer	79321a195c	archiver: remove dead attribute from FutureNode	2022-07-23 14:16:23 +02:00
Michael Eischer	5a6f2f9fa0	Fix S3 legacy layout migration	2022-07-23 11:19:32 +02:00
Michael Eischer	04e49924fb	checker: Fix S3 legacy layout detection	2022-07-23 11:19:32 +02:00
Michael Eischer	fcb3ddf181	check: Complain about usage of s3 legacy layout	2022-07-23 11:19:32 +02:00
Michael Eischer	8b8bd4e8ac	check: complain about mixed pack files	2022-07-23 11:19:32 +02:00
MichaelEischer	443cc49afd	Merge pull request #3830 from MichaelEischer/cleanup-repo Extract Load/SaveTree/JSONUnpacked from repository	2022-07-23 10:46:13 +02:00
MichaelEischer	1f5369e072	Merge pull request #3831 from MichaelEischer/move-code Move code out of the restic package and consolidate backend specific code	2022-07-23 10:33:05 +02:00
Michael Eischer	9729e6d7ef	backend: extract readerat from restic package	2022-07-17 15:29:09 +02:00
Michael Eischer	c44b21d366	restorer: extract hardlinks index from restic package	2022-07-17 13:45:42 +02:00
Michael Eischer	8c11fc3ec9	crypto: move crypto buffer helpers	2022-07-17 13:42:23 +02:00
Michael Eischer	a0cef9f247	limiter: move to internal/backend	2022-07-17 13:40:15 +02:00
Michael Eischer	163ab9c025	mock: move to internal/backend	2022-07-17 13:40:06 +02:00
Michael Eischer	89d3ce852b	repository: extract Load/StoreJSONUnpacked A Load/Store method for each data type is much clearer. As a result the repository no longer needs a method to load / store json.	2022-07-17 13:22:00 +02:00
Michael Eischer	fbcbd5318c	repository: extract LoadTree/SaveTree The repository has no real idea what a Tree is. So these methods never belonged there.	2022-07-17 13:11:28 +02:00
Michael Eischer	5639c41b6a	azure: Strip ? prefix from sas token	2022-07-16 23:55:18 +02:00
Roger Gammans	64a7ec5341	azure: add SAS authentication option	2022-07-16 23:55:18 +02:00
Lorenz Bausch	d6e3c7f28e	Wording: change repo to repository	2022-07-08 20:05:35 +02:00
Michael Eischer	ce89018902	Fix data race in blob_saver After the `BlobSaver` job is submitted, the buffer can be released and reused by another `FileSaver` even before `BlobSaver.Save` returns. That FileSaver will the change `buf.Data` leading to wrong backup statistics. Found by `go test -race ./...`: WARNING: DATA RACE Write at 0x00c0000784a0 by goroutine 41: github.com/restic/restic/internal/archiver.(FileSaver).saveFile() /home/michael/Projekte/restic/restic/internal/archiver/file_saver.go:176 +0x789 github.com/restic/restic/internal/archiver.(FileSaver).worker() /home/michael/Projekte/restic/restic/internal/archiver/file_saver.go:242 +0x2af github.com/restic/restic/internal/archiver.NewFileSaver.func2() /home/michael/Projekte/restic/restic/internal/archiver/file_saver.go:88 +0x5d golang.org/x/sync/errgroup.(Group).Go.func1() /home/michael/go/pkg/mod/golang.org/x/sync@v0.0.0-20210220032951-036812b2e83c/errgroup/errgroup.go:57 +0x91 Previous read at 0x00c0000784a0 by goroutine 29: github.com/restic/restic/internal/archiver.(BlobSaver).Save() /home/michael/Projekte/restic/restic/internal/archiver/blob_saver.go:57 +0x1dd github.com/restic/restic/internal/archiver.(BlobSaver).Save-fm() <autogenerated>:1 +0xac github.com/restic/restic/internal/archiver.(FileSaver).saveFile() /home/michael/Projekte/restic/restic/internal/archiver/file_saver.go:191 +0x855 github.com/restic/restic/internal/archiver.(FileSaver).worker() /home/michael/Projekte/restic/restic/internal/archiver/file_saver.go:242 +0x2af github.com/restic/restic/internal/archiver.NewFileSaver.func2() /home/michael/Projekte/restic/restic/internal/archiver/file_saver.go:88 +0x5d golang.org/x/sync/errgroup.(Group).Go.func1() /home/michael/go/pkg/mod/golang.org/x/sync@v0.0.0-20210220032951-036812b2e83c/errgroup/errgroup.go:57 +0x91	2022-07-03 14:47:53 +02:00
Michael Eischer	6f53ecc1ae	adapt workers based on whether an operation is CPU or IO-bound Use runtime.GOMAXPROCS(0) as worker count for CPU-bound tasks, repo.Connections() for IO-bound task and a combination if a task can be both. Streaming packs is treated as IO-bound as adding more worker cannot provide a speedup. Typical IO-bound tasks are download / uploading / deleting files. Decoding / Encoding / Verifying are usually CPU-bound. Several tasks are a combination of both, e.g. for combined download and decode functions. In the latter case add both limits together. As the backends have their own concurrency limits restic still won't download more than repo.Connections() files in parallel, but the additional workers can decode already downloaded data in parallel.	2022-07-03 12:19:26 +02:00
Michael Eischer	753e56ee29	repository: Limit to a single pending pack file Use only a single not completed pack file to keep the number of open and active pack files low. The main change here is to defer hashing the pack file to the upload step. This prevents the pack assembly step to become a bottleneck as the only task is now to write data to the temporary pack file. The tests are cleaned up to no longer reimplement packer manager functions.	2022-07-02 22:42:34 +02:00
Michael Eischer	fa25d6118e	archiver: Reduce tree saver concurrency Large amount of tree savers have no obvious benefit, however they can increase the amount of (potentially large) trees kept in memory.	2022-07-02 22:42:34 +02:00
Michael Eischer	bba1e81719	archiver: Limit blob saver count to GOMAXPROCS Now with the asynchronous uploaders there's no more benefit from using more blob savers than we have CPUs. Thus use just one blob saver for each CPU we are allowed to use.	2022-07-02 22:42:34 +02:00
Michael Eischer	120ccc8754	repository: Rework blob saving to use an async pack uploader Previously, SaveAndEncrypt would assemble blobs into packs and either return immediately if the pack is not yet full or upload the pack file otherwise. The upload will block the current goroutine until it finishes. Now, the upload is done using separate goroutines. This requires changes to the error handling. As uploads are no longer tied to a SaveAndEncrypt call, failed uploads are signaled using an errgroup. To count the uploaded amount of data, the pack header overhead is no longer returned by `packer.Finalize` but rather by `packer.HeaderOverhead`. This helper method is necessary to continue returning the pack header overhead directly to the responsible call to `repository.SaveBlob`. Without the method this would not be possible, as packs are finalized asynchronously.	2022-07-02 22:42:34 +02:00
MichaelEischer	3e1de52e0a	Merge pull request #3805 from greatroar/global cmd/restic, limiter: Move config knowledge to internal packages	2022-07-02 21:56:35 +02:00
MichaelEischer	621023a50b	Merge pull request #3772 from MichaelEischer/fix-mixed-index rebuild-index: correctly rebuild index for mixed packs	2022-07-02 20:10:02 +02:00
MichaelEischer	90e9c5c4cc	Merge pull request #3729 from MichaelEischer/full-ids-in-check Include full IDs in check output	2022-07-02 20:09:39 +02:00
Michael Eischer	cdaf9b4f26	Don't crash if SecretString is uninitialized	2022-07-02 19:44:28 +02:00
Michael Eischer	5e0f1c3cef	check: remove dead code	2022-07-02 19:28:57 +02:00
Michael Eischer	0df022fa6d	check: Print full ids The short ids are not always unique. In addition, recovering from damages is easier when having the full ids as that makes it easier to access the corresponding files.	2022-07-02 19:28:57 +02:00
Michael Eischer	04c23fa95d	rebuild-index: correctly rebuild index for mixed packs For mixed packs, data and tree blobs were stored in separate index entries. This results in warning from the check command and maybe other problems.	2022-07-02 19:24:02 +02:00
MichaelEischer	bb5f196b09	Merge pull request #3733 from restic/improve-stats Improve stats	2022-07-02 19:07:31 +02:00
MichaelEischer	c16f989d4a	Merge pull request #3470 from MichaelEischer/sanitize-debug-log Sanitize debug log	2022-07-02 19:00:54 +02:00
Michael Eischer	a6e9e08034	Account for pack header overhead at each entry This will miss the pack header crypto overhead and the length field, which only amount to a few bytes per pack file.	2022-07-02 18:55:58 +02:00
Alexander Neumann	6c4ceaf1e7	Print number of bytes added to the repo This includes optional compression and crypto overhead.	2022-07-02 18:55:12 +02:00
Alexander Neumann	99634c0936	Return real size from SaveBlob	2022-07-02 18:55:12 +02:00
MichaelEischer	fdc53a9d32	Merge pull request #3787 from MichaelEischer/refactor-repository repository: (Mostly) index-related cleanups	2022-07-02 18:54:04 +02:00
Michael Eischer	6923353c43	redact swift auth token in debug output	2022-07-02 18:47:35 +02:00
Michael Eischer	5a11d14082	redacted keys/token in backend config debug log	2022-07-02 18:47:35 +02:00
Michael Eischer	0936d864a4	redact http authorization header in debug log output	2022-07-02 18:47:35 +02:00
Michael Eischer	ec7c9ce88b	drop unused repository.Loader interface	2022-07-02 18:39:59 +02:00
Michael Eischer	2cd7e90ad1	repository: cleanup	2022-07-02 18:39:59 +02:00
Michael Eischer	c1a8fa4290	repository: remove unused packIDToIndex field	2022-07-02 18:39:59 +02:00
Michael Eischer	e68c3a4e62	repository: simplify CreateIndexFromPacks	2022-07-02 18:39:59 +02:00
Michael Eischer	1974ad7ce2	repository: hide MasterIndex.FinalizeFullIndexes / FinalizeNotFinalIndexes	2022-07-02 18:39:59 +02:00
Michael Eischer	ef53ca4a5a	repository: remove MasterIndex.All()	2022-07-02 18:39:59 +02:00
Michael Eischer	bf81bf0795	repository: Properly set id for finalized index As MergeFinalIndex and index uploads can occur concurrently, it is necessary for MergeFinalIndex to check whether the IDs for an index were already set before merging it. Otherwise, we'd loose the ID of an index which is set _after_ uploading it.	2022-07-02 18:39:59 +02:00
Michael Eischer	e0a7852b8b	repository: remove unused (Master)Index.Count	2022-07-02 18:39:58 +02:00
Michael Eischer	8ef2968f28	repository: remove unused index.ListPack	2022-07-02 18:39:12 +02:00
Michael Eischer	e4f20dea61	repository: inline index.encode	2022-07-02 18:39:12 +02:00
Michael Eischer	fe5a8e137a	repository: remove unused index.Store	2022-07-02 18:39:12 +02:00
Michael Eischer	628ae799ca	repository: make flushPacks private	2022-07-02 18:39:12 +02:00
Michael Eischer	ed8aa15376	repository: add Save method to MasterIndex interface	2022-07-02 18:38:56 +02:00
Michael Eischer	a77d5c4d11	repository: index saving belongs into the MasterIndex	2022-07-02 18:38:56 +02:00

1 2 3 4 5 ...

1407 Commits