pgbackrest

mirror of https://github.com/pgbackrest/pgbackrest.git synced 2024-12-16 10:20:02 +02:00

Author	SHA1	Message	Date
David Steele	b7e17d80ea	More efficient memory allocation for Strings and String Variants. The vast majority of Strings are never modified so for most cases allocate memory for the string with the object. This results in one allocation in most cases instead of two. Use strNew() if strCat*() functions are needed. Update varNewStr() in the same way since String Variants can never be modified. This results in one allocation in all cases instead of three. Also update varNewStrZ() to use STR() instead of strNewZ() to save two more allocations.	2021-10-07 19:43:28 -04:00
Stefan Fercot	34f7873432	Report backup file validation errors in backup.info. Currently errors found during the backup are only available in text output when specifying --set. Add a flag to backup.info that is available in both the text and json output when --set is not specified. This at least provides the basic info that an error was found in the cluster during the backup, though details are still only available as described above.	2021-10-04 13:45:53 -04:00
David Steele	9346895f5b	Rename page checksum error to error list in info text output. "error list" makes it clearer that other errors may be reported. For example, if checksum-page is true in the manifest but no checksum-page-error list is provided then the error is in alignment, i.e. the file size is not a multiple of the page size, with allowances made for a valid-looking partial page at the end of the file. It is still not possible to differentiate between alignment and page checksum errors in the output but this will be addressed in a future commit.	2021-09-29 09:58:47 -04:00
David Steele	3c8819e10f	Add CodeQL static code analysis. Also fix some minor issues identified, specifically using gmtime_r()/localtime_r() vs gmtime()/localtime().	2021-07-09 14:16:10 -04:00
Eric Radman	23bdc3deb6	Fix documentation and comment typos. Identified using `ag -l \| igor`.	2021-07-01 11:50:03 -04:00
David Steele	aed3d468a1	Rename strNew() to strNewZ() and add parameter-less strNew(). Replace all instances of strNew("") with strNew() and use strNewZ() for non-empty zero-terminated strings. Besides saving a useless parameter, this will allow smarter memory allocation in a future commit by signaling intent, in general, to append or not. In the tests use STRDEF() or VARSTRDEF() where more appropriate rather than blindly replacing with strNewZ(). Also replace strLstAdd() with strLstAddZ() where appropriate for the same reason.	2021-05-21 17:36:43 -04:00
David Steele	5464ac83d1	Convert option values in commands to StringId. Convert most of the remaining options that benefit from being StringIds. Since all the command modules can include config.h directly it makes sense to auto-generate these values instead of manually creating an enum for each one. For the time being StringIds are not being auto-generated because the StringId code does not exist in Perl. However, the *_Z zero-terminated constants for each allowed option value are now auto-generated.	2021-05-11 17:24:30 -04:00
David Steele	87df6d7a58	Convert BackupType enum to StringId. Allows removal of backupType()/backupTypeStr() and improves debug logging of the enum. Move BackupType enum and string constants to info/infoBackup.h so they are available to more modules. Also convert InfoBackup to use BackupType instead of a String.	2021-05-03 12:15:39 -04:00
David Steele	85fc3da4c3	Update CipherType/CipherMode to StringId. As in `6cc521b`, this allows option values and enums to be easily mapped together.	2021-04-28 11:36:20 -04:00
David Steele	bcc925b740	Replace misused kvAdd() with kvPut(). Although kvAdd() works like kvPut() on the first call, kvPut() is more efficient when a key has a single value. Update the comment to clarify that kvAdd() is seldom required.	2021-04-22 20:04:27 -04:00
Cynthia Shang	2789d3b620	Improve info command fault tolerance. This improvement reduces the number of errors thrown; these errors will now be reported as a status for the stanza or repo as appropriate. Invalid option configurations are still thrown but all other errors are caught, formatted and reported. This was necessary for multiple repositories so that the command can complete gathering information from each repository and report the results rather than immediately aborting when an error occurs. Two new error codes were introduced: 6 = requested backup not found 99 = other, which is used to indicate an error has occurred that requires more details to be provided A new stanza name of "[invalid]" was created for instances where a stanza was not specified and no stanza can be found. If there is only one repository configured the error will move up to the stanza level with the standard error formatting of 'error (message)' where the message will be "other" and the details of the error will be listed on the next line(s): stanza: stanza1 status: error (other) [CryptoError] unable to load info file '/var/lib/pgbackrest/repo/backup/stanza1/backup.info' or '/var/lib/pgbackrest/repo/backup/stanza1/backup.info.copy': CryptoError: cipher header invalid HINT: is or was the repo encrypted? FileMissingError: unable to open missing file '/var/lib/pgbackrest/repo/backup/stanza1/backup.info.copy' for read HINT: backup.info cannot be opened and is required to perform a backup. HINT: has a stanza-create been performed? HINT: use option --stanza if encryption settings are different for the stanza than the global cipher: aes-256-cbc If a backup set is requested but is not found on any repo, a stanza-level status error of 'requested backup not found' is reported when there are no other errors: pgbackrest info --stanza=demo --set=bogus stanza: demo status: error (requested backup not found) cipher: mixed repo1: aes-256-cbc repo2: none If there are multiple repositories configured and a single repo is in error but the other repos are ok or have a different error: pgbackrest info --stanza=demo --set=20210322-171211F stanza: demo status: mixed repo1: error [CryptoError] unable to load info file '/var/lib/pgbackrest/repo/backup/stanza1/backup.info' or '/var/lib/pgbackrest/repo/backup/stanza1/backup.info.copy': CryptoError: cipher header invalid HINT: is or was the repo encrypted? FileMissingError: unable to open missing file '/var/lib/pgbackrest/repo/backup/stanza1/backup.info.copy' for read HINT: backup.info cannot be opened and is required to perform a backup. HINT: has a stanza-create been performed? HINT: use option --stanza if encryption settings are different for the stanza than the global repo2: ok cipher: mixed repo1: aes-256-cbc repo2: none db (current) wal archive min/max (12): 000000010000000000000001/000000010000000000000003 full backup: 20210322-171211F timestamp start/stop: 2021-03-22 17:12:11 / 2021-03-22 17:12:28 wal start/stop: 000000010000000000000002 / 000000010000000000000002 database size: 23.4MB, database backup size: 23.4MB repo2: backup set size: 2.8MB, backup size: 2.8MB database list: postgres (13359) Json output will include the repository information and any error information. If no stanzas are found, then [invalid] will be set as the name: [ { "archive":[], "backup":[], "cipher":"none", "db":[], "name":"[invalid]", "repo":[ { "cipher":"none", "key":1, "status":{ "code":99, "message":"[PathOpenError] unable to list file info for path '/var/lib/pgbackrest/repo2/backup': [13] Permission denied" } } ], "status":{ "code":99, "lock":{"backup":{"held":false}}, "message":"other" } } ]	2021-03-25 12:29:36 -04:00
Cynthia Shang	065b2ff230	Refactor info command repoMin/Max.	2021-02-23 16:27:05 -05:00
David Steele	7d6c0319f0	Add lstEmpty(), strLstEmpty(), and varLstEmpty(). This seems more readable than lstSize() == 0. Hopefully this will also eliminate usage of lstSize() > 0/lst*Size() != 0 variants for the inverse.	2021-01-29 14:27:56 -05:00
David Steele	456a300bb7	Remove too-verbose braces in switch statements. The original intention was to enclose complex code in braces but somehow braces got propagated almost everywhere. Document the standard for braces in switch statements and update the code to reflect the standard.	2021-01-26 12:10:24 -05:00
Cynthia Shang	00fac1c0d1	Improve info command text output and --set handling. The info command provides total sizes for files in the backup on the database as well as the repository. The text output and associated user documentation has been updated to provide more clarity regarding the sizes being displayed. In addition, the info command is updated to allow a user to optionally specify the repository when requesting a specific backup set. In this case, the text output will reflect the status of the stanza, the cipher types and archive min/max over all the repositories instead of a single repository when the repo option is specified.	2021-01-25 09:19:05 -05:00
Cynthia Shang	f32eb9b94e	Partial multi-repository implementation. Multi-repository implementations for the archive-push, check, info, stanza-create, stanza-upgrade, and stanza-delete commands. Multi-repo configuration is disabled so there should be no behavioral changes between these commands and their current single-repo implementations. Multi-repo documentation and integration tests are still in the multi-repo development branch. All unit tests work as multi-repo since they are able to bypass the configuration restrictions.	2021-01-21 15:21:50 -05:00
David Steele	298cc4d5e5	Remove non-conforming periods and reformat some comments.	2021-01-14 10:39:25 -05:00
Cynthia Shang	cc90163233	Add empty archive array to info command JSON when stanza is missing. There is an inconsistency when the JSON is output for the case when a stanza is requested and it does not exist in the repo. This was the only case where the archive array was not added to the JSON. Adding it will simplify the upcoming multi-repo support code. Also, a redundant test was removed rather than updating it for this case.	2020-12-30 16:17:56 -05:00
Stefan Fercot	5488de8b6a	Report page checksum errors in info command text output. This feature currently only works for text output. JSON output is planned for the future.	2020-11-25 12:14:03 -05:00
David Steele	7fda83b31e	Allow multiple remote locks from the same main process. Improve locking on remote processes by introducing an exec-id that is unique to the main process and passed to all remote processes. This allows the remote processes to determine if a lock is held by a remote from the same main process. If so, the lock is allowed. The exec-id is also useful for associating remote logs with main logs for debugging purposes.	2020-11-23 12:41:54 -05:00
Stefan Fercot	abe9d90c89	Improve info command output when a stanza is specified but missing. Return a path missing error when a stanza is specified for the info command but the stanza does not exist in the repository. Previously [] was returned, which is still the case if no stanza is specified and the repository does not exist.	2020-10-27 08:34:18 -04:00
David Steele	cde2c756ea	Rename handle to fd. Pretty much everywhere handle is used what is really meant is file descriptor (fd). This terminology got migrated over from Perl and is just not quite correct, or at least not as correct as fd. There were also plenty of places fd was used so now all uses are consistent. The Perl code was not updated but might be in a future commit.	2020-08-05 18:25:07 -04:00
David Steele	3e9dce0d76	Rename strPtr()/strPtrNull() to strZ()/strZNull(). We use the Z suffix in many functions to indicate that we are expecting a zero-terminated string so make this function conform to the pattern. As a bonus the new name is a bit shorter, which is a good quality in a commonly-used function.	2020-07-30 07:49:06 -04:00
David Steele	45d9b03136	Add strCatZ(). strCat() did not follow our convention of appending Z to functions that accept zero-terminated strings rather than String objects. Add strCatZ() to accept zero-terminated strings and update strCat() to accept String objects. Use LF_STR where appropriate but don't use other String constants because they do not improve readability.	2020-06-24 12:09:24 -04:00
David Steele	c4fe09dabe	Fix incorrect param log types.	2020-06-16 19:25:16 -04:00
David Steele	ce55866714	Enforce non-null for most string options. There have been a number of segfaults reported because a string option expected to be non-null was actually null. This is generally due to options that are expected to be set but are in fact optional. Protect against this by creating cfgOptionStrNull() to get options that can be null, while changing cfgOptionStr() to always expect non-null. There are relatively few places where nulls are expected. There is definitely a chance for breakage here as null options might currently be working in the field but will be caught by this new check. Hopefully introducing the check early in the release cycle will allow us to catch any issues.	2020-04-30 10:34:44 -04:00
Stefan Fercot	e92eb709d6	Add backup/expire running status to the info command. This is implemented by checking for a backup lock on the host where info is running so there are a few limitations: * It is not currently possible to know which command is running: backup, expire, or stanza-. The stanza commands are very unlikely to be running so it's pretty safe to guess backup/expire. Command information may be added to the lock file to improve the accuracy of the reported command. If the info command is run on a host that is not participating in the backup, e.g. a standby, then there will be no backup lock. This seems like a minor limitation since running info on the repo or primary host is preferred.	2020-04-24 08:00:00 -04:00
David Steele	1aca2cc902	Move extern function comments to headers. This has been the policy for some time but due to migration pressure only new functions and refactors have been following this rule. Now it seems sensible to make a clean sweep and move all the comments that have not been moved already (i.e. most of them). Only obvious typos and gross inaccuracies in the comments have been fixed. For this most part this was a copy and paste operation. Useless comments, e.g. "New object", were not copied. Even so, there are surely many deficient comments left. Some rearranging was done where needed and functions were placed in the proper sections, e.g. "Constructors", "Functions", etc. A few function prototypes were found that not longer had an implementation. These were removed, but there may be more. The coding document has been updated to reflect this policy, which is not new but has never been documented.	2020-04-03 18:01:28 -04:00
David Steele	ec173f12fb	Add MEM_CONTEXT_PRIOR() block and update current call sites. This macro block encapsulates the common pattern of switching to the prior (formerly called old) mem context to return results from a function. Also rename MEM_CONTEXT_OLD() to memContextPrior(). This violates our convention of macros being in all caps but memContextPrior() will become a function very soon so this will reduce churn.	2020-01-17 13:29:49 -07:00
David Steele	f0ef73db70	pgBackRest is now pure C. Remove embedded Perl from the distributed binary. This includes code, configure, Makefile, and packages. The distributed binary is now pure C. Remove storagePathEnforceSet() from the C Storage object which allowed Perl to write outside of the storage base directory. Update mock/all and real/all integration tests to use storageLocal() where they were violating this rule. Remove "c" option that allowed the remote to tell if it was being called from C or Perl. Code to convert options to JSON for passing to Perl (perl/config.c) has been moved to LibC since it is still required for Perl integration tests. Update build and installation instructions in the user guide. Remove all Perl unit tests. Remove obsolete Perl code. In particular this included all the Perl protocol code which required modifications to the Perl storage, manifest, and db objects that are still required for integration testing but only run locally. Any remaining Perl code is required for testing, documentation, or code generation. Rename perlReq to binReq in define.yaml to indicate that the binary is required for a test. This had been the actual meaning for quite some time but the key was never renamed.	2019-12-13 17:55:41 -05:00
David Steele	c5a6631d27	Rearrange manifest module. Put functions with related functions, move getters above the helper functions, and rename manifestPgPath() to manifestPathPg().	2019-11-21 11:44:40 -05:00
David Steele	1db9e3b144	Remove *MP() macros variants. Adding a dummy column which is always set by the P() macro allows a single macro to be used for parameters or no parameters without violating C's prohibition on the {} initializer. -Wmissing-field-initializers remains disabled because it still gives wildly different results between versions of gcc.	2019-11-17 15:10:40 -05:00
Cynthia Shang	c5b76d213b	Modify InfoBackupData struct to use time_t for backup start/stop times. The uint64_t types worked but were not consistent with how timestamps are handled in other parts of the code.	2019-11-12 17:05:09 -05:00
Cynthia Shang	db1dc4f275	Remove pretty-printing from jsonFromKv() and jsonFromVar(). Now that pretty-printing has been removed from the info command it no longer has a purpose, so remove it.	2019-10-11 13:03:52 -04:00
Cynthia Shang	d90b2724f8	JSON output from the info command is no longer pretty-printed. Monitoring systems can more easily ingest the JSON without linefeeds. External tools such as jq can be used to pretty-print if desired.	2019-10-11 12:56:03 -04:00
David Steele	309ae66e2f	Remove unneeded static declarations and use sizeof() where appropriate.	2019-10-01 08:47:56 -04:00
Cynthia Shang	f96c54c4ba	Add info command set option for detailed text output. The additional details include databases that can be used for selective restore and a list of tablespaces and symlinks with their default destinations. This information is not included in the JSON output because it requires reading the manifest which is too IO intensive to do for all manifests. We plan to include this information for JSON in a future release.	2019-09-30 12:39:38 -04:00
Cynthia Shang	56bf9d0566	Update HINT messages to conform to new standard detailed in CODING.md.	2019-09-14 12:21:08 -04:00
David Steele	4d84820021	Improve performance of info file load/save. Info files required three copies in memory to be loaded (the original string, an ini representation, and the final info object). Not only was this memory inefficient but the Ini object does sequential scans when searching for keys making large files very slow to load. This has not been an issue since archive.info and backup.info are very small, but it becomes a big deal when loading manifests with hundreds of thousands of files. Instead of holding copies of the data in memory, use a callback to deliver the ini data directly to the object when loading. Use a similar method for save to avoid having an intermediate copy. Save is a bit complex because sections/keys must be written in alpha order or older versions of pgBackRest will not calculate the correct checksum. Also move the load retry logic to helper functions rather than embedding it in the Info object. This allows for more flexibility in loading and ensures that stack traces will be available when developing unit tests. Reviewed by Cynthia Shang.	2019-09-06 13:48:28 -04:00
Josh Soref	8074ca6a26	Fix typos in variable names. Contributed by Josh Soref.	2019-08-26 12:30:22 -04:00
Josh Soref	c2771e5469	Fix comment typos. This includes some variable names in tests which don't seem important enough for their own commits. Contributed by Josh Soref.	2019-08-26 12:05:36 -04:00
Cynthia Shang	44bafc127d	Rename infoNew() functions to infoNewLoad(). These names more accurately reflect what the functions do and follow the convention started in Info and InfoPg. Also remove the ignoreMissing parameter since it was never used. Contributed by Cynthia Shang.	2019-06-17 06:47:15 -04:00
David Steele	96770c529b	storageList() returns an empty list by default for missing paths. The prior behavior was to return NULL so the caller would know the path was missing, but this is rarely useful, complicates the calling code, and increases the chance of segfaults. The .nullOnMissing param has been added to enable the prior behavior.	2019-05-24 13:12:56 -04:00
David Steele	027c263871	Add configure script for improved multi-platform support. Use autoconf to provide a basic configure script. WITH_BACKTRACE is yet to be migrated to configure and the unit tests still use a custom Makefile. Each C file must include "build.auto.conf" before all other includes and defines. This is enforced by test.pl for includes, but it won't detect incorrect define ordering. Update packages to call configure and use standard flags to pass options.	2019-04-26 08:08:23 -04:00
David Steele	81f652137c	Add separate functions to encode/decode each JSON type. In most cases the JSON type is known so this is more efficient than converting to Variant first, both in terms of memory and time. Also rename some of the existing functions for consistency.	2019-04-22 18:41:01 -04:00
David Steele	47491e3c47	varNewKv() accepts a KeyValue object rather than creating one. This allows for more flexibility about when the Variant is created.	2019-04-22 16:04:04 -04:00
David Steele	0c866f52c6	Update code to use new unsigned int Variant type and config methods.	2019-04-19 11:40:39 -04:00
David Steele	4c13955c05	Add macros to create constant Variant types. These work almost exactly like the String constant macros. However, a struct per variant type was required which meant custom constructors and destructors for each type. Propagate the variant constants out into the codebase wherever they are useful.	2019-04-17 08:04:22 -04:00
David Steele	2ef5ad70a2	Move crypto module to common/crypto. It makes sense for the crypto code to be in common since it is not pgBackRest-specific. Also combine the crypto tests into a single module.	2019-03-10 13:27:30 +02:00
Stefan Fercot	80df1114bd	Fix info command missing WAL min/max when stanza specified. This issue was a result of STORAGE_REPO_PATH prepending an extra stanza when the stanza was specified on the command line. The tests missed this because by some strange coincidence the WAL dirs were empty for each test that specified a stanza. Add new tests to prevent a regression. Fixed by Stefan Fercot.	2019-02-21 12:09:12 +02:00

1 2

60 Commits