pgbackrest

mirror of https://github.com/pgbackrest/pgbackrest.git synced 2024-12-12 10:04:14 +02:00

Author	SHA1	Message	Date
David Steele	a79034ae2f	Add read range to all storage drivers. The range feature allows reading out an arbitrary chunk of a file and will be important for efficient small file support. Now that all drivers are required to support ranges remove the storageFeatureLimitRead feature flag that was implemented only by the Posix driver.	2022-01-11 14:42:53 -05:00
David Steele	2cddbbdee0	Remove obsolete cfgOptionHostPort()/cfgOptionIdxHostPort(). These functions were made obsolete by the refactor in `6a124584`.	2022-01-10 17:20:48 -05:00
David Steele	aeecb500f5	Improve implementation of cfgOptionIdxName(). Cache option names after they are generated rather than regenerating them each time.	2022-01-10 14:47:29 -05:00
David Steele	aced5d47ed	Replace cfgOptionGroupIdxToKey() with cfgOptionGroupName(). Do the replacement anywhere cfgOptionGroupIdxToKey() is being used to construct a group name in a message. cfgOptionGroupName() is better for this case since it also includes the name of the group so that it does not need to be repeated in each message.	2022-01-10 09:10:06 -05:00
David Steele	e4b48eb430	Fix inconsistent group display names in messages. In other instances there are no dashes, e.g. repo1 or pg1. Make these messages match.	2022-01-09 19:43:44 -05:00
David Steele	5f78a5fc18	Add ioCopy(). Functionality to copy from IoRead to IoWrite is frequently used so centralize it. This also simplifies coverage testing in places where a loop was required before.	2022-01-09 13:19:43 -05:00
David Steele	47954774c6	Combine encrypted backupFile() tests with unencrypted tests. This makes it easier to comment out all the tests while developing without getting unused variable errors.	2022-01-09 10:11:00 -05:00
Stefan Fercot	d866dd5c29	Add backup LSNs to info command output. The backup LSNs are useful for performing LSN-based PITR. LSNs will not be displayed in the general text output (without --set) because they are probably not useful enough to deserve their own line.	2022-01-07 14:09:58 -05:00
David Steele	bb4b30ddd3	Remove support for PostgreSQL 8.3/8.4. There is no evidence that users need 8.3/8.4 anymore but it does cost us in terms of development and testing, especially now that we have a number of new backup/restore features planned. It seems to make sense to remove this support now. If there are users who need to use/migrate from these versions they can use an older version of pgBackRest.	2022-01-06 15:34:04 -05:00
Reid Thompson	fdbeb8e7d6	Fix typo in error message.	2022-01-06 14:22:56 -05:00
Reid Thompson	a82f0179cd	Note that replications slots are not restored. Update documentation and help to note that replication slots are not restored and reference the PostgreSQL documentation to explain why.	2022-01-04 16:11:27 -05:00
David Steele	f18f2d9991	v2.37: TLS Server Bug Fixes: * Fix restore delta link mapping when path/file already exists. (Reviewed by Reid Thompson. Reported by Younes Alhroub.) * Fix socket leak on connection retries. (Reviewed by Reid Thompson. Reported by James Coleman.) Features: * Add TLS server. (Reviewed by Stephen Frost, Reid Thompson, Andrew L'Ecuyer.) * Add --cmd option. (Contributed by Reid Thompson. Reviewed by Stefan Fercot, David Steele. Suggested by Virgile CREVON.) Improvements: * Check archive immediately after backup start. (Reviewed by Reid Thompson, David Christensen.) * Add timeline and checkpoint checks to backup. (Reviewed by Stefan Fercot, Reid Thompson.) * Check that clusters are alive and correctly configured during a backup. (Reviewed by Stefan Fercot.) * Error when restore is unable to find a backup to match the time target. (Reviewed by Reid Thompson, Douglas J Hunley. Suggested by Douglas J Hunley.) * Parse protocol/port in S3/Azure endpoints. (Contributed by Reid Thompson. Reviewed by David Steele.) * Add warning when checkpoint_timeout exceeds db-timeout. (Contributed by Stefan Fercot. Reviewed by David Steele.) * Add verb to HTTP error output. (Contributed by Christoph Berg. Reviewed by David Steele.) * Allow y/n arguments for boolean command-line options. (Contributed by Reid Thompson. Reviewed by David Steele.) * Make backup size logging exactly match info command output. (Contributed by Reid Thompson. Reviewed by David Steele. Suggested by Mahomed Hussein.) Documentation Improvements: * Display size option default and allowed values with appropriate units. (Reviewed by Reid Thompson.) * Fix typos and improve documentation for the tablespace-map-all option. (Reviewed by Reid Thompson. Suggested by Reid Thompson.) * Remove obsolete statement about future multi-repository support. (Suggested by David Christensen.)	2022-01-03 08:43:55 -05:00
David Steele	d6ebf6e2d6	Remove dead test code.	2021-12-30 18:54:36 -05:00
Reid Thompson	6a12458440	Parse protocol/port in S3/Azure endpoints. Utilize httpUrlNewParseP() to parse endpoint and port from the URL in the S3 and Azure helpers to avoid issues where protocol was not expected to be part of the URL.	2021-12-16 10:30:59 -05:00
David Steele	f06101de77	Add TLS server documentation. Add documentation and make the feature visible.	2021-12-16 09:47:04 -05:00
David Steele	615bdff403	Fix socket leak on connection retries. This leak was caused by the file descriptor variable getting clobbered after a long jump. Mark it as volatile to fix. Testing this is a bit complex because the issue only happens in optimized builds, if at all. Put the test into the performance suite, which is always optimized, until a better idea presents itself.	2021-12-14 14:53:41 -05:00
David Steele	a73fe4eb96	Fix restore delta link mapping when path/file already exists. If a path/file was remapped to a link using either --link-map or --link-all there would be no affect if the path/file already existed. If a link existed it would be properly updated and converting a link to a path/file also worked. The issue happened during delta cleanup, which failed to check if the existing path/file had been remapped to a link. Add checks for newly mapped path/file links and remove the old path/file we required.	2021-12-10 15:53:40 -05:00
David Steele	19a7ec69de	Close expect log file when unit test completes. This did not cause any issues, but it is better to explicitly close open files.	2021-12-10 15:04:55 -05:00
Christoph Berg	c38e2d3170	Add verb to HTTP error output. This makes it easier to debug HTTP errors.	2021-12-08 15:00:19 -05:00
David Steele	be4ac3923c	Error when restore is unable to find a backup to match the time target. This was previously a warning but the warning is easy to miss so a lot of time may be lost restoring and recovering a backup that will not hit the target. Since this is technically a breaking change, add an "important note" about the change to the release.	2021-12-08 13:57:26 -05:00
Stefan Fercot	6723305937	Add warning when checkpoint_timeout exceeds db-timeout. In the backup command, add a warning if start-fast is disabled and the PostgreSQL checkpoint_timeout is greater than db-timeout. In such cases, we might timeout before the checkpoint occurs and the backup really starts.	2021-12-08 12:29:20 -05:00
David Steele	bd2ba802db	Check that clusters are alive and correctly configured during a backup. Fail the backup if a cluster stops or the standby is promoted. Previously, shutting down the primary would cause an error but it was not detected until the end of the backup. Now the error will happen sooner and a promotion on the standby will also cause an error.	2021-12-08 10:16:41 -05:00
David Steele	7b3ea883c7	Add SIGTERM and SIGHUP handling to TLS server. SIGHUP allows the configuration to be reloaded. Note that the configuration will not be updated in child processes that have already started. SIGTERM terminates the server process gracefully and sends SIGTERM to all child processes. This also gives the tests an easy way to stop the server.	2021-12-07 18:18:43 -05:00
David Steele	49145d72ba	Add timeline and checkpoint checks to backup. Add the following checks: * Checkpoint is updated in pg_control after pg_start_backup(). This helps ensure that PostgreSQL and pgBackRest have a consistent view of the storage and that PGDATA paths match. * Timeline of backup start WAL file matches pg_control. Hard to see how this one could get hit, but we have the power... * Standby is on the same timeline as the primary. If not, this standby is not following the primary. * Last standby checkpoint is not greater than the backup checkpoint. If so, this standby is not following the primary. This also requires some additional plumbing to read/write timeline/checkpoint from pg_control and parse timelines from WAL filenames. There were some changes in the backup tests caused by the fact that pg_control now has different contents for each backup. The check to ensure that the required checkpoint was reached on the standby should also be updated to use pg_control (it currently uses pg_control_checkpoint()), but that requires non-trivial changes to the test harness and will need to wait.	2021-12-07 09:21:07 -05:00
David Steele	9c76056dd0	Add error type and message to CHECK() macro. A CHECK() worked exactly like ASSERT() except that it was compiled into production code. However, over time many checks have been added that should not throw AssertError, which should be reserved for probable coding errors. Allow the error code to be specified so other error types can be thrown. Also add a human-readable message since many of these could be seen by users even when there is no coding error. Update coverage exceptions for CHECK() to match ASSERT() since all conditions will never be covered.	2021-11-30 16:21:15 -05:00
David Steele	0895cfcdf7	Add HRN_PG_CONTROL_PUT() and HRN_PG_CONTROL_TIME(). These macros simplify management of pg_control test files. Centralize time updates for pg_control in the command/backup module. This caused some time updates in the logs. Finally, move the postgres module after the storage module so it can use storage macros.	2021-11-30 13:23:11 -05:00
David Steele	01ac6b6cac	Autogenerate test system identifiers. hrnPgControlToBuffer() and hrnPgWalToBuffer() now generate the system id based on the version of Postgres. If a value less than 100 is specified for systemId then it will be added to the default system id so there can be multiple ids for a single version of PostgreSQL. Add constants to represent version system ids in tests. These will eventually be auto-generated. This changes some checksums and we no longer have big-endian tests systems, so X those checksums out so it is obvious they are no longer valid.	2021-11-30 08:28:36 -05:00
David Steele	3f7409019d	Ensure ASSERT() macro is always available in test modules. Tests that run without DEBUG for performance did not have ASSERT() and were using CHECK() instead. Instead ensure that the ASSERT() macro is always available in tests.	2021-11-24 16:09:45 -05:00
David Steele	7e35245dc3	Use ASSERT() or TEST_RESULT*() instead of CHECK() in test modules.	2021-11-23 08:07:31 -05:00
Reid Thompson	a3d7a23a9d	Use infoBackupDataByLabel() to log backup size. Eliminate summing and passing of copied files sizes for logging backup size. Instead, utilize infoBackupDataByLabel() to pull the backup size for the log message.	2021-11-22 12:52:37 -05:00
Reid Thompson	1a0560d363	Allow y/n arguments for boolean command-line options. This allows boolean boolean command-line options to work like their config file equivalents. At least for now this behavior will remain undocumented since all examples in the documentation will continue to use the standard syntax. The idea is that it will "just work" when options are copied out of config files rather than generating an error.	2021-11-19 12:22:09 -05:00
David Steele	2d963ce947	Rename server-start command to server.	2021-11-18 17:23:11 -05:00
David Steele	1f14f45dfb	Check archive immediately after backup start. Previously the archive was only checked at the end of the backup to ensure all WAL required to make the backup consistent was present. The problem was that if archiving was not functioning then the backup had to complete before the user found out, which could be a while if the database was large enough. Add an archive check immediately after backup start so failures are reported earlier. The trick is to determine which WAL to check. If the repo is new there may not be any WAL in it and pg_start_backup() will not switch the WAL segment if it is empty. These are both likely scenarios when setting up and/or testing pgBackRest. If the WAL segment is switched by pg_start_backup(), then check the archive for the segment that was detected prior to backup start. This should be common on normal running clusters with regular activity. Note that this might not be the segment immediately prior to the backup start segment if WAL volume is high. If pg_start_backup() did not switch the WAL then we can force a switch on PostgreSQL >= 9.3 by creating a restore point. In that case the WAL to check will be the backup start WAL. This is most likely to happen on idle systems, during testing, or immediately after a repo switch. An advantage of this approach other than earlier notification is that the backup directory will not be created so no resume will be attempted on the next backup. Note that some additional churn was created in backup.c because the load of archive.info needs to be done earlier.	2021-11-18 16:18:10 -05:00
David Steele	809f0bbc63	Add infoBackupLabelExists(). This is easier to read than using infoBackupDataByLabel() != NULL. It also allows an assertion to be added to infoBackupDataByLabel() to ensure that a NULL return value is not used unsafely.	2021-11-16 11:34:53 -05:00
David Steele	b3a5f7a8e2	Add tablespace_map file to command/backup test module. The code worked fine but better to have explicit tests for this file.	2021-11-15 14:32:22 -05:00
David Steele	43cfa9cef7	Revive archive performance test. This test was lost due to a syntax issue in `a58635ac`. Update the test to use system() to better mimic what postgres does and add logging so pgBackRest timing can be determined.	2021-11-10 12:14:41 -05:00
Reid Thompson	6e635764a6	Match backup log size with size reported by info command. Properly log the size of files copied during the backup, matching the backup size returned from the info command. In the reference issue, the incremental backup after switchover logs the size of all files evaluated rather than only the size of the files copied in the backup.	2021-11-09 13:24:56 -05:00
David Steele	d05d6b8714	Do not delete manifests individually during stanza delete. This appears to have been an attempt to not delete files that we don't recognize, but it only works in narrow cases and could leave the user is a position of not being able to complete the stanza delete without manual intervention. It seems better just to proceed with the delete, especially since the info files have already been removed. In addition, deleting the manifests individually could be slow on object stores if there were a very large number of backups.	2021-11-08 09:39:58 -05:00
David Steele	676b9d95dd	Optional parameters for tlsClientNew(). There are a number of optional parameters with the same type so this makes them easier to track and reduces churn when new ones are added.	2021-11-04 08:19:18 -04:00
David Steele	038abaa71d	Display size option default and allowed values with appropriate units. Size option default and allowed values were displayed in bytes, which was confusing for the user. This also lays the groundwork for adding units to time options. Move option parsing functions into a common module so they can be used from the build module.	2021-11-03 15:23:08 -04:00
Reid Thompson	2a576477b3	Add --cmd option. Allows users to provide an executable to be used when pgbackrest generates command strings that expect to invoke pgbackrest. These generated commands are written to files by pgbackrest, e.g. recovery.conf.	2021-11-03 11:36:34 -04:00
David Steele	c5b5b58806	Simplify error handler. The error handler used a loop to process try, catch, and finally blocks. This worked fine but static analysis tools like Coverity did not understand that the finally block would always run and so there were false positives about double-free, unfreed resource, etc. This implementation removes the loop, which simplifies everything, and makes it clear that the finally block will always run. This cuts down on Coverity false positives. This implementation also catches lack of coverage on empty catch blocks so a few test fixes were committed separately in `d74fe7a`. A small refactor in backup.c is required because gcc 10.3.1 on Fedora 33 complains that the reason variable may be used uninitialized. It's not clear why this is the case, but reducing the scope of the TRY block fixes the issue.	2021-11-03 10:36:31 -04:00
David Steele	7f6c513be9	Add StringId as an option type. Rather the converting String to StringIds at runtime, store defaults in StringId format in parse.auto.c and convert user input to StringId during parsing.	2021-11-03 07:27:26 -04:00
David Steele	b13844086d	Use cfgOptionStrId() instead of cfgOptionStr() where appropriate. The compress-type, repo-type and log-level-* options have allow lists, which means it is more efficient to treat them as StringIds. For compress-type and log-level-* also update the functions that convert them to enums.	2021-11-01 17:35:19 -04:00
David Steele	bc352fa6a8	Simplify strIdFrom() functions. The strIdFrom() forced the caller to pick an encoding, which led to a number of TRY...CATCH blocks in the code. In practice the caller does not care which encoding is used as long as the string is valid for some encoding. Update the strIdFrom*() function to try all possible encodings and only throw an error when the string is not valid for any of them.	2021-11-01 10:08:56 -04:00
David Steele	42fd6ce4e0	v2.36: Minor Bug Fixes and Improvements Bug Fixes: * Allow "global" as a stanza prefix. (Reviewed by Stefan Fercot. Reported by Younes Alhroub.) * Fix segfault on invalid GCS key file. (Reviewed by Stephen Frost. Reported by Henrik Feldt.) Improvements: * Allow link-map option to create new links. (Reviewed by Don Seiler, Stefan Fercot, Chris Bandy. Suggested by Don Seiler.) * Increase max index allowed for pg/repo options to 256. (Reviewed by Cynthia Shang.) * Add WebIdentity authentication for AWS S3. (Reviewed by James Callahan, Reid Thompson, Benjamin Blattberg, Andrew L'Ecuyer.) * Report backup file validation errors in backup.info. (Contributed by Stefan Fercot. Reviewed by David Steele.) * Add recovery start time to online backup restore log. (Reviewed by Tom Swartz, Stefan Fercot. Suggested by Tom Swartz.) * Report original error and retries on local job failure. (Reviewed by Stefan Fercot.) * Rename page checksum error to error list in info text output. (Reviewed by Stefan Fercot.) * Add hints to standby replay timeout message. (Reviewed by Cynthia Shang, Stefan Fercot. Suggested by Leigh Downs.)	2021-11-01 08:59:14 -04:00
David Steele	c32e000ab9	Use Rocky Linux for documentation builds instead of CentOS. Since CentOS 8 will be EOL at the end of the year it makes sense to do this now. The centos:8 image is still used in documentation.xml because changes there require manual testing, which will need to be done at a later date. The changes are not user-facing, however, and can be done at any time. Also update CentOS references to RHEL since that is what we are emulating for testing purposes.	2021-10-28 15:15:49 -04:00
David Steele	adc09ffc3b	Minor fix for lower-casing of option summaries. This works with existing cases and fixes "I/O".	2021-10-28 08:10:43 -04:00
David Steele	fa564ee196	Improve documentation for cmd-ssh, repo-host-cmd, pg-host-cmd options. Use "command" instead of "exe" and make the descriptions more consistent.	2021-10-27 11:08:32 -04:00
David Steele	e1f6c066b3	Improve documentation for buffer-size option.	2021-10-27 10:52:39 -04:00
David Steele	1f7c7b7dda	Fix test descriptions in common/typeVariantTest.	2021-10-26 16:56:44 -04:00
David Steele	d74fe7a222	Add coverage for empty CATCH() blocks. Currently empty CATCH() blocks are always marked as covered because of the loop structure of error handling. A prototype implementation of error handling without looping has shown that these CATCH() blocks are not covered without new tests. Whether or not that prototype gets committed it is worth adding the tests.	2021-10-26 13:53:44 -04:00
David Steele	7fb99c59c8	Use externed instead of extern'd in comments. This is mostly to revert some comment changes in `b11ab9f7` that will break the ppc64le patch, but at the same time keep the spelling consistent in all comments and documentation. Also revert some space changes for the same reason.	2021-10-26 07:46:48 -04:00
David Steele	653ffcf8d9	Adjustments for new breaking change in Azurite. Azurite released another breaking change (see `fbd018cd`, `096829b3`, `c38d6926`, and Azurite issue 1039) so make adjustments as needed to documentation and tests. Also remove some dead code that hid the repo-storage-host option and was made obsolete by all these changes.	2021-10-25 15:42:28 -04:00
David Steele	3879bc69b8	Add WebIdentity authentication for AWS S3. This allows credentials to be automatically acquired in an EKS environment.	2021-10-22 18:31:55 -04:00
David Steele	51785739f4	Store config values as a union instead of a variant. The variants were needed to easily serialize configurations for the Perl code. Unions are more efficient and will allow us to add new types that are not supported by variants, e.g. StringId.	2021-10-22 18:02:20 -04:00
David Steele	b11ab9f799	Fix typos.	2021-10-21 13:31:22 -04:00
David Steele	5dfdd6dd5b	Add -Werror -Wfatal-errors -g flags to configure --enable-test. These flags are used for all tests but it was not possible to add them to configure before the change in `046d6643`. This is especially important for adhoc tests to ensure the flags are not forgotten. Remove the flags from test make commands where they were being applied. There is no change for production builds.	2021-10-19 12:45:20 -04:00
David Steele	ccc255d3e0	Add TLS Server. The TLS server is an alternative to using SSH for protocol connections to remote hosts. This command is currently experimental and intended only for trial and testing. As such, the new commands and options will not show up in the command-line help unless directly requested.	2021-10-18 14:32:41 -04:00
David Steele	90f7f11a9f	Add missing static keywords in test modules.	2021-10-18 12:22:48 -04:00
David Steele	4570c7e275	Allow error buffer to be resized for testing. Some tests can generate very large error messages for diffs and they often get cut off before the end. Also fix a test so it does not create too large a buffer on the stack.	2021-10-18 11:32:53 -04:00
David Steele	838ee3bd08	Increase some storage test timeouts. 32-bit Debian 9 is sometimes timing out on these tests so increase the timeouts to make the tests more reliable.	2021-10-18 11:05:53 -04:00
David Steele	6b9e19d423	Convert configuration optional rules to pack format. The previous format was custom for configuration parsing and was not as expressive as the pack format. An immediate benefit is that commands with the same optional rules are merged. Defaults are now represented correctly (not multiplied), which simplifies the option default functions used by help.	2021-10-16 12:35:47 -04:00
David Steele	360cff94e4	Update 32-bit test container to Debian 9. Also rebalance PostgreSQL version integration tests.	2021-10-16 12:33:31 -04:00
David Steele	144469b977	Add const buffer functions to Pack type. These allow packs to be created without allocating a buffer in the case that the buffer already exists or the data is in a global constant. Also fix a rendering issue in hrnPackReadToStr().	2021-10-15 15:50:55 -04:00
David Steele	66bfd1327e	Rename SSH connection control parameters in integration tests.	2021-10-13 19:48:41 -04:00
David Steele	447b24309d	Update RHEL package URL.	2021-10-13 19:43:40 -04:00
David Steele	01b20724da	Rename PostgreSQL pid file constants and tests.	2021-10-13 19:36:59 -04:00
David Steele	5701620408	Rename manifest file primary flag in tests.	2021-10-13 19:02:58 -04:00
David Steele	a44f9e373b	Update Vagrantfile to Ubuntu 20.04.	2021-10-13 13:21:04 -04:00
David Steele	5e84645ac0	Update comments referring to the PostgreSQL primary.	2021-10-13 12:16:47 -04:00
David Steele	90c73183ea	Add libc6-dbg required by updated valgrind to Vagrantfile/Dockerfile.	2021-10-13 09:37:03 -04:00
David Steele	c2d4552b73	Add debug options to code generation make in test.pl.	2021-10-13 08:51:58 -04:00
David Steele	610bfd736e	Increase tolerance for 0ms sleep in common/time test.	2021-10-09 12:34:45 -04:00
David Steele	ed68792e76	Rename strNewN() to strNewZN(). Make the function name consistent with other functions that accept zero-terminated strings, e.g. strNewZ() and strCatZN().	2021-10-07 19:57:28 -04:00
David Steele	b7e17d80ea	More efficient memory allocation for Strings and String Variants. The vast majority of Strings are never modified so for most cases allocate memory for the string with the object. This results in one allocation in most cases instead of two. Use strNew() if strCat*() functions are needed. Update varNewStr() in the same way since String Variants can never be modified. This results in one allocation in all cases instead of three. Also update varNewStrZ() to use STR() instead of strNewZ() to save two more allocations.	2021-10-07 19:43:28 -04:00
David Steele	208641ac7f	Use constant string for user/group in performance/type test. It is not safe to return strings created with STRDEF() from a function.	2021-10-07 18:50:56 -04:00
David Steele	498902e885	Allow "global" as a stanza prefix. A stanza name like global_stanza was not allowed because the code was not selective enough about how a global section should be formatted. Update the config parser to correctly recognize global sections.	2021-10-07 12:18:24 -04:00
David Steele	fb3f6928c9	Add configurable storage helpers to create repository storage. Remove the hardcoded storage helpers from storageRepoGet() except for the the built-in Posix helper and the special remote helper. The goal is to make storage driver development a bit easier by isolating as much of the code as possible into the driver module. This also makes coverage reporting much simpler for additional drivers since they do not need to provide coverage for storage/helper. Consolidate the CIFS tests into the Posix tests since CIFS is just a special case of the Posix. Test all storage features in the Posix test so that other storage driver tests do not need to provide coverage for storage/storage. Remove some dead code in the storage/s3 test.	2021-10-06 19:27:04 -04:00
David Steele	68c5f3eaf1	Allow link-map option to create new links. Currently link-map only allows links that exist in the backup manifest to be remapped to a new destination. Allow link-map to create a new link as long as a valid path/file from the backup is referenced.	2021-10-05 17:59:05 -04:00
David Steele	f2aeb30fc7	Add state to ProtocolClient. This is currently only useful for debugging, but in the future the state may be used for resetting the protocol when something goes wrong.	2021-10-05 14:06:59 -04:00
David Steele	6af827cbb1	Report original error and retries on local job failure. The local process will retry jobs (e.g. backup file) but after a certain number of failures gives up. Previously, the last error was reported but generally the first error is far more valuable. The last error is likely to be a cascade failure such as the protocol being out of sync. Report the first error (and stack trace) and append the retry errors to the first error without stack trace information.	2021-10-05 09:00:16 -04:00
Stefan Fercot	34f7873432	Report backup file validation errors in backup.info. Currently errors found during the backup are only available in text output when specifying --set. Add a flag to backup.info that is available in both the text and json output when --set is not specified. This at least provides the basic info that an error was found in the cluster during the backup, though details are still only available as described above.	2021-10-04 13:45:53 -04:00
David Steele	57c6231546	Add arm64 testing on Cirrus CI. These tests run in a container without permissions to mount tempfs, so add an option to ci.pl to not create tempfs. Also add some packages not in the base image.	2021-10-02 17:27:33 -04:00
David Steele	f1ed8f0e51	Sort WAL segment names when reporting duplicates. Make the output consistent even when files are listed in a different order. This is purely for testing purposes, but there is no harm in consistent output. Found on arm64.	2021-10-02 16:29:31 -04:00
David Steele	ae40ed6ec9	Add jobRetry parameter to HRN_CFG_LOAD(). Allow the default of 0 to be overridden to test retry behavior for commands.	2021-10-01 17:15:36 -04:00
David Steele	136d309dd4	Allow stack trace to be specified for errorInternalThrow(). This allows the stack trace to be set when an error is received by the protocol, rather than appending it to the message. Now these errors will look no different than any other error and the stack trace will be reported in the same way. One immediate benefit is that test.pl --vm-out --log-level-test=debug will work for tests that check expect log results. Previously, the test would error at the first check because the stack trace included in the message would not match the expected log output.	2021-10-01 15:29:31 -04:00
David Steele	62f6fbe2a9	Update file mode in info/manifest test to 0600. 0400 is not a very realistic mode. It may have become the default due to copy-pasting.	2021-10-01 10:15:34 -04:00
David Steele	cf1a57518f	Refactor restoreManifestMap() to be driven by link-map. This will allow new links to be added in a future commit. The current implementation is driven by the links that already exist in the manifest, which would make the new use case more complex to implement. Also, add a more helpful error when a tablespace link is specified.	2021-09-30 14:29:49 -04:00
David Steele	d89a67776c	Refactor restoreManifestMap() tests in the command/restore unit. Add test titles, new tests, and rearrange. Also manifestTargetFindDefault(), which will soon be used by core code in a refactoring commit.	2021-09-30 13:39:29 -04:00
David Steele	815377cc60	Finalize catalog number for PostgreSQL 14 release.	2021-09-30 13:27:14 -04:00
Stefan Fercot	baf186bfb0	Fix comment typos.	2021-09-29 12:03:01 -04:00
David Steele	9e79f0e64b	Add recovery start time to online backup restore log. This helps give an idea of how much recovery needs to be done to reach the end of the WAL stream and is easier to read than the backup label.	2021-09-29 10:31:51 -04:00
David Steele	9346895f5b	Rename page checksum error to error list in info text output. "error list" makes it clearer that other errors may be reported. For example, if checksum-page is true in the manifest but no checksum-page-error list is provided then the error is in alignment, i.e. the file size is not a multiple of the page size, with allowances made for a valid-looking partial page at the end of the file. It is still not possible to differentiate between alignment and page checksum errors in the output but this will be addressed in a future commit.	2021-09-29 09:58:47 -04:00
David Steele	b7ef12a76f	Add hints to standby replay timeout message.	2021-09-28 15:55:13 -04:00
David Steele	096829b3b2	Add repo-azure-uri-style option. Azurite introduced a breaking change in 8f63964e to use automatically host-style URIs when the endpoint appears to be a multipart hostname. This option allows the user to configure which style URI will be used, but changing the endpoint might cause breakage if Azurite decides to use a different style. Future changes to Azurite may also cause breakage.	2021-09-27 09:01:53 -04:00
David Steele	c8ea17c68f	Convert page checksum filter result to a pack. The pack is both more compact and more efficient than a variant. Also aggregate the page error info in the main process rather than in the filter to allow additional LSN filtering, to be added in a future commit.	2021-09-24 17:40:31 -04:00
David Steele	ac1f6db4a2	Centralize and optimize tag stack management. The push and pop code was duplicated in four places, so centralize the code into pckTagStackPop() and pckTagStackPush(). Also create a default bottom item for the stack to avoid allocating a list if there will only ever be the default container, which is very common. This avoids the extra time and memory to allocate a list.	2021-09-23 14:06:00 -04:00
David Steele	15e7ff10d3	Add Pack pseudo-type. Rather than working directly with Buffer types, define a new Pack pseudo-type that represents a Buffer containing a pack. This makes it clearer that a pack is being stored and allows stronger typing.	2021-09-23 08:31:32 -04:00
David Steele	131ac0ab5e	Rename pckReadNew()/pckWriteNew() to pckReadNewIo()/pckWriteNewIo(). These names more accurately describe the purpose of the constructors.	2021-09-22 11:18:12 -04:00

1 2 3 4 5 ...

2123 Commits