rust/neon - neon - Gitea: Git with a cup of tea

rust/neon

mirror of https://github.com/neondatabase/neon.git synced 2026-01-08 05:52:55 +00:00

Author	SHA1	Message	Date
Dmitry Rodionov	557e3024cd	Forward pageserver connection string from compute to safekeeper This is needed for implementation of tenant rebalancing. With this change safekeeper becomes aware of which pageserver is supposed to be used for replication from this particular compute.	2021-12-06 21:28:49 +03:00
Arseny Sher	bd34d7ecfc	Bump safekeeper control file version and allow reading the previous one. Should have been a part of `cba4da3f4d` to provide upgrade for previously existing clusters. Separates version independent header (magic + version) out of SafeKeeperState to choose what to deserialize.	2021-12-06 19:47:55 +03:00
Dmitry Ivanov	0a8c672630	[CI] Fix benchmarks Too bad we don't have a --dry-run in PRs :(	2021-12-06 13:52:28 +03:00
Dmitry Ivanov	b87ab17d05	Bump rust version to 1.56.1 Apparently, code coverage doesn't work that well in 1.55.	2021-12-06 13:27:52 +03:00
Dmitry Ivanov	d874675955	Collect coverage in CI	2021-12-06 13:27:52 +03:00
Dmitry Ivanov	5d37560308	Add bespoke glue script leveraging LLVM coverage tools	2021-12-06 13:27:52 +03:00
Dmitry Ivanov	7cec13d1df	Improve shutdown story for code coverage This patch introduces fixes for several problems affecting LLVM-based code coverage: * Daemonizing parent processes should call _exit() to prevent coverage data file corruption (.profraw) due to concurrent writes. Implement proper shutdown handlers in safekeeper.	2021-12-06 13:27:52 +03:00
anastasia	b7685eb6ba	Enable backpressure	2021-12-06 12:49:42 +03:00
anastasia	c7f3b4e62c	Clarify the meaning of StandbyReply LSNs: write_lsn - The last LSN received and processed by pageserver's walreceiver. flush_lsn - same as write_lsn. At pageserver it doesn't guarantees data persistence, but it's fine. We rely on safekeepers. apply_lsn - The LSN at which pageserver guaranteed persistence of all received data (disk_consistent_lsn).	2021-12-06 12:49:42 +03:00
Heikki Linnakangas	5bad2deff8	Don't hold 'timelines' lock over checkpoint. It was very noticeable that you while the checkpointer was busy, you could not e.g. open a new connection.	2021-12-03 07:42:10 -05:00
Arseny Sher	d39608c367	Fix passing start_offset to find_end_of_wal_segment.	2021-12-03 12:43:57 +03:00
Arseny Sher	cba4da3f4d	Add term history to safekeepers. Persist full history of term switches on safekeepers instead of storing only the single term of the highest entry (called epoch). This allows easily and correctly find the divergence point of two logs and truncate the obsolete part before overwriting it with entries of the newer proposer(s). Full history of the proposer is transferred in separate message before proposer starts streaming; it is immediately persisted by safekeeper, though he might not yet have entries for some older terms there. That's because we can't atomically append to WAL and update the control file anyway, so locally available WAL must be taken into account when looking at the history. We should sometimes purge term history entries beyond truncate_lsn; this is not done here. Per https://github.com/zenithdb/rfcs/pull/12 Closes #296. Bumps vendor/postgres.	2021-12-03 12:43:57 +03:00
Dmitry Rodionov	2669d140f8	use full commit sha for version info for builds in docker this is not needed, since environment variable with commit sha already contains full version	2021-12-01 17:35:57 +03:00
Heikki Linnakangas	f49ad33f1b	Initialize 'loaded' correctly in DeltaLayer. While we're at it, reuse the Book and the VirtualFile that's backing it even over unload() calls. Previously, we would keep the Book open, but on load(), we would re-open it anyway, which didn't make much sense. Now we reuse it it. Alternatively, perhaps we should close it on unload() to save some memory, but this I'm not going to think too hard about it right now as the whole load/unload thing is a bit of a hack and needs to be rewritten. This is hard to reproduce ATM, because the incorrect state would get fixed by an unload(). A checkpoint creates the DeltaLayer, and it also calls unload() afterwards, so the window is not very large. I hit it occasionally with a scale 1000 pgbench test, after I had modified InMemoryLayer::write_to_disk() to not write an image layer every time, which made the DeltaLayers be accessed more often.	2021-11-30 22:23:59 +02:00
Kirill Bulatov	670205e17a	Evict excessively failing sync tasks, improve processing for the rest of the tasks	2021-11-30 13:58:49 +02:00
Konstantin Knizhnik	f72d4814b1	Extract page images from FPI WAL records (#949 ) * Extract page images from FPI WAL records * Fix issues reported in review	2021-11-30 12:57:26 +03:00
Heikki Linnakangas	5ecf0664cc	Fix off-by-one error in check for future delta layers. This doesnt show up at the moment, because we never create a delta layer with end-LSN equal to the last LSN. We always create an image layer at that LSN instead. For example, if the latest processed LSN is 100, we would create a delta layer with end LSN 100 (exclusive), and an image layer at 100. But that's just how InMemoryLayer::write_to_disk happens to work at the moment, there's no fundamental reason it needs to always create that image layer. I noticed this bug when I tried to change the logic in InMemoryLayer::write_to_disk to only create an image layer after a few delta layers.	2021-11-29 14:35:24 +02:00
Heikki Linnakangas	7cae265447	Fix dump_layerfile. The VirtualFile machinery panics if it's not initialized	2021-11-29 11:26:54 +02:00
Heikki Linnakangas	5aa969a588	Replace in-memory layers and OOM-triggered eviction with temp files. The "in-memory layer" is misnomer now, each in-memory layer is now actually backed by a file. The files are ephemeral, in that they don't survive page server crash or shutdown. To avoid reading the file for every operation, "ephemeral files" are cached in a page cache. This includes changes from 'inmemory-layer-chunks' branch to serialize / the page versions when they are added to the open layer. The difference is that they are not serialized to the expandable in-memory "chunk buffer", but written out to the file.	2021-11-26 17:25:17 +03:00
Arthur Petukhovsky	93cc40584d	Shutdown socket on CopyFail (#938 ) Fixes #935	2021-11-26 16:48:27 +03:00
Dmitry Rodionov	130184fee9	Prohibit branch creation and basebackup at out of scope lsns Out of scope LSNs include pre initdb LSNs, and LSNs prior to latest_gc_cutoff. To get there there was also two cleanups: * Fix error handling in Execute message handler. This fixes behaviour when basebackup retured an error. Previously pageserver thread just died. * Remove "ancestor" file which previously contained ancestor id and branch lsn. Currently the same data can be obtained from metadata file. And just the way we handled ancestor file in the code introduced the case when branching fails timeline directory is created but there is no data in it except ancestor file. And this confused gc because it scans directories. So it is better to just remove ancestor file and clean up this timeline directory creation so it happens after all validity checks have passed	2021-11-25 15:27:16 +03:00
Heikki Linnakangas	d47f610606	Fix pageserver CLI parameter names and document them	2021-11-25 13:31:52 +02:00
Dmitry Rodionov	0650e51f0b	add test one more case for layer visibility	2021-11-22 11:39:20 +03:00
Dmitry Rodionov	737a557f09	add check to python tests that afteer gc number of rows is unchanged in all branches	2021-11-22 11:39:20 +03:00
Dmitry Rodionov	6f7ebe6e01	preserve data in parent branch that might be referenced in child branch	2021-11-22 11:39:20 +03:00
Dmitry Rodionov	70ab0d5b1f	add missing script	2021-11-19 00:10:40 +03:00
Dmitry Rodionov	6ac76248cf	Save performance test results from perfirmance test suit runs. Also render reports for both staging and local runs.	2021-11-19 00:00:19 +03:00
Kirill Bulatov	b32da3b42e	Use less pageserver-specific method in RemoteStorage trait	2021-11-18 22:53:40 +02:00
Dmitry Ivanov	0ccfc62e88	[proxy] Pass PostgreSQL version to client Fixes #779	2021-11-17 16:28:44 +03:00
Dmitry Ivanov	b55cf773a8	[proxy] Streamline control- and dataflow	2021-11-17 16:28:44 +03:00
Dmitry Ivanov	43ded1c54b	[proxy] Minor cleanup	2021-11-17 16:28:44 +03:00
Heikki Linnakangas	f8702d4625	Fix checking for whether segment exists on a frozen in-memory layer. Ever since we've had frozen in-memory layers, having an 'end_lsn' no longer means that the layer has been dropped. Need to check the 'dropped' flag explicitly. This was reliably causing a failure on the new 'test_parallel_copy' test in https://github.com/zenithdb/zenith/pull/864. I'm not sure why it doesn't happen on main branch, but the bug is pretty straightforward when you see it.	2021-11-15 20:19:15 +02:00
Dmitry Rodionov	44111e3ba3	Prohibit branch creation at lsn that was already garbage collected. This introduces new timeline field latest_gc_cutoff. It is updated before each gc iteration. New check is added to branch_timelines to prevent branch creation with start point less than latest_gc_cutoff. Also this adds a check to get_page_at_lsn which asserts that lsn at which the page is requested was not garbage collected. This check currently is triggered for readonly nodes which are pinned to specific lsn and because they are not tracked in pageserver garbage collection can remove data that still might be referenced. This is a bug and will be fixed separately.	2021-11-15 20:03:16 +03:00
Patrick Insinger	298bc588f9	pageserver - don't try to GC InMemoryLayers	2021-11-15 09:01:45 -08:00
Heikki Linnakangas	4ba521f53f	Add performance test case for parallel COPY TO	2021-11-15 14:49:53 +02:00
Heikki Linnakangas	431d32756b	Add a buffer cache, and use it to store materialized pages. The buffer cache is shared across all tenants, allowing memory to be dynamically allocated where it's needed the most. The cache works on 8 kB pages, and uses the clock algorithm for replacement policy; same as the PostgreSQL buffer cache. One peculiarity is that the materialized page versions can be looked up by an inexact LSN, to find the latest page version with an LSN >= the search key. The code is structured to support caching other kinds of pages in the same cache in the future, but with a different mapping key. Co-authored-by: Patrick Insinger <patrick@zenith.tech>	2021-11-12 11:02:12 -08:00
Heikki Linnakangas	3d172d98a3	Improve layered repo README. Add an informal overview of how it works.	2021-11-12 19:59:31 +02:00
Heikki Linnakangas	849ac791a6	Bandaid fix for "page not found" errors, when a table is loaded. During parallel load of a table, Postgres sometimes requests a page from the page server for which no WAL has been generated yet. That's normal; Postgres expects the page to be full of zeros. There was a special case for that in LayeredTimeline::materialize_page, but the problem remained when you're crossing a segment boundary, so that there's no layer for the segment at all. It would be nice to have a more robust cross-check for this case. That might need help from the Postgres side. But this extends the bandaid fix we had in materialize_page() to the case where cross segment boundary. Fixes https://github.com/zenithdb/zenith/issues/841	2021-11-12 18:47:39 +02:00
Alexey Kondratov	de5e6a15ae	Set LD_LIBRARY_PATH in the check_restored_datadir_content() psql call Otherwise we may use outdated system libpq. Also print stdout/stderr if basebackup failed in check_restored_datadir_content()	2021-11-12 16:27:43 +03:00
Alexey Kondratov	0d6bf14ecb	Use vendor/postgres rebased on top of REL_14_1	2021-11-12 16:27:43 +03:00
Heikki Linnakangas	d1e79c4af3	Fix locking issues in VirtualFile machinery. There were two separate locking issues that could lead to a deadlock, both related to holding a lock for longer than necessary: 1. In the loop in `VirtualFile::with_file`, the "handle_guard" was held across iterations of the loop. Because of that, if the handle was changed by a concurrent thread, the loop would try to acquire the handle lock, when it was still holding the lock from previous iteration. To fix, release the lock earlier. There was no need to hold it across iterations, it was just accidental. 2. In the same function, we also held the "slot_guard" longer than necessary. It's only needed in the first part of the loop, where we check if the current handle is valid. If it's not, the slot lock can be immediately released. But it was not, it was kept over the acquisition of the handle lock. I'm not sure if that alone could cause problems, but let's release the lock as soon as possible anyway. Add a test case, based on Konstantin's test program to demonstrate the deadlock.	2021-11-11 20:12:59 +02:00
Kirill Bulatov	abb2ac5246	Better context when erroring	2021-11-11 19:22:05 +02:00
Kirill Bulatov	99dbbe5f18	Allow downloading remote files partially	2021-11-11 18:51:34 +02:00
Arseny Sher	e7ca8ef5a8	Use PG timelineid 1 everywhere. As changing it doesn't have useful meaning in Zenith. ref #824	2021-11-11 13:53:39 +03:00
Patrick Insinger	1ce4976e36	pageserver - track size of VecMaps	2021-11-10 11:09:34 -08:00
Heikki Linnakangas	9300107cdf	Cache Book objects, use virtual files to avoid running out of fds. Currently, whenever a page version is needed from an image or delta layer, we open the file and read and parse the bookfile headers. That's pretty expensive. To reduce the overhead, introduce a cache of open file descriptors, and use that to cache the Book objects so that we don't need to read the metadata on every access.	2021-11-10 17:19:37 +02:00
Arthur Petukhovsky	9aaa02bc9a	Fix high CPU usage in walproposer (#860 ) * Bump vendor/postgres * Update time limits for test_restarts_under_load	2021-11-10 17:18:07 +03:00
Arseny Sher	5603259c53	In wal_proposer_recovery, don't wait outcoming WAL to be committed. Otherwise we're deadlocking ourselves. Oversight of `33007cc`.	2021-11-10 01:38:25 +03:00
Arseny Sher	ce15c62f35	Fix 'send WAL up to' debug logging.	2021-11-10 01:38:25 +03:00
Egor Suvorov	eaff0cd568	Check python for the whole repository and improve docs (#813 )	2021-11-09 22:23:29 +03:00

1 2 3 4 5 ...

1085 Commits