scylladb

Author	SHA1	Message	Date
Tomasz Grabiec	c9e1694c58	Merge "Some optimizations on cache entry lookup" from Pavel Emelyanov The set contains 3 small optimizations: - avoid copying of partition key on lookup path - reduce number of args carried around when creating a new entry - save one partition key comparison on reader creation Plus related satellite cleanups. * https://github.com/xemul/scylla/tree/br-row-cache-less-copies: row_cache: Revive do_find_or_create_entry concepts populating reader: Do not copy decorated key too early populating reader: Less allocator switching on population populating reader: Fix indentation after previous patch row_cache: Move missing entry creation into helper test: Lookup an existing entry with its own helper row_cache: Do not copy partition tombstone when creating cache entry row_cache: Kill incomplete_tag row_cache: Save one key compare on direct hit	2020-09-15 17:49:47 +02:00
Nadav Har'El	3322328b21	alternator test: fix two tests that failed in HTTPS mode When the test suite is run with Scylla serving in HTTPS mode, using test/alternator/run --https, two Alternator Streams tests failed. With this patch fixing a bug in the test, the tests pass. The bug was in the is_local_java() function which was supposed to detect DynamoDB Local (which behaves in some things differently from the real DynamoDB). When that detection code makes an HTTPS request and does not disable checking the server's certificate (which on Alternator is self-signed), the request fails - but not in the way that the code expected. So we need to fix the is_local_java() to allow the failure mode of the self-signed certificate. Anyway, this case is not DynamoDB Local so the detection function would return false. Fixes #7214 Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20200910194738.125263-1-nyh@scylladb.com>	2020-09-11 08:10:13 +02:00
Nadav Har'El	02ee0483b2	alternator test: add reproducing tests for several issues This patch adds regression tests for four recently-fixed issues which did not yet have tests: Refs #7157 (LatestStreamArn) Refs #7158 (SequenceNumber should be numeric) Refs #7162 (LatestStreamLabel) Refs #7163 (StreamSpecification) I verified that all the new tests failed before these issues were fixed, but now pass. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20200907155334.562844-1-nyh@scylladb.com>	2020-09-10 17:36:23 +02:00
Avi Kivity	0e03c979d2	Merge 'Fix ignoring cells after null in appending hash' from Piotr Sarna " This series fixes a bug in `appending_hash<row>` that caused it to ignore any cells after the first NULL. It also adds a cluster feature which starts using the new hashing only after the whole cluster is aware of it. The series comes with tests, which reproduce the issue. Fixes #4567 Based on #4574 " * psarna-fix_ignoring_cells_after_null_in_appending_hash: test: extend mutation_test for NULL values tests/mutation: add reproducer for #4567 gms: add a cluster feature for fixed hashing digest: add null values to row digest mutation_partition: fix formatting appending_hash<row>: make publicly visible	2020-09-10 15:35:38 +03:00
Piotr Sarna	fe5cd846b5	test: extend mutation_test for NULL values The test is extended for another possible corner case: [1, NULL, 2] vs [1, 2, NULL] should have different digests. Also, a check for legacy behavior is added.	2020-09-10 13:16:44 +02:00
Paweł Dziepak	287d0371fa	tests/mutation: add reproducer for #4567	2020-09-10 13:16:44 +02:00
Dejan Mircevski	9d02f10c71	cql3: Fix NULL reference in get_column_defs_for_filtering There was a typo in get_column_defs_for_filtering(): it checked the wrong pointer before dereferencing. Add a test exposing the NULL dereference and fix the typo. Tests: unit (dev) Fixes #7198. Signed-off-by: Dejan Mircevski <dejan@scylladb.com>	2020-09-10 08:45:07 +02:00
Wojciech Mitros	66e8214606	cql: Forbid adding new fields to UDTs used in partition key columns Changing a user type may allow adding apparently duplicate rows to tables where this type is used in a partitioning key. Fix by checking all types of existing partitioning columns before allowing to add new fields to the type. Fixes #6941	2020-09-08 16:08:07 +03:00
Avi Kivity	7ac59dcc98	lsa: decay reserves The log-structured allocator (LSA) reserves memory when performing operations, since its operations are performed with reclaiming disabled and if it runs out, it cannot evict cache to gain more. The amount of memory to reserve is remembered across calls so that it does not have to repeat the fail/increase-reserve/retry cycle for every operation. However, we currently lack decaying the amount to reserve. This means that if a single operation increased the reserve in the distant past, all current operations also require this large reserve. Large reserves are expensive since they can cause large amounts of cache to be evicted. This patch adds reserve decay. The time-to-decay is inversely proportional to reserve size: 10GB/reserve. This means that a 20MB reserve is halved after 500 operations (10GB/20MB) while a 20kB reserve is halved after 500,000 operations (10GB/20kB). So large, expensive reserves are decayed quickly while small, inexpensive reserves are decayed slowly to reduce the risk of allocation failures and exceptions. A unit test is added. Fixes #325.	2020-09-08 15:59:25 +03:00
Etienne Adam	63a1a4cbb9	redis: add hgetall and hdel commands This patch adds support for 2 hash commands HDEL and HGETALL. Internally it introduces the hashes_result_builder class to read hashes and stored them in a std::map. Other changes: - one exception return string was fixed - tests now use pytest.raises Signed-off-by: Etienne Adam <etienne.adam@gmail.com> Message-Id: <20200907202528.4985-1-etienne.adam@gmail.com>	2020-09-08 11:59:52 +03:00
Piotr Grabowski	ffd8c8c505	utf8: Print invalid UTF-8 character position Add new validate_with_error_position function which returns -1 if data is a valid UTF-8 string or otherwise a byte position of first invalid character. The position is added to exception messages of all UTF-8 parsing errors in Scylla. validate_with_error_position is done in two passes in order to preserve the same performance in common case when the string is valid.	2020-09-07 18:11:21 +03:00
Botond Dénes	c01af1d9d2	tests/boost/multishard_mutation_query_test: remove last BOOST_REQUIRE* macros Previous patches removed those `BOOST_REQUIRE` macros that could be invoked from shards other than 0. The reason is that said macros are not thread-safe, so calling them from multiple shards produces mangled output to stdout as well as the XML report file. It was assumed that only these invocations -- from a non-0 shard -- are problematic, but it turns out even these can race with seastar log messages emitted from other shards. This patch removes all such macros, replacing them with the thread safe `require` functions from `test/lib/test_utils.hh`. Signed-off-by: Botond Dénes <bdenes@scylladb.com> Message-Id: <20200907125309.1199104-1-bdenes@scylladb.com>	2020-09-07 17:07:26 +03:00
Tomasz Grabiec	bcdcf06ec7	Merge "lwt: for each statement in cas_request provide a row in CAS result set" from Pavel Solodovnikov Previously batch statement result set included rows for only those updates which have a prefetch data present (i.e. there was an "old" (pre-existing) row for a key). Also, these rows were sorted not in the order in which statements appear in the batch, but in the order of updated clustering keys. If we have a batch which updates a few non-existent keys, then it's impossible to figure out which update inserted a new key by looking at the query response. Not only because the responses may not correspond to the order of statements in the batch, but even some rows may not show up in the result set at all. Please see #7113 on Github for detailed description of the problem: https://github.com/scylladb/scylla/issues/7113 The patch set proposes the following fix: For conditional batch statements the result set now always includes a row for each LWT statement, in the same order in which individual statements appear in the batch. This way we can always tell which update did actually insert a new key or update the existing one. Technically, the following changes were made: * `update_parameters::prefetch_data::row::is_in_cas_result_set` member removed as well as the supporting code in `cas_request::applies_to` which iterated through cas updates and marked individual `prefetch_data` rows as "need to be in cas result set". * `cas_request::applies_to` substantially simplified since it doesn't do anything more than checking `stmt.applies_to()` in short-circuiting manner. * `modification_statement::build_cas_result_set` method moved to `cas_request`. This allows to easily iterate through individual `cas_row_update` instances and preserve the order of the rows in the result set. * A little helper `cas_request::find_old_row` is introduced to find a row in `prefetch_data` based on the (pk, ck) combination obtained from the current `cas_request` and a given `cas_row_update`. * A few tests for the issue #7113 are written, other lwt-batch-related tests adjusted accordingly.	2020-09-04 16:09:45 +02:00
Pavel Solodovnikov	92fd515186	lwt: for each statement in cas_request provide a row in CAS result set Previously batch statement result set included rows for only those updates which have a prefetch data present (i.e. there was an "old" (pre-existing) row for a key). Also, these rows were sorted not in the order in which statements appear in the batch, but in the order of updated clustering keys. If we have a batch which updates a few non-existent keys, then it's impossible to figure out which update inserted a new key by looking at the query response. Not only because the responses may not correspond to the order of statements in the batch, but even some rows may not show up in the result set at all. The patch proposes the following fix: For conditional batch statements the result set now always includes a row for each LWT statement, in the same order in which individual statements appear in the batch. This way we can always tell which update did actually insert a new key or update the existing one. `update_parameters::prefetch_data::row::is_in_cas_result_set` member variable was removed as well as supporting code in `cas_request::applies_to` which iterated through cas updates and marked individual `prefetch_data` rows as "need to be in cas result set". Instead now `cas_request::applies_to` is significantly simplified since it doesn't do anything more than checking `stmt.applies_to()` in short-circuiting manner. A few tests for the issue are written, other lwt-batch-related tests were adjusted accordingly to include rows in result set for each statement inside conditional batches. Tests: unit(dev, debug) Co-authored-by: Konstantin Osipov <kostja@scylladb.com> Signed-off-by: Pavel Solodovnikov <pa.solodovnikov@scylladb.com>	2020-09-04 13:13:26 +03:00
Pavel Emelyanov	84a6d439ad	test: Lookup an existing entry with its own helper The only caller of find_or_create() in tests works on already existing (.populate()-d) entry, so patch this place for explicity and for the sake of next patching. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-09-03 21:13:21 +03:00
Avi Kivity	64c7c81bac	Merge "Update log messages to {fmt} rules" from Pavel E " Before seastar is updated with the {fmt} engine under the logging hood, some changes are to be made in scylla to conform to {fmt} standards. Compilation and tests checked against both -- old (current) and new seastar-s. tests: unit(dev), manual " * 'br-logging-update' of https://github.com/xemul/scylla: code: Force formatting of pointer in .debug and .trace code: Format { and } as {fmt} needs streaming: Do not reveal raw pointer in info message mp_row_consumer: Provide hex-formatting wrapper for bytes_view heat_load_balance: Include fmt/ranges.h	2020-09-03 15:10:09 +03:00
Nadav Har'El	1d06da18fc	alternator test: test for the TRIM_HORIZON stream iterator This patch adds a test for the TRIM_HORIZON option of GetShardIterator in Alternator Streams. This option asks to fetch again all the available history in this shard stream. We had an implementation for it, but not a test - so this patch adds one. The test passes. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20200830131458.381350-1-nyh@scylladb.com>	2020-09-02 18:37:06 +02:00
Nadav Har'El	3d4183863a	alternator test: add tests for sequence-number based iterators Alternator Streams already support the AT_SEQUENCE_NUMBER and AFTER_SEQUENCE_NUMBER options for iterators. These options allow to replay a stream of changes from a known position or after that known position. However, we never had a test verifying that these features actually work as intended, beyond just checking syntax. Having such tests is important because recently we changed the implementation of these iterators, but didn't have a test verifying that they still work. So in this patch we add such tests. The tests pass (as usual, on both Alternator and DynamoDB). Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20200830115817.380075-1-nyh@scylladb.com>	2020-09-02 18:36:59 +02:00
Nadav Har'El	c879b23b82	alternator test: add another passing Alternator Streams test We had a test, test_streams_last_result, that verifies that after reading from an Alternator Stream the last event, reading again will find nothing. But we didn't actually have a test which checks that if at that point a new event does arrive, we can read it. This test checks this case, and it passes (we don't have a bug there, but it's good as a regression test for NextShardIterator). This test also verifies that after reading an event for a particular key on a a specific stream "shard", the next event for the same key will arrive on the same shard. This test passes on both Alternator and DynamoDB. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20200830105744.378790-1-nyh@scylladb.com>	2020-09-02 18:36:50 +02:00
Kamil Braun	ff78a3c332	cdc: rename CDC description tables... again Commit `a6ad70d3da` changed the format of stream IDs: the lower 8 bytes were previously generated randomly, now some of them have semantics. In particular, the least significant byte contains a version (stream IDs might evolve with further releases). This is a backward-incompatible change: the code won't properly handle stream IDs with all lower 8 bytes generated randomly. To protect us from subtle bugs, the code has an assertion that checks the stream ID's version. This means that if an experimental user used CDC before the change and then upgraded, they might hit the assertion when a node attempts to retrieve a CDC generation with old stream IDs from the CDC description tables and then decode it. In effect, the user won't even be able to start a node. Similarly as with the case described in `d89b7a0548`, the simplest fix is to rename the tables. This fix must get merged in before CDC goes out of experimental. Now, if the user upgrades their cluster from a pre-rename version, the node will simply complain that it can't obtain the CDC generation instead of preventing the cluster from working. The user will be able to use CDC after running checkAndRepairCDCStreams. Since a new table is added to the system_distributed keyspace, the cluster's schema has changed, so sstables and digests need to be regenerated for schema_digest_test.	2020-08-31 11:33:14 +03:00
Calle Wilund	78236c015a	cdc: Remove cdc delta_mode::off Fixes #7128 CDC logs are not useful without at least delta_mode==keys, since pre/post image data has no info on _what_ was actually done to base table in source mutation.	2020-08-31 07:59:40 +00:00
Etienne Adam	19683d04c6	redis: add hget and hset commands hget and hset commands using hashes internally, thus they are not using the existing write_strings() function. Limitations: - hset only supports 3 params, instead of multiple field/value list that is available in official redis-server. - hset should return 0 when the key and field already exists, but I am not sure it's possible to retrieve this information without doing read-before-write, which would not be atomic. I factorized a bit the query_* functions to reduce duplication, but I am not 100% sure of the naming, it may still be a bit confusing between the schema used (strings, hashes) and the returned format (currently only string but array should come later with hgetall). Signed-off-by: Etienne Adam <etienne.adam@gmail.com> Message-Id: <20200830190128.18534-1-etienne.adam@gmail.com>	2020-08-30 22:05:41 +03:00
Rafael Ávila de Espíndola	d18af34205	everywhere: Use future::get0 when appropriate This works with current seastar and clears most of the way for updating to a version that doesn't use std::tuple in futures. Signed-off-by: Rafael Ávila de Espíndola <espindola@scylladb.com> Message-Id: <20200826231947.1145890-1-espindola@scylladb.com>	2020-08-27 15:05:51 +03:00
Nadav Har'El	1da3af5420	alternator test: enable a passing test After issue #7107 was fixed (regarding the correctness of OldImage and NewImage in Alternator Streams) we forgot to remove the "xfail" tag from one of the tests for this issue. This test now passes, as expected, so in this patch we remove the xfail tag. Refs #7107 Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20200827103054.186555-1-nyh@scylladb.com>	2020-08-27 14:15:32 +03:00
Nadav Har'El	95afadfe21	merge: alternator_streams: Include keys in OldImage/NewImage Merged pull request https://github.com/scylladb/scylla/pull/7063 By Calle Wilund: Fixes #6935 DynamoDB streams for some reason duplicate the record keys into both the "Keys" and "OldImage"/"NewImage" sub-objects when doing GetRecords. But only if there is other data to include. This patch appends the pk/ck parts into old/new image iff we had any record data. Updated to handle keys-only updates, and distinguish creating vs. updating rows. Changes cdc to not generate preimage for non-existent/deleted rows, and also fixes missing operations/ttls in keys-only delta mode. alternator_streams: Include keys in OldImage/NewImage cdc: Do not generate pre/post image for non-existent rows	2020-08-27 11:23:35 +03:00
Calle Wilund	678ecc7469	alternator_streams: Include keys in OldImage/NewImage Fixes #6935 Fixes #7107 DynamoDB streams for some reason duplicate the record keys into both the "Keys" and "OldImage"/"NewImage" sub-objects when doing GetRecords. This patch appends the pk/ck parts into old/new image, and also removes the previous restrictions on image generation since cdc now generates more consistent pre/post image data.	2020-08-26 18:14:09 +00:00
Calle Wilund	e50911e5b0	cdc: Do not generate pre/post image for non-existent rows Fixes #7119 Fixes #7120 If preimage select came up empty - i.e. the row did not exist, either due to never been created, or once delete, we should not bother creating a log preimage row for it. Esp. since it makes it harder to interpret the cdc log. If an operation in a cdc batch did a row delete (ranged, ck, etc), do not generate postimage data, since the row does no longer exist. Note that we differentiate deleting all (non-pk/ck) columns from actual row delete.	2020-08-26 18:14:09 +00:00
Pavel Emelyanov	812eed27fe	code: Force formatting of pointer in .debug and .trace ... and tests. Printin a pointer in logs is considered to be a bad practice, so the proposal is to keep this explicit (with fmt::ptr) and allow it for .debug and .trace cases. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-08-26 20:44:11 +03:00
Pavel Emelyanov	366b4e8a8f	code: Format { and } as {fmt} needs There are two places that want to print "{<text>}" strings, but do not format the curly braces the {fmt}-way. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-08-26 20:44:11 +03:00
Avi Kivity	3daa49f098	Merge "materialized views: Fix undefined behavior on base table schema changes" from Tomasz " The view_info object, which is attached to the schema object of the view, contains a data structure called "base_non_pk_columns_in_view_pk". This data structure contains column ids of the base table so is valid only for a particular version of the base table schema. This data structure is used by materialized view code to interpret mutations of the base table, those coming from base table writes, or reads of the base table done as part of view updates or view building. The base table schema version of that data structure must match the schema version of the mutation fragments, otherwise we hit undefined behavior. This may include aborts, exceptions, segfaults, or data corruption (e.g. writes landing in the wrong column in the view). Before this patch, we could get schema version mismatch here after the base table was altered. That's because the view schema did not change when the base table was altered. Another problem was that view building was using the current table's schema to interpret the fragments and invoke view building. That's incorrect for two reasons. First, fragments generated by a reader must be accessed only using the reader's schema. Second, base_non_pk_columns_in_view_pk of the recorded view ptrs may not longer match the current base table schema, which is used to generate the view updates. Part of the fix is to extract base_non_pk_columns_in_view_pk into a third entity called base_dependent_view_info, which changes both on base table schema changes and view schema changes. It is managed by a shared pointer so that we can take immutable snapshots of it, just like with schema_ptr. When starting the view update, the base table schema_ptr and the corresponding base_dependent_view_info have to match. So we must obtain them atomically, and base_dependent_view_info cannot change during update. Also, whenever the base table schema changes, we must update base_dependent_view_infos of all attached views (atomically) so that it matches the base table schema. Fixes #7061. Tests: - unit (dev) - [v1] manual (reproduced using scylla binary and cqlsh) " * tag 'mv-schema-mismatch-fix-v2' of github.com:tgrabiec/scylla: db: view: Refactor view_info::initialize_base_dependent_fields() tests: mv: Test dropping columns from base table db: view: Fix incorrect schema access during view building after base table schema changes schema: Call on_internal_error() when out of range id is passed to column_at() db: views: Fix undefined behavior on base table schema changes db: views: Introduce has_base_non_pk_columns_in_view_pk()	2020-08-26 17:37:52 +03:00
Nadav Har'El	d4b452002a	alternator test: tests for NewImage feature of Alternator Streams This patch adds tests for the "NewImage" attribute in Alternator Streams in NEW_IMAGE and NEW_AND_OLD_IMAGES mode. It reproduces issue #7107, that items' key attributes are missing in the NewImage. It also verifies the risky corner cases where the new item is "empty" and NewImage should include just the key, vs. the case where the item is deleted, so NewImage should be missing. This test currently passes on AWS DynamoDB, and xfails on Alternator. Refs #7107. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20200825113857.106489-1-nyh@scylladb.com>	2020-08-26 11:33:23 +03:00
Nadav Har'El	8e06734893	redis test: add default host and port test/redis/README.md suggests that when running "pytest" the default is to connect to a local redis on localhost:6379. This default was recently lost when options were added to use a different host and port. It's still good to have the default suggested in README.md. It also makes it easier to run the tests against the standard redis, which by default runs on localhost:6379 - by just running "pytest". Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20200825195143.124429-1-nyh@scylladb.com>	2020-08-26 11:33:23 +03:00
Calle Wilund	5ed3d6892d	cdc: Remove stored (postimage) data when doing row delete Fixes #6900 Clustered range deletes did not clear out the "row_states" data associated with affected rows (might be many). Adds a sweep through and erases relevant data. Since we do pre- and postimage in "order", this should only affect postimage.	2020-08-25 12:27:18 +03:00
Piotr Jastrzebski	f01ce1458f	cdc: Preserve metadata columns when geting only keys for delta Fixes #7095 Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2020-08-25 10:41:54 +03:00
Raphael S. Carvalho	1c29f0a43d	cql3/statements: verify that counter column cannot be added into non-counter table A check, to validate that counter column cannot be added into non-counter table, is missing for alter table statement. Validation is performed when building new schema, but it's limited to checking that a schema will not contain both counter and non-counter columns. Due to lack of validation, the added counter column could be incorrectly persisted to the schema, but this results in a crash when setting the new schema to its table. On restart, it can be confirmed that the schema change was indeed persisted when describing the table. This problem is fixed by doing proper validation for the alter table statement, which consists of making sure a new counter column cannot be added to a non-counter table. The test cdc_disallow_cdc_for_counters_test is adjusted because one of its tests was built on the assumption that counter column can be added into a non-counter table. Fixes #7065. Signed-off-by: Raphael S. Carvalho <raphaelsc@scylladb.com> Message-Id: <20200824155709.34743-1-raphaelsc@scylladb.com>	2020-08-25 10:41:54 +03:00
Takuya ASADA	fe8679a6ee	test/redis: make redis tests runnable from test.py Just like test/alternator, make redis-test runnable from test.py. For this we move the redis tests into a subdirectory of tests/, and create a script to run them: tests/redis/run. These tests currently fail, so we did not yet modify test.py to actually run them automatically. Fixes #6331	2020-08-23 20:31:45 +03:00
Pavel Emelyanov	a6e6856e1f	compaction: Keep database reference on cleanup options The database is available at both places that create the options -- tests and API perform_cleanup call. Options object doesn't over-survive the returned future, so it's safe to keep the reference on it. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-08-21 14:58:40 +03:00
Tomasz Grabiec	c44455d514	Merge "Miscellaneous schema code cleanups" from Rafael	2020-08-20 15:19:42 +02:00
Tomasz Grabiec	617ccc5408	tests: mv: Test dropping columns from base table Reproduces #7061.	2020-08-20 14:53:07 +02:00
Nadav Har'El	a1453303f8	alternator test: test for OLD_IMAGE of an empty item We already have a test, test_streams.py::test_streams_updateitem_old_image, for issue #6935: It tests that the OLD_IMAGE in Alternator Streams should contain the item's key. However this test was missing one corner case, which is the first solution for this issue did incorrectly. So in this patch we add a test for this corner case, test_streams_updateitem_old_image_empty_item: This corner case about the item existing, but empty, i.e., having just the key but no other attribute. In this case, OLD_IMAGE should return that empty item - including its key. Not nothing. As usual, this test passes on DynamoDB and xfails on Alternator, and the "xfail" mark will be removed when issue #6935 is fixed. Refs #6935. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20200819155229.34475-1-nyh@scylladb.com>	2020-08-20 13:22:40 +02:00
Avi Kivity	392e24d199	Merge "Unglobal messaging service" from Pavel E " The messaging service is (as many other services) present in the global namespace and is widely accessed from where needed with global get(_local)?_messaging_service() calls. There's a long-term task to get rid of this globality and make services and componenets reference each-other and, for and due-to this, start and stop in specific order. This set makes this for the messaging service. The service is very low level and doesn't depend on anything. It's used by gossiper, streaming, repair, migration manager, storage proxy, storage service and API. According to this dependencies the set consists of several parts: patches 1-9 are preparatory, they encapsulate messaging service init/fini stuff in its own module and decouple it from the db::config patch 10-12 introduce local service reference in main and set its init/fini calls at the early stage so that this reference can later be passed to those depending on it patches 13-42 replace global referencing of messaging service from other subsystems with local references initialized from main. patch 43 finalizes tests. patch 44 wraps things up with removing global messaiging service instance along with get(_local)?_messaging_service calls. The service's stopping part is deliberately left incomplete (as it is now), the sharded service remains alive, only the instance's stop() method is called (and is empty for a while). Since the messaging service's users still do not stop cleanly, its instances should better continue leaking on exit. Once (if) the seastar gets the helper rpc::has_handlers() method merged the messaging_service::stop() will be able to check if all the verbs had been unregistered (spoiler: not yet, more fixes to come). For debugging purposes the pointer on now-local messaging service instance is kept in service::debug namespace. tests: unit(dev) dtest(dev: simple_boot_shutdown, repair, update_cluster_layout) manual start-stop " * 'br-unglobal-messaging-service-2' of https://github.com/xemul/scylla: (44 commits) messaging_service: Unglobal messaging service instance tests: Use own instances of messaging_service storage_service: Use local messaging reference storage_service: Keep reference on sharded messaging service migration_manager: Add messaging service as argument to get_schema_definition migration_manager: Use local messaging reference in simple cases migration_manager: Keep reference on messaging migration_manager: Make push_schema_mutation private non-static method migration_manager: Move get_schema_version verb handling from proxy repair: Stop using global messaging_service references repair: Keep sharded messaging service reference on repair_meta repair: Keep sharded messaging service reference on repair_info repair: Keep reference on messaging in row-level code repair: Keep sharded messaging service in API repair: Unset API endpoints on stop repair: Setup API endpoints in separate helper repair: Push the sharded<messaging_service> reference down to sync_data_using_repair repair: Use existing sharded db reference repair: Mark repair.cc local functions as static streaming: Keep messaging service on send_info ...	2020-08-20 12:20:36 +03:00
Rafael Ávila de Espíndola	6363716799	schema: Pass an rvalue to set_compaction_strategy_options This produces less code and makes sure every caller moves the value. Signed-off-by: Rafael Ávila de Espíndola <espindola@scylladb.com>	2020-08-19 14:02:35 -07:00
Pavel Emelyanov	ee41645a1a	tests: Use own instances of messaging_service The global one is going away, no core code uses it, so all tests can be safely switched to use their own instances. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-08-19 20:50:53 +03:00
Pavel Emelyanov	4ea3c2797c	storage_service: Keep reference on sharded messaging service It is a bit step backward in the storage-service decompsition campaign, but... Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-08-19 20:50:53 +03:00
Pavel Emelyanov	6c49127d04	migration_manager: Keep reference on messaging That's another user of messaging service, init it with private reference. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-08-19 20:50:53 +03:00
Pavel Emelyanov	24cb1b781f	storage_proxy: Keep reference on messaging The proxy is another user of messaging, so keep the reference on it. Its real usage will come in next patches. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-08-19 20:50:52 +03:00
Pavel Emelyanov	65bd54604d	gossiper: Use messaging service by reference Gossiper needs messaging service, the messaging is started before the gossiper, so we can push the former reference into it. Gossiper is not stopped for real, neither the messaging service is, so the memory usage is still safe. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-08-19 20:50:52 +03:00
Botond Dénes	6ad80f0adb	test/lib/cql_test_env: set debug::db pointer To allow using scylla-gdb.py scripts for debugging tests. These scripts expect a valid database pointer in `debug::db`. Signed-off-by: Botond Dénes <bdenes@scylladb.com> Message-Id: <20200819145632.2423462-1-bdenes@scylladb.com>	2020-08-19 19:13:05 +03:00
Avi Kivity	6f986df458	Merge "Fix TWCS compaction aggressiveness due to data segregation" from Raphael " After data segregation feature, anything that cause out-of-order writes, like read repair, can result in small updates to past time windows. This causes compaction to be very aggressive because whenever a past time window is updated like that, that time window is recompacted into a single SSTable. Users expect that once a window is closed, it will no longer be written to, but that has changed since the introduction of the data segregation future. We didn't anticipate the write amplification issues that the feature would cause. To fix this problem, let's perform size-tiered compaction on the windows that are no longer active and were updated because data was segregated. The current behavior where the last active window is merged into one file is kept. But thereafter, that same window will only be compacted using STCS. Fixes #6928. " * 'fix_twcs_agressiveness_after_data_segregation_v2' of github.com:raphaelsc/scylla: compaction/twcs: improve further debug messages compaction/twcs: Improve debug log which shows all windows test: Check that TWCS properly performs size-tiered compaction on past windows compaction/twcs: Make task estimation take into account the size-tiered behavior compaction/stcs: Export static function that estimates pending tasks compaction/stcs: Make get_buckets() static compact/twcs: Perform size-tiered compaction on past time windows compaction/twcs: Make strategy easier to extend by removing duplicated knowledge compaction/twcs: Make newest_bucket() non-static compaction/twcs: Move TWCS implementation into source file	2020-08-19 17:19:01 +03:00
Pavel Emelyanov	dc0918e255	tests: Keep local reference on global messaging Some tests directly reference the global messaging service. For the sake of simpler patching wrap this global reference with a local one. Once the global messaging service goes away tests will get their own instances. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-08-19 13:08:12 +03:00

1 2 3 4 5 ...

782 Commits