scylladb

mirror of https://github.com/scylladb/scylladb.git synced 2026-06-05 06:23:03 +00:00

Author	SHA1	Message	Date
Benny Halevy	d38206587e	table: enable_auto_compaction: trigger compaction auto_compaction has been disabled so sstables may have already been accumulated and require compaction. Do not wait for new sstables to be written to trigger compaction, trigger compaction right away. Refs #9784 Signed-off-by: Benny Halevy <bhalevy@scylladb.com> Message-Id: <20211212090632.1257829-1-bhalevy@scylladb.com>	2021-12-14 11:15:23 +02:00
Nadav Har'El	41c7b2fb4b	test/cql-pytest run: fix inaccurate comment The code in test/cql-pytest/run.py can start Scylla (or Cassandra, or Redis, etc.) in a random IP address in 127.... We explained in a comment that 127.0.0. is used by CCM so we avoid it in case someone runs both dtest and test.py in parallel on the same machine. But this observation was not accurate: Although the original CCM did use only 127.0.0., in Scylla's CCM we added in 2017, in commit 00d3ba5562567ab83190dd4580654232f4590962, the ability to run multiple copies of CCM in parallel; CCM now uses 127.0.., not just 127.0.0.. So we need to correct this in the comment. Luckily, the code doesn't need to change! We already avoided the entire 127.0.. for simplicity, not just 127.0.0.*. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20211212151339.1361451-1-nyh@scylladb.com>	2021-12-13 18:12:11 +02:00
Avi Kivity	e44a28dce4	Merge "compaction: Allow data from different buckets (e.g. windows) to be compacted together" from Raphael " Today, data from different buckets (e.g. windows) cannot be compacted together because mutation compactor happens inside each consumer, where each consumer is done on behalf of a particular bucket. To solve this problem, mutation compaction process is being moved from consumer into producer, such that interposer consumer, which is responsible for segregation, will be feeded with compacted data and forward it into the owner bucket. Fixes #9662. tests: unit(debug). " * 'compact_across_buckets_v2' of github.com:raphaelsc/scylla: tests: sstable_compaction_test: add test_twcs_compaction_across_buckets compaction: Move mutation compaction into producer for TWCS compaction: make enable_garbage_collected_sstable_writer() more precise	2021-12-12 15:07:15 +02:00
Raphael S. Carvalho	7c90088152	tests: sstable_compaction_test: add test_twcs_compaction_across_buckets Verify that TWCS compaction can now compact data across time windows, like a tombstone which will cause all shadowed data to be purged once they're all compacted together. Signed-off-by: Raphael S. Carvalho <raphaelsc@scylladb.com>	2021-12-10 17:14:45 -03:00
Raphael S. Carvalho	9b8aa1e9ae	compaction: Move mutation compaction into producer for TWCS If interposer is enabled, like the timestamp-based one for TWCS, data from different buckets (e.g. windows) cannot be compacted together because mutation compaction happens inside each consumer, where each consumer will be belong to a different bucket. To remove this limitation, let's move the mutation compactor from consumer into producer, such that compacted data will be feeded into the interposer, before it segregates data. We're short-circuiting this logic if TWCS isn't in use as compacting reader adds overhead to compaction, given that this reader will pop fragments from combined sstable reader, compact them using mutation_compactor and finally push them out to the underlying reader. without compacting reader (e.g. STCS + no interposer): 228255.92 +- 1519.53 partitions / sec (50 runs, 1 concurrent ops) 224636.13 +- 1165.05 partitions / sec (100 runs, 1 concurrent ops) 224582.38 +- 1050.71 partitions / sec (100 runs, 1 concurrent ops) with compacting reader (e.g. TWCS + interposer): 221376.19 +- 1282.11 partitions / sec (50 runs, 1 concurrent ops) 216611.65 +- 985.44 partitions / sec (100 runs, 1 concurrent ops) 215975.51 +- 930.79 partitions / sec (100 runs, 1 concurrent ops) So the cost of compacting data across buckets is ~3.5%, which happens only with interposer enabled and GC writer disabled. Fixes #9662. Signed-off-by: Raphael S. Carvalho <raphaelsc@scylladb.com>	2021-12-10 17:14:44 -03:00
Pavel Emelyanov	b0a8c153f7	select_statement: Remove unused proxy args and captures The generate_view_paging_state_from_base_query_results() has unused proxy argument that's carried over quite a long stack for nothing. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com> Message-Id: <20211210175203.26197-1-xemul@scylladb.com>	2021-12-10 20:39:55 +02:00
Raphael S. Carvalho	484269cd8f	compaction: make enable_garbage_collected_sstable_writer() more precise we only want to enable GC writer if incremental compaction is required. let's make it more precise by checking that size limit for sstable isn't disabled, so GC writer will only be enabled for compaction strategies that really need it. So strategies that don't need it won't pay the penalty. Signed-off-by: Raphael S. Carvalho <raphaelsc@scylladb.com>	2021-12-10 15:22:08 -03:00
Botond Dénes	04306d762f	tools/scylla-sstables: remove unused variables and captures Signed-off-by: Botond Dénes <bdenes@scylladb.com> Message-Id: <20211210142949.527545-1-bdenes@scylladb.com>	2021-12-10 18:24:08 +03:00
Juliusz Stasiewicz	351f142791	cdc/check_and_repair_cdc_streams: ignore LEFT endpoints When `check_and_repair_cdc_streams` encountered a node with status LEFT, Scylla would throw. This behavior is fixed so that LEFT nodes are simply ignored. Fixes #9771 Closes #9778	2021-12-10 15:28:14 +01:00
Raphael S. Carvalho	e0758fded1	compaction_manager: make get_compaction_state() private internal method that should never be directly used by the outside world. Signed-off-by: Raphael S. Carvalho <raphaelsc@scylladb.com> Message-Id: <20211210120806.19233-1-raphaelsc@scylladb.com>	2021-12-10 17:19:24 +03:00
Tomasz Grabiec	4d302dfa1a	Merge "Fix exception safety of rows insertion" from Pavel Emelyanov There are several places that (still) use throwing b-tree .insert_before() method and don't manage the inserted object lifetime. Some of those places also leave the leaked rows_entry on the LRU delaying the assertion failure by the time those entries get evicted (#9728) To prevent such surprises in the future, the set removes the non-safe inserters from the B-tree code. Actually most of this set is that removal plus preparations for reviewability. * xemul/br-rows-insertion-exception-safety-2: btree: Earnestly discourage from insertion of plain references row-cache: Handle exception (un)safety of rows_entry insertion partition_snapshot_row_cursor: Shuffle ensure_result creation mutation_partition: Use B-tree insertion sugar tests: Make B-tree tests use unique-ptrs for insertion	2021-12-10 13:55:18 +01:00
Pavel Emelyanov	6b4b170025	btree: Earnestly discourage from insertion of plain references Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-12-10 12:35:12 +03:00
Pavel Emelyanov	ee103636ac	row-cache: Handle exception (un)safety of rows_entry insertion The B-tree's insert_before() is throwing operation, its caller must account for that. When the rows_entry's collection was switched on B-tree all the risky places were fixed by `ee9e1045`, but few places went under the radar. In the cache_flat_mutation_reader there's a place where a C-pointer is inserted into the tree, thus potentially leaking the entry. In the partition_snapshot_row_cursor there are two places that not only leak the entry, but also leave it in the LRU list. The latter it quite nasty, because those entry can be evicted, eviction code tries to get rows_entry iterator from "this", but the hook happens to be unattached (because insertion threw) and fails the assert. fixes: #9728 Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-12-10 12:35:12 +03:00
Pavel Emelyanov	9fd8db318d	partition_snapshot_row_cursor: Shuffle ensure_result creation Both places get the C-pointer on the freshly allocated rows_entry, insert it where needed and return back the dereferenced pointer. The C-pointer is going to become smart-pointer that would go out of scope before return. This change prepares for that by constructing the ensure_result from the iterator, that's returned from insertion of the entry. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-12-10 12:35:12 +03:00
Pavel Emelyanov	e03f7191d9	mutation_partition: Use B-tree insertion sugar The B-tree insertion methods accept smart pointers and automatically release the ownership after exception-risky part is passed. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-12-10 12:35:12 +03:00
Pavel Emelyanov	5a405a4273	tests: Make B-tree tests use unique-ptrs for insertion The non-smart-pointers overloads are going away, prepare tests for that. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-12-10 12:35:12 +03:00
Nadav Har'El	03d67440ef	alternator: test additional metrics and fix another broken counter In issue #9406 we noticed that a counter for BatchGetItem operations was missing. When we fixed it, we added a test which checked this counter - but only this counter. It was left as a TODO to test the rest of the Alternator metrics, and this is what this patch does. Here we add a comprehensive test for all of the operations supported by Scylla and how they increase the appropriate operation counter. With this test we discovered a new bug: the DescribeTimeToLive operation incremented the UpdateTimeToLiveCounter :-( So in this patch we also include a fix for that bug, and the new test verifies that it is fixed. In addition to the operation counters, Alternator also has additional metric and we also added tests for some of them - but not all. The remaining untested metrics are listed in a TODO comment. Message-Id: <20211206154727.1170112-1-nyh@scylladb.com>	2021-12-10 08:08:54 +02:00
Benny Halevy	cca956bce2	database_test: snapshot_with_quarantine_works: get the list of sstables from table object Rather than the filesystem, to reduce flakiness. Also, add some test logging. Fixes #9763 Test: database_test(debug, release) Signed-off-by: Benny Halevy <bhalevy@scylladb.com> Message-Id: <20211209175144.854896-1-bhalevy@scylladb.com>	2021-12-09 21:01:25 +02:00
Nadav Har'El	006fa588a3	alternator ttl: correct misleading typo in error message Alternator's support for the DynamoDB API TTL features is experimental, so if a user attempts to use one the TTL API requests, an error message is returned that the experimental feature must be turned on first. The message incorrectly said that the name of the experimental flag to turn on is "alternator_ttl", with an underscore. But that's a type - it should be "alternator-ttl" with a hyphen. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20211209183428.1336526-1-nyh@scylladb.com>	2021-12-09 20:47:05 +02:00
Benny Halevy	8728fd480d	database_test: do_with_some_data: get the return func future do_with_some_data runs a function in a seastar thread. It needs to get() the future func returns rather than propagating it. This solves a secondary failure due to abandoned future when the test case fails, as seen in https://jenkins.scylladb.com/view/master/job/scylla-master/job/next/4254/artifact/testlog/x86_64_debug/database_test.snapshot_with_quarantine_works.381.log ``` test/boost/database_test.cc(903): fatal error: in "snapshot_with_quarantine_works": critical check expected.empty() has failed WARN 2021-12-08 00:35:16,300 [shard 0] seastar - Exceptional future ignored: boost::execution_aborted, backtrace: 0x10935e50 0x16ff2d8d 0x16ff2a4d 0x16ff5033 0x16ff5ec2 0x162d4ce9 0x10a2bdb5 0x10a2bd24 0x10a54ca4 0x10a27cf3 0x10a22151 0x10a67c9d 0x10a67a78 0x163ac37e 0x163b29e9 0x163b7690 0x163b51c1 0x17c212df 0x17c1f097 0x17bf8b4c 0x17bf83f2 0x17bf82a2 0x17bf7d52 0x10f8bf5a 0x166db84b /lib64/libpthread.so.0+0x9298 /lib64/libc.so.6+0x100352 ... *** 1 abandoned failed future(s) detected Failing the test because fail was requested by --fail-on-abandoned-failed-futures ``` Signed-off-by: Benny Halevy <bhalevy@scylladb.com> Message-Id: <20211209174512.851945-1-bhalevy@scylladb.com>	2021-12-09 21:11:56 +03:00
Nadav Har'El	c6f2afb93d	Merge 'cql3: Allow to skip EQ restricted columns in ORDER BY' from Jan Ciołek In queries like: ```cql SELECT * FROM t WHERE p = 0 AND c1 = 0 ORDER BY (c1 ASC, c2 ASC) ``` we can skip the requirement to specify ordering for `c1` column. The `c1` column is restricted by an `EQ` restriction, so it can have at most one value anyway, there is no need to sort. This commit makes it possible to write just: ```cql SELECT * FROM t WHERE p = 0 AND c1 = 0 ORDER BY (c2 ASC) ``` I reorganized the ordering code, I feel that it's now clearer and easier to understand. It's possible to only introduce a small change to the existing code, but I feel like it becomes a bit too messy. I tried it out on the [`orderby_disorder_small`](https://github.com/cvybhu/scylla/commits/orderby_disorder_small) branch. The diff is a bit messy because I moved all ordering functions to one place, it's better to read [select_statement.cc](https://github.com/cvybhu/scylla/blob/orderby_disorder/cql3/statements/select_statement.cc#L1495-L1658) lines 1495-1658 directly. In the new code it would also be trivial to allow specifying columns in any order, we would just have to sort them. For now I commented out the code needed to do that, because the point of this PR was to fix #2247. Allowing this would require some more work changing the existing tests. Fixes: #2247 Closes #9518 * github.com:scylladb/scylla: cql-pytest: Enable test for skipping eq restricted columns in order by cql3: Allow to skip EQ restricted columns in ORDER BY cql3: Add has_eq_restriction_on_column function cql3: Reorganize orderings code	2021-12-09 21:11:56 +03:00
Nadav Har'El	36c3b92b19	alternator, schema_loader: get rid of deprecation warnings Seastar moved the read_entire_stream(), read_entire_stream_contiguous() and skip_entire_stream() from the "httpd" namespace to the "util" namespace. Using them with their old names causes deprecation warnings when compiling alternator/server.cc. This patch fixes the namespace (and adds the new include) to get rid of the deprecation warnings. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20211209132759.1319420-1-nyh@scylladb.com>	2021-12-09 21:11:56 +03:00
Avi Kivity	242e19195f	Merge "table: Prevent resurrecting data from memtable on compaction" from Mikołaj " Mutations are not guaranteed to come in the order of their timestamps. If there is an expired tombstone in the sstable and a repair inserts old data into memtable, the compaction would not consider memtable data and purge the tombstone leading to data resurrection. The solution is to disallow purging tombstones newer than min memtable timestamp. If there are no memtables, max timestamp is used. " * 'check-memtable-at-compact-tombstone-discard/v2' of github.com:mikolajsieluzycki/scylla: table: Prevent resurrecting data from memtable on compaction table: Add min_memtable_timestamp function to table	2021-12-09 21:11:56 +03:00
Piotr Sarna	2ec36a6c53	alternator,ttl: limit parallelism to 1 page Right now we do not really have any parallelism in the alternator TTL service, but in order to be future-proof, a semaphore is instantiated to ensure that we only handle 1 page of a scan at a time, regardless of how many tables are served. This commit also removes the FIXME regarding the service permit - using an empty permit is a conscious decision, because the parallelism is limited by other means (see above). Tests: unit(release) Message-Id: <b5f0c94f1afbead1f940a210911cc05f70900dcd.1638990637.git.sarna@scylladb.com>	2021-12-09 21:11:55 +03:00
Asias He	9859c76de1	storage_service: Wait for seastar::get_units in node_ops The seastar::get_units returns a future, we have to wait for it. Fixes #9767 Closes #9768	2021-12-09 21:11:55 +03:00
Jan Ciolek	13d367dada	cql-pytest: Enable test for skipping eq restricted columns in order by This test was marked as xfail, but now the functionality it tests has been implemented. In my opinion the expected error message makes no sense, the message was: "Order by currently only supports the ordering of columns following their declared order in the PRIMARY KEY" In cases where there was missing restriction on one column. This has been changed to: "Unsupported order by relation - column {} doesn't have an ordering or EQ relation." Because of that I had to modify the test to accept messages from both Scylla and Cassandra. The expected error message pattern is now "rder by", because that's the largest common part. Signed-off-by: Jan Ciolek <jan.ciolek@scylladb.com>	2021-12-09 14:59:47 +01:00
Mikołaj Sielużycki	504efe0607	table: Prevent resurrecting data from memtable on compaction Mutations are not guaranteed to come in the order of their timestamps. If there is an expired tombstone in the sstable and a repair inserts old data into memtable, the compaction would not consider memtable data and purge the tombstone leading to data resurrection. The solution is to disallow purging tombstones newer than min memtable timestamp.	2021-12-09 13:22:14 +01:00
Mikołaj Sielużycki	7ce0ca040d	table: Add min_memtable_timestamp function to table	2021-12-09 13:14:38 +01:00
Jan Ciolek	a548c2dac4	cql3: Allow to skip EQ restricted columns in ORDER BY In queries like: SELECT * FROM t WHERE p = 0 AND c1 = 0 ORDER BY (c1 ASC, c2 ASC) we can skip the requirement to specify ordering for c1 column. The c1 column is restricted by an EQ restriction, so it can have only one value anyway, there is no need to sort. This commit makes it possible to write just: SELECT * FROM t WHERE p = 0 AND c1 = 0 ORDER BY (c2 ASC) Fixes: #2247 Signed-off-by: Jan Ciolek <jan.ciolek@scylladb.com>	2021-12-09 12:07:02 +01:00
Jan Ciolek	7bbfa48bc5	cql3: Add has_eq_restriction_on_column function Adds a function that checks whether a given expression has eq restrction on the specified column. It finds restrictions like col = ... or (col, col2) = ... IN restrictions don't count, they aren't EQ restrictions Signed-off-by: Jan Ciolek <jan.ciolek@scylladb.com>	2021-12-09 12:06:43 +01:00
Jan Ciolek	f76a1cd4bf	cql3: Reorganize orderings code Reorganized the code that handles column ordering (ASC or DESC). I feel that it's now clearer and easier to understand. Added an enum that describes column ordering. It has two possible values: ascending or descending. It used to be a bool that was sometimes called 'reversed', which could mean multiple things. Instead of column.type->is_reversed() != <ordering bool> there is now a function called are_column_select_results_reversed. Split checking if ordering is reversed and verifying whether it's correct into two functions. Before all of this was done by is_reversed() This is a preparation to later allow skipping ORDER BY restrictions on some columns. Adding this to the existing code caused it to get quite complex, but this new version is better suited for the task. The diff is a bit messy because I moved all ordering functions to one place, it's better to read select_statement.cc lines 1495-1651 directly. Signed-off-by: Jan Ciolek <jan.ciolek@scylladb.com>	2021-12-09 12:06:42 +01:00
Nadav Har'El	f9673309aa	docs: protocols.md - add information on Redis listening address The description in protocols.md of the Redis protocol server in Scylla explains how its port can be configured, but not how the listening IP address can be configured. It turns out that the same "rpc_address" that controls CQL's and Thrift's IP address also applies to Redis. So let's document that. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20211208160206.1290916-1-nyh@scylladb.com>	2021-12-08 20:14:52 +01:00
Nadav Har'El	e032f92c5c	Merge 'api/storage service: validate table names' from Benny Halevy This series fixes a couple issues around generating and handling of no_such_keyspace and no_such_column_family exceptions. First, it removes std::throw_with_nested around their throw sites in the respective database::find_* functions. Fixes #9753 And then, it introduces a `validate_tables` helper in api/storage_service.cc that generates a `bad_param_exception` in order to set the correct http response status if a non-existing table name is provided in the `cf` http request parameter. Fixes #9754 The series also adds a test for the REST API under test/rest_api that verifies the storage_service enable/disable auto_compaction api and checks the error codes for non-existing keyspace or table. Test: unit(dev) Closes #9755 * github.com:scylladb/scylla: api: storage_service: add parse_tables database: un-nest no_such_keyspace and no_such_column_family exceptions database: throw internal error when failing uuid returned by find_uuid database: find_uuid: throw no_such_column_family exception if ks/cf were not found test: rest_api: add storage_service test test: add basic rest api test test: cql-pytest: wait for rest api when starting scylla	2021-12-08 16:54:48 +02:00
Benny Halevy	ff63ad9f6e	api: storage_service: add parse_tables Splits and validate the cf parameter, containing an optional comma-separated list of table names. If any table is not found and a no_such_column_family exception is thrown, wrap it in a `bad_param_exception` so it will translate to `reply::status_type::bad_request` rather than `reply::status_type::internal_server_error`. With that, hide the split_cf function from api/api.hh since it was used only from api/storage_service and new use sites should use validate_tables instead. Fixes #9754 Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-12-08 16:42:40 +02:00
Benny Halevy	a3bd7806e7	database: un-nest no_such_keyspace and no_such_column_family exceptions These were thrown in the respective database::find_* function as nested exception since `d3fe0c5182`. Wrapping them in nested exceptions just makes it harder to figure out and work with and apprently serves no purpose. Without these nested_exception we can correctly detect internal errors when synchronously failing to find a uuid returned by find_uuid(ks, cf). Fixes #9753 Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-12-08 16:35:38 +02:00
Benny Halevy	ac49e5fff1	database: throw internal error when failing uuid returned by find_uuid find_uuid returns a uuid found for ks_name.table_name. In some cases, we immediately and synchronously use that uuid to lookup other information like the table& or the schema. Failing to find that uuid indicates an internal error when no preemption is possible. Note that yielding could allow deletion of the table to sneak in and invalidate the uuis. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-12-08 16:35:38 +02:00
Benny Halevy	8cbecb1c21	database: find_uuid: throw no_such_column_family exception if ks/cf were not found Rather than masquerading all errors as std::out_of_range("") convert only the std::out_of_range error from _ks_cf_to_uuid.at() to no_such_column_family(ks, cf). That relieves all callers of fund_uuid from doing that conversion themselves. For example, get_uuid in api/column_family now only deals with converting no_such_column_family to bad_param_exception, as it needs to do at the api level, rather than generating a similar error from scratch. Other call sites required no intervention. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-12-08 16:35:38 +02:00
Benny Halevy	5eb32aa57c	test: rest_api: add storage_service test FIXME: negative tests for not-found tables should result in a requests.codes.bad_request but currently result in requests.codes.internal_server_error. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-12-08 16:35:36 +02:00
Piotr Sarna	d486b496a6	alternator,ttl: start scans from a random token range This patch addresses yet another FIXME from alternator/ttl.cc. Namely, scans are now started from a random, owned token range instead of always starting with the first range. This mechanism is expected to reduce the probability of some ranges being starved when the scanning process is often restarted, e.g. due to nodes failing. Should the mechanism prove insufficient for some users, a more complete solution is to regularly persist the state of the scanning process in a table (distributed if we want to allow other nodes to pick up from where a dead node left off), but that induces overhead. Tests: unit(release) (including a long loop over the ttl pytest) Message-Id: <7fc3f6525ceb69725c41de10d0fb6b16188349e3.1638387924.git.sarna@scylladb.com> Message-Id: <db198e743ca9ed1e5cc659e73da342fbce2c882a.1638473143.git.sarna@scylladb.com>	2021-12-08 16:15:53 +02:00
Benny Halevy	26257cfa6d	test: add basic rest api test Test system/uptime_ms to start with. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-12-08 16:05:33 +02:00
Benny Halevy	01f2e8b391	test: cql-pytest: wait for rest api when starting scylla Some of the tests, like nodetool.py, use the scylla REST API. Add a check_rest_api function that queries http://<node_addr>:10000/ that is served once scylla starts listening on the API port and call it via run.wait_for_services. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-12-08 16:05:32 +02:00
Piotr Sarna	26288c1a86	test,alternator: make TTL tests less prone to false negatives On my local machine, a 3 second deadline proved to cause flakiness of test_ttl_expiration case, because its execution time is just around 3 seconds. This patch addresse the problem by bumping the local timeout to 10 (and 15 for test_ttl_expiration_long, since it's dangerously near the 10 second deadline on my machine as well). Moreover, some test cases short-circuited once they detected that all needed items expired, but other ones lacked it and always used their full time slot. Since 10 seconds is a little too long for a single test case, even one marked with --veryslow, this patch also adds a couple of other short-circuits. One exception is test_ttl_expiration_hash_wrong_type, which actually depends on the fact that we should wait for the whole loop to finish. Since this case was never flaky for me with the 3 second timeout, it's left as is. Theoretically, test_ttl_expiration also kind of depends on checking the condition more than once (because the TTL of one of the values is bumped on each iteration), but empirical evidence shows that multiple iterations always occur in this test case anyway - for me, it always spinned at least 3 times. Tests: unit(release) Message-Id: <a0a479929dac37daace744e0a970567a8aa3b518.1638431933.git.sarna@scylladb.com>	2021-12-08 16:02:45 +02:00
Raphael S. Carvalho	c3c23dd1e5	multishard_mutation_query: make multi_range_reader::fill_buffer() work even after EOS if fill_buffer() is called after EOS, underlying reader will be fast forwarded to a range pointed to by an invalid iterator, so producing incorrect results. fill_buffer() is changed to return early if EOS was found, meaning that underlying reader already fast forwarded to all ranges managed by multi_range_reader. Usually, consume facilities check for EOS, before calling fill_buffer() but most reader impl check for EOS to avoid correctness issues. Let's do the same here. Signed-off-by: Raphael S. Carvalho <raphaelsc@scylladb.com> Message-Id: <20211208131423.31612-1-raphaelsc@scylladb.com>	2021-12-08 15:39:11 +02:00
Avi Kivity	f28552016f	Update seastar submodule * seastar f8a038a0a2...8d15e8e67a (21): > core/program_options: preserve defaultness of CLI arguments > log: Silence logger when logging > Include the core/loop.hh header inside when_all.hh header > http: Fix deprecated wrappers > foreign_ptr: Add concept > util: file: add read_entire_file > short_streams: move to util > Revert "Merge: file: util: add read_entire_file utilities" > foreign_ptr: declare destroy as a static method > Merge: file: util: add read_entire_file utilities > Merge "output_stream: handle close failure" from Benny > net: bring local_address() to seastar::connected_socket. > Merge "Allow programatically configuring seastar" from Botond > Merge 'core: clean up memory metric definitions' from John Spray > Add PopOS to debian list in install-dependencies.sh > Merge "make shared_mutex functions exception safe and noexcept" from Benny > on_internal_error: set_abort_on_internal_error: return current state > Implementation of iterator-range version of when_any > net: mark functions returning ethernet_address noexcept > net: ethernet_address: mark functions noexcept > shared_mutex: mark wake and unlock methods noexcept Contains patch from Botond Dénes <bdenes@scylladb.com>: db/config: configure logging based on app_template::seastar_options Scylla has its own config file which supports configuring aspects of logging, in addition to the built-in CLI logging options. When applying this configuration, the CLI provided option values have priority over the ones coming from the option file. To implement this scylla currently reads CLI options belonging to seastar from the boost program options variable map. The internal representation of CLI options however do not constitute an API of seastar and are thus subject to change (even if unlikely). This patch moves away from this practice and uses the new shiny C++ api: `app_template::seastar_options` to obtain the current logging options.	2021-12-08 14:21:11 +02:00
Tomasz Grabiec	5eaca85e4b	Merge "wire up schema raft state machine" from Gleb This series wires up the schema state machine to process raft commands and transfer snapshots. The series assumes that raft group zero is used for schema transfer only and that single raft command contains single schema change in a form of canonical_mutation array. Both assumptions may change in which case the code will be changed accordingly, but we need to start somewhere. * scylla-dev/gleb/schema-raft-sm-v2: schema raft sm: request schema sync on schema_state_machine snapshot transfer raft service: delegate snapshot transfer to a state machine implementation schema raft sm: pass migration manager to schema_raft_state_machine and merge schema on apply()	2021-12-08 13:14:28 +01:00
Nadav Har'El	92e7fbe657	test/alternator: check correct error for unknown operation Add a short test verifying that Alternator responds with the correct error code (UnknownOperationException) when receiving an unknown or unsupported operation. The test passes on both AWS and Alternator, confirming that the behavior is the same. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20211206125710.1153008-1-nyh@scylladb.com>	2021-12-08 13:56:38 +02:00
Gleb Natapov	f25424edcd	storage_service: remove unused function. is_auto_bootstrap() function is no longer used. Message-Id: <YbCVXPI4hE8wgT4T@scylladb.com>	2021-12-08 13:55:32 +02:00
Botond Dénes	0aa4e5e726	test/cql-pytest: mv virtual_tables.py -> test_virtual_tables.py For consistency with the other tests. Signed-off-by: Botond Dénes <bdenes@scylladb.com> Message-Id: <20211208102108.126492-1-bdenes@scylladb.com>	2021-12-08 12:23:22 +02:00
Gavin Howell	c6e0a807b4	Update wasm.md Grammar correction, sentence re-write. Closes #9760	2021-12-08 10:24:53 +01:00
Tomasz Grabiec	2a36377bb3	Merge "test: raft: randomized_nemesis_test: introduce server stop/crash nemesis" from Kamil We begin by preparing the `persistence` class so that the storage can be reused across different Raft server instances: the test keeps a shared pointer to the storage so that when a server stops, a new server with the same ID can be reconstructed with this storage. We then modify `environment` so that server instances can be removed and replaced in middle of operations. Finally we prepare a nemesis operation which gracefully stops or immediately crashes a randomly picked server and run this operation periodically in `basic_generator_test`. One important change that changes the API of `raft::server` is included: the metrics are not automatically registered in `start()`. This is because metric registration modifies global data structures, which cannot be done twice with the same set of metrics (and we would do it when we restart a server with the same ID). Instead, `register_metrics()` is exposed in the `raft::server` interface to be called when running servers in production. * kbr/crashes-v3: raft: server: print the ID of aborted server test: raft: randomized_nemesis_test: run stop_crash nemesis in `basic_generator_test` test: raft: randomized_nemesis_test: introduce `stop_crash` operation test: raft: randomized_nemesis_test: environment: implement server `stop` and `crash` raft: server: don't register metrics in `start()` test: raft: randomized_nemesis_test: raft_server: return `stopped_error` when called during abort test: raft: randomized_nemesis_test: handle `raft::stopped_error` test: raft: randomized_nemesis_test: handle missing servers in `environment` call functions test: raft: randomized_nemesis_test: environment: split `new_server` into `new_node` and `start_server` test: raft: randomized_nemesis_test: remove `environment::get_server` test: raft: randomized_nemesis_test: construct `persistence_proxy` outside `raft_server<M>::create` test: raft: randomized_nemesis_test: persistence_proxy: store a shared pointer to `persistence` test: raft: randomized_nemesis_test: persistence: split into two classes test: raft: logical_timer: introduce `sleep_until`	2021-12-07 22:16:23 +01:00

1 2 3 4 5 ...

29355 Commits