scylladb

mirror of https://github.com/scylladb/scylladb.git synced 2026-06-01 12:36:56 +00:00

Author	SHA1	Message	Date
Pavel Solodovnikov	7c229998e8	raft: unit-tests for `raft_address_map` Signed-off-by: Pavel Solodovnikov <pa.solodovnikov@scylladb.com>	2021-03-26 20:22:44 +03:00
Tomasz Grabiec	ef06a939c4	Merge "raft: seven etcd unit tests ported" from Alejo Seven etcd unit tests as boost tests. * alejo/raft-tests-etcd-08-v4-communicate-v5: raft: etcd unit tests: test proposal handling scenarios raft: etcd unit tests: test old messages ignored raft: etcd unit tests: test single node precandidate raft: etcd unit tests: test dueling precandidates raft: etcd unit tests: test dueling candidates raft: etcd unit tests: test cannot commit without new term raft: etcd unit tests: test single node commit raft: etcd unit tests: update test_leader_election_overwrite_newer_logs raft: etcd unit tests: fix test_progress_leader raft: testing: log comparison helper functions raft: testing: helper to make fsm candidate raft: testing: expose log for test verification raft: testing: use server_address_set raft: testing: add prevote configuration raft: testing: make become_follower() available for tests	2021-03-25 20:27:07 +01:00
Alejo Sanchez	ace0ee514f	raft: etcd unit tests: test proposal handling scenarios TestProposal For multiple scenarios, check proposal handling. Note, instead of expecting an explicit result for each specified case, the test automatically checks for expected behavior when quorum is reached or not. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:29 -04:00
Alejo Sanchez	77163ea76a	raft: etcd unit tests: test old messages ignored TestOldMessages Checks an append request from a leader from a previous term is ignored. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:29 -04:00
Alejo Sanchez	bf65b19803	raft: etcd unit tests: test single node precandidate TestSingleNodePreCandidate Checks a single node configuration with precandidate on works to automatically elect the node. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:29 -04:00
Alejo Sanchez	de7051467b	raft: etcd unit tests: test dueling precandidates TestDuelingPreCandidates In a configuration of 3 nodes, two nodes don't see each other and they compete for leadership. Loser (3) should revert to follower when prevote is rejected and revert to term 1. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:29 -04:00
Alejo Sanchez	aa7d23f86b	raft: etcd unit tests: test dueling candidates TestDuelingCandidates In a configuration of 3 nodes, two nodes don't see each other and they compete for leadership. Once reconnected, loser should not disrupt. But note it will remain candidate with current algorithm without prevoting and other fsms will not bump term. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:29 -04:00
Alejo Sanchez	1eac94e7d6	raft: etcd unit tests: test cannot commit without new term TestCannotCommitWithoutNewTermEntry tests the entries cannot be committed when leader changes, no new proposal comes in and ChangeTerm proposal is filtered. NOTE: this doesn't check committed but it's implicit for next round; this could also use communicate() providing committed output map Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:29 -04:00
Alejo Sanchez	b421fe3605	raft: etcd unit tests: test single node commit Port etcd TestSingleNodeCommit In a single node configuration elect the node, add 2 entries and check number of committed entries. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:29 -04:00
Alejo Sanchez	9b4538476b	raft: etcd unit tests: update test_leader_election_overwrite_newer_logs Make test_leader_election_overwrite_newer_logs use newer communicate() and other new helpers. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:29 -04:00
Alejo Sanchez	368eec1190	raft: etcd unit tests: fix test_progress_leader Make implementation follow closer to original test. Use newer boost test helpers. NOTE: in etcd it seems a leader's self progress is in PIPELINE state. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:28 -04:00
Alejo Sanchez	ba29970e29	raft: testing: log comparison helper functions Two helper functions to compare logs. For now only index, term, and data type are used. Data content comparison does not seem to be necessary for now. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:28 -04:00
Alejo Sanchez	aeab4cf4a9	raft: testing: helper to make fsm candidate Current election_timeout() helper might bump the term twice. It's convenient and less error prone to have a more fine grained helper that stops right when candidate state is reached. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:04:19 -04:00
Alejo Sanchez	7a6616f1cb	raft: testing: expose log for test verification Let derived classes access the log to verify its contents. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:03:46 -04:00
Alejo Sanchez	05b1f57e67	raft: testing: use server_address_set Use server_address_set in local namespace for brevity. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:01:12 -04:00
Alejo Sanchez	9d0a7d8ccf	raft: testing: add prevote configuration Provide a generic prevote configuration for tests. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-25 15:00:28 -04:00
Dejan Mircevski	b2a04985f7	cql-pytest: Drop needless INSERT in test_null One INSERT statement was unnecessary for the test, so delete it. Another was necessary, so explain it. Tests: cql-pytest/test_null on both Scylla and Cassandra Signed-off-by: Dejan Mircevski <dejan@scylladb.com> Closes #8304	2021-03-25 16:37:00 +01:00
Tomasz Grabiec	7b30d31d77	Merge "raft: test configuration changes" from Kostja Test raft configuration changes: a node with empty configuration, transitioning to an entirely different cluster, transitioning in presence of down nodes, leader change during configuration change, stray replies, etc. * scylla-dev/raft-empty-confchange-v5: (21 commits) raft: (testing) stray replies from removed followers raft: always return a non-zero configuration index from the log raft: (testing) leader change during configuration change raft: (testing) test confchange {ABCDE} -> {ABCDEFG} raft: (testing) test confchange {ABCDEF} -> {ABCGH} raft: (testing) test confchange {ABC} -> {CDE} raft: (testing) test confchange {AB} -> {CD} raft: (testing) test confchange {A} -> {B} raft: (testing) test a server with empty configuration raft: (testing) introduce testing utilities raft: (testing) simplify id allocation in test raft: (testing) add select_leader() helper raft: (testing) introduce communicate() helper raft: (testing) style cleanup in raft_fsm_test raft: (testing) fix bug in election_threshold raft: minor style changes & comments raft: do not assert when transitioning to empty config raft: assert we never apply a snapshot over uncommitted entries (leader) raft: improve tracing raft: add fsm_output::empty() helper to aid testing ...	2021-03-25 14:01:09 +01:00
Alejo Sanchez	7e6807e8fc	raft: testing: make become_follower() available for tests Some etcd tests need to force a follower with a specific leader. Signed-off-by: Alejo Sanchez <alejo.sanchez@scylladb.com>	2021-03-24 19:11:09 -04:00
Piotr Wojtczak	c1daf2bb24	column_family: Make toppartitions queries more generic Right now toppartitions can only be invoked on one column family at a time. This change introduces a natural extension to this functionality, allowing to specify a list of families. We provide three ways for filtering in the query parameter "name_list": 1. A specific column family to include in the form "ks:cf" 2. A keyspace, telling the server to include all column families in it. Specified by omitting the cf name, i.e. "ks:" 3. All column families, which is represented by an empty list The list can include any amount of one or both of the 1. and 2. option. Fixes #4520 Closes #7864	2021-03-24 17:54:05 +02:00
Konstantin Osipov	1a1d7ab662	raft: (testing) stray replies from removed followers	2021-03-24 14:05:55 +03:00
Konstantin Osipov	0295163f6f	raft: always return a non-zero configuration index from the log Return snapshot index for last configuration index if there is no configuration in the log.	2021-03-24 14:05:55 +03:00
Konstantin Osipov	cec59e53ef	raft: (testing) leader change during configuration change	2021-03-24 14:05:36 +03:00
Konstantin Osipov	a203c8833f	raft: (testing) test confchange {ABCDE} -> {ABCDEFG}	2021-03-24 14:04:18 +03:00
Konstantin Osipov	40e117d36e	raft: (testing) test confchange {ABCDEF} -> {ABCGH}	2021-03-24 14:04:18 +03:00
Konstantin Osipov	14b2d5d308	raft: (testing) test confchange {ABC} -> {CDE} Test leader change during configuration change.	2021-03-24 14:04:18 +03:00
Konstantin Osipov	3c718a175e	raft: (testing) test confchange {AB} -> {CD}	2021-03-24 14:04:18 +03:00
Konstantin Osipov	2e30c8540e	raft: (testing) test confchange {A} -> {B} Test non-restart and leader restart scenario.	2021-03-24 14:04:18 +03:00
Konstantin Osipov	e23da06fef	raft: (testing) test a server with empty configuration Try becoming a candidate for such server, or adding it to an existing configuration.	2021-03-24 14:04:18 +03:00
Konstantin Osipov	b18599c630	raft: (testing) introduce testing utilities Add a discrete_failure_detector, to be able to mark a single server dead.	2021-03-24 14:04:18 +03:00
Konstantin Osipov	8d26d24370	raft: (testing) simplify id allocation in test	2021-03-24 14:04:18 +03:00
Konstantin Osipov	322a15ec33	raft: (testing) add select_leader() helper With leader stepdown extension, leadership transfer can happen to any follower with long enough log. Add a helper to select that follower from a list.	2021-03-24 14:04:18 +03:00
Konstantin Osipov	4a00da276d	raft: (testing) introduce communicate() helper Allow to communicate between arbitrary number of FSMs. Drop messages to FSMs which are not in the argument list. Stop communication upon predicate.	2021-03-24 14:04:18 +03:00
Konstantin Osipov	7182323ac0	raft: (testing) style cleanup in raft_fsm_test 1) Avoid memory violations on test failure 2) Print better diagnostics on failure (BOOST_CHECK_EQUAL vs BOOST_CHECK)	2021-03-24 14:04:18 +03:00
Konstantin Osipov	f0f25bf7fb	raft: (testing) fix bug in election_threshold election_threshold was ticking one extra tick, causing the follower to become candidate in some cases. This was rendering tests unstable.	2021-03-24 14:04:18 +03:00
Piotr Sarna	23057dd186	Merge 'Implement RAFT's leader stepdown extension' from Gleb This series implements leader stepdown extension. See patch 4 for justification for its existence. First three patches either implement cleanups to existing code that future patch will touch or fix bugs that need to be fixed in order for stepdown test to work. * 'raft-leader-stepdown-v3' of github.com:scylladb/scylla-dev: raft: add test for leader stepdown raft: introduce leader stepdown procedure raft: fix replication when leader is not part of current config raft: do not update last election time if current leader is not a part of current configuration raft: move log limiting semaphore into the leader state	2021-03-22 09:45:19 +01:00
Avi Kivity	3c44445c07	Merge "Introduce off-strategy compaction for repair-based bootstrap and replace" from Raphael " Scylla suffers with aggressive compaction after repair-based operation has initiated. That translates into bad latency and slowness for the operation itself. This aggressiveness comes from the fact that: 1) new sstables are immediately added to the compaction backlog, so reducing bandwidth available for the operation. 2) new sstables are in bad shape when integrated into the main sstable set, not conforming to the strategy invariant. To solve this problem, new sstables will be incrementally reshaped, off the compaction strategy, until finally integrated into the main set. The solution takes advantage there's only one sstable per vnode range, meaning sstables generated by repair-based operations are disjoint. NOTE: off-strategy for repair-based decommission and removenode will follow this series and require little work as the infrastructure is introduced in this series. Refs #5226. " * 'offstrategy_v7' of github.com:raphaelsc/scylla: tests: Add unit test for off-strategy sstable compaction table: Wire up off-strategy compaction on repair-based bootstrap and replace table: extend add_sstable_and_update_cache() for off-strategy sstables/compaction_manager: Add function to submit off-strategy work table: Introduce off-strategy compaction on maintenance sstable set table: change build_new_sstable_list() to accept other sstable sets table: change non_staging_sstables() to filter out off-strategy sstables table: Introduce maintenance sstable set table: Wire compound sstable set table: prepare make_reader_excluding_sstables() to work with compound sstable set table: prepare discard_sstables() to work with compound sstable set table: extract add_sstable() common code into a function sstable_set: Introduce compound sstable set reshape: STCS: preserve token contiguity when reshaping disjoint sstables	2021-03-22 10:43:13 +02:00
Gleb Natapov	272cb1c1e6	raft: add test for leader stepdown	2021-03-22 10:31:16 +02:00
Gleb Natapov	9d6bf7f351	raft: introduce leader stepdown procedure Section 3.10 of the PhD describes two cases for which the extension can be helpful: 1. Sometimes the leader must step down. For example, it may need to reboot for maintenance, or it may be removed from the cluster. When it steps down, the cluster will be idle for an election timeout until another server times out and wins an election. This brief unavailability can be avoided by having the leader transfer its leadership to another server before it steps down. 2. In some cases, one or more servers may be more suitable to lead the cluster than others. For example, a server with high load would not make a good leader, or in a WAN deployment, servers in a primary datacenter may be preferred in order to minimize the latency between clients and the leader. Other consensus algorithms may be able to accommodate these preferences during leader election, but Raft needs a server with a sufficiently up-to-date log to become leader, which might not be the most preferred one. Instead, a leader in Raft can periodically check to see whether one of its available followers would be more suitable, and if so, transfer its leadership to that server. (If only human leaders were so graceful.) The patch here implements the extension and employs it automatically when a leader removes itself from a cluster.	2021-03-22 10:28:43 +02:00
Benny Halevy	f562c9c2f3	test: sstable_datafile_test: tombstone_purge_test: use a longer ttl As seen in next-3319 unit testing on jenkins The cell ttl may expire during the test (presuming that the test machine was overloaded), leading to: ``` INFO 2021-03-21 10:05:23,048 [shard 0] compaction - [Compact tests.tombstone_purge 2fcaf680-8a1c-11eb-b1b9-97020c5d261e] Compacting [/jenkins/workspace/scylla-master/next/scylla/testlog/release/scylla-af8644ec-7f07-4ffe-80bf-6703a942e435/la-17-big-Data.db:level=0:origin=, ] INFO 2021-03-21 10:05:23,048 [shard 0] compaction - [Compact tests.tombstone_purge 2fcaf680-8a1c-11eb-b1b9-97020c5d261e] Compacted 1 sstables to []. 4kB to 0 bytes (~0% of original) in 0ms = 0 bytes/s. ~128 total partitions merged to 0. ./test/lib/mutation_assertions.hh(108): fatal error: in "tombstone_purge_test": Mutations differ, expected {table: 'tests.tombstone_purge', key: {'id': alpha, token: -7531858254489963}, mutation_partition: { rows: [ { cont: true, dummy: false, position: { bound_weight: 0, }, 'value': { atomic_cell{1,ts=1616313953,expiry=1616313958,ttl=5} }, }, ] } } ...but got: {table: 'tests.tombstone_purge', key: {'id': alpha, token: -7531858254489963}, mutation_partition: { rows: [ { cont: true, dummy: false, position: { bound_weight: 0, }, 'value': { atomic_cell{DEAD,ts=1616313953,deletion_time=1616313953} }, }, ] } } ``` This corresponds to: ``` 2395 auto mut2 = make_expiring(alpha, ttl); 2396 auto mut3 = make_insert(beta); ... 2399 auto sst2 = make_sstable_containing(sst_gen, {mut2, mut3}); ``` Extend (logical) ttl to 10 seconds to reduce flakiness due to real-time timing. Test: sstable_datafile_test(dev) Signed-off-by: Benny Halevy <bhalevy@scylladb.com> Message-Id: <20210321142931.1226850-1-bhalevy@scylladb.com>	2021-03-21 16:42:00 +02:00
Avi Kivity	1e820687eb	Merge "reader_concurrency_semaphore: limit non-admitted inactive reads" from Botond " Due to bad interaction of recent changes (`913d970` and `4c8ab10`) inctive readers that are not admitted have managed to completely fly under the radar, avoiding any sort of limitation. The reason is that pre-admission the permits don't forward their resource cost to the semaphore, to prevent them possibly blocking their own admission later. However this meant that if such a reader is registered as inactive, it completely avoids the normal resource based eviction mechanism and can accumulate without bounds. The real solution to this is to move the semaphore before the cache and make all reads pass admission before they get started (#4758). Although work has been started towards this, it is still a while until it lands. In the meanwhile this patchset provides a workaround in the form of a new inactive state, which -- like admitted -- causes the permit to forward its cost to the semaphore, making sure these un-admitted inactive reads are accounted for and evicted if there is too much of them. Fixes: #8258 Tests: unit(release), dtest(oppartitions_test.py:TestTopPartitions.test_read_by_gause_key_distribution_for_compound_primary_key_and_large_rows_number) " * 'reader-concurrency-semaphore-limit-inactive-reads/v4' of https://github.com/denesb/scylla: test: mutation_reader_test: add test for permit cleanup test: querier_cache_test: add memory based cache eviction test reader_permit: add inactive state querier: insert(): account immediately evicted querier as resource based eviction reader_concurrency_semaphore: fix clear_inactive_reads() reader_concurrency_semaphore: make inactive_read_handle a weak reference reader_concurrency_semaphore: make evict() noexcept reader_concurrency_semaphore: update out-of-date comments	2021-03-21 16:24:54 +02:00
Nadav Har'El	ab75226626	test/cql-pytest: remove xfail from passing test After commit `0bd201d3ca` ("cql3: Skip indexed column for CK restrictions") fixed issue #7888, the test cassandra_tests/validation/entities/frozen_collections_test.py::testClusteringColumnFiltering began passing, as expected. So we can remove its "xfail" label. Refs #7888. cassandra_tests/validation/entities/frozen_collections_test.py::testClusteringColumnFiltering Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20210321080522.1831115-1-nyh@scylladb.com>	2021-03-21 16:02:30 +02:00
Nadav Har'El	10bf2ba60a	cql-pytest: translate Cassandra's reproducers for issue #2962 This is a translation of Cassandra's CQL unit test source file validation/entities/SecondaryIndexOnMapEntriesTest.java into our our cql-pytest framework. This test file checks various features of indexing (with secondary index) individual entries of maps. All these tests pass on Cassandra, but fail on Scylla because of issue #2962 - we do not yet support indexing of the content of unfrozen collections. The failing test currently fail as soon as they try to create the index, with the message: "Cannot create secondary index on non-frozen collection or UDT column v". Refs #2962. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20210310124638.1653606-1-nyh@scylladb.com>	2021-03-21 12:30:00 +02:00
Avi Kivity	a78f43b071	Merge 'tracing: fast slow query tracing' from Ivan Prisyazhnyy The set of patches introduces a new tracing mode - `fast slow query tracing`. In this mode, Scylla tracks only tracing sessions and omits all tracing events if the tracing context does not have a `full_tracing` state set. Fixes #2572 Motivation --- We want to run production systems with that option always enabled so we could always catch slow queries without an overhead. The next step is we are gonna optimize further the costs of having tracing enabled to minimize session context handling overhead to allow it to be as transparent for the end-user as possible. Fast tracing mode --- To read the status do $ curl -v http://localhost:10000/storage_service/slow_query To enable fast slow-query tracing $ curl -v --request POST http://localhost:10000/storage_service/slow_query\?fast=true\&enable=true Potential optimizations --- - remove tracing::begin(lazy_eval) - replace tracing::begin(string) for enum to remove copying and memory allocations - merge parameters allocations - group parameters check for trace context - delay formatting - reuse prepared statement shared_ptr instead of both copying it and copying its query Performance --- 100% cache hits --- 1 Core: ``` $ SCYLLA_HOME=/home/sitano.public/Projects/scylla build/release/scylla --smp 1 --cpuset 7 --log-to-syslog 0 --log-to-stdout 1 --default-log-level info --network-stack posix --workdir /home/sitano.public/Projects/scylla --developer-mode 1 --listen-address 0.0.0.0 --api-address 0.0.0.0 --rpc-address 0.0.0.0 --broadcast-rpc-address 172.18.0.1 --broadcast-address 127.0.0.1 ./cassandra-stress write n=100000 no-warmup -pop seq=1..100000 -node 127.0.0.1 -log level=verbose -rate threads=1 -mode native cql3 curl --request POST http://localhost:10000/storage_service/slow_query\?fast\=false\&enable\=false for i in $(seq 5); do taskset -c 2,3,4,5 ./cassandra-stress read duration=5m -pop seq=1..100000 -node 127.0.0.1 -log level=verbose -rate threads=4 throttle=30000/s -mode native cql3 done curl --request POST http://localhost:10000/storage_service/slow_query\?fast\=true\&enable\=true for i in $(seq 5); do taskset -c 2,3,4,5 ./cassandra-stress read duration=5m -pop seq=1..100000 -node 127.0.0.1 -log level=verbose -rate threads=4 throttle=30000/s -mode native cql3 done curl --request POST http://localhost:10000/storage_service/slow_query\?fast\=false\&enable\=true for i in $(seq 5); do taskset -c 2,3,4,5 ./cassandra-stress read duration=5m -pop seq=1..100000 -node 127.0.0.1 -log level=verbose -rate threads=4 throttle=30000/s -mode native cql3 done ``` \| qps \| \| \| -- \| -- \| -- \| -- \| -- \| baseline \| fast, slow \| nofast, slow \| %[1-fastslow/baseline] \| 29,018 \| 26,468 \| 23,591 \| 8.79% \| 28,909 \| 26,274 \| 23,584 \| 9.11% \| 28,900 \| 26,547 \| 23,598 \| 8.14% \| 28,921 \| 26,669 \| 23,596 \| 7.79% \| 28,821 \| 26,385 \| 23,601 \| 8.45% stdev \| 70.24030182 \| 150.9678774 \| 6.670832032 \| avg \| 28,914 \| 26,469 \| 23,594 \| stderr \| 0.24% \| 0.57% \| 0.03% \| %[avg/baseline] \| \| 8.46% \| 18.40% \| 8.46% performance degradation in `fast slow query mode` for pure in-memory workload with minimum traces. 18.40% performance degradation in `original slow query mode` for pure in-memory workload with minimum traces. 0% cache hits --- 1GB memory, 1 Core: $ SCYLLA_HOME=/home/sitano.public/Projects/scylla build/release/scylla --memory 1G --smp 1 --cpuset 7 --log-to-syslog 0 --log-to-stdout 1 --default-log-level info --network-stack posix --workdir /home/sitano.public/Projects/scylla --developer-mode 1 --listen-address 0.0.0.0 --api-address 0.0.0.0 --rpc-address 0.0.0.0 --broadcast-rpc-address 172.18.0.1 --broadcast-address 127.0.0.1 2.4GB, 10000000 keys data: $ ./cassandra-stress write n=10000000 no-warmup -pop seq=1..10000000 -node 127.0.0.1 -log level=verbose -rate threads=4 -mode native cql3 $ curl --request POST http://localhost:10000/storage_service/slow_query\?fast\=true\&enable\=true CASSANDRA_STRESS prepared statements with BYPASS CACHE $ taskset -c 2,3,4,5 ./cassandra-stress read duration=5m -pop seq=1..10000000 -node 127.0.0.1 -log level=verbose -rate threads=4 throttle=30000/s -mode native cql3 20000 reads IOPS, 100MB/s from disk \| qps \| \| \| -- \| -- \| -- \| -- \| -- \| baseline reads \| fast, slow reads \| %[1-fastslow/baseline] \| \| 9,575 \| 9,054 \| 5.44% \| \| 9,614 \| 9,065 \| 5.71% \| \| 9,610 \| 9,066 \| 5.66% \| \| 9,611 \| 9,062 \| 5.71% \| \| 9,614 \| 9,073 \| 5.63% \| stdev \| 16.75410397 \| 6.892024376 \| avg \| 9,605 \| 9,064 \| stderr \| 0.17% \| 0.08% \| %[avg/baseline] \| \| 5.63% \| 5.63% performance degradation in `fast slow query mode` for pure on-disk workload with minimum traces. Closes #8314 * github.com:scylladb/scylla: tracing: fast mode unit test tracing: rest api for lightweight slow query tracing tracing: omit tracing session events and subsessions in fast mode	2021-03-21 12:15:17 +02:00
Dejan Mircevski	318f773d81	types: Unreverse tuple subtype for serialization When a tuple value is serialized, we go through every element type and use it to serialize element values. But an element type can be reversed, which is artificially different from the type of the value being read. This results in a server error due to the type mismatch. Fix it by unreversing the element type prior to comparing it to the value type. Fixes #7902 Tests: unit (dev) Signed-off-by: Dejan Mircevski <dejan@scylladb.com> Closes #8316	2021-03-21 12:07:29 +02:00
Dejan Mircevski	0bd201d3ca	cql3: Skip indexed column for CK restrictions When querying an index table, we assemble clustering-column restrictions for that query by going over the base table token, partition columns, and clustering columns. But if one of those columns is the indexed column, there is a problem; the indexed column is the index table's partition key, not clustering key. We end up with invalid clustering slice, which can cause problems downstream. Fix this by skipping the indexed column when assembling the clustering restrictions. Tests: unit (dev) Fixes #7888 Signed-off-by: Dejan Mircevski <dejan@scylladb.com> Closes #8320	2021-03-21 09:52:06 +02:00
Avi Kivity	58b7f225ab	keys: convert trichotomic comparators to return std::strong_ordering A trichotomic comparator returning an int an easily be mistaken for a less comparator as the return types are convertible. Use the new std::strong_ordering instead. A caller in cql3's update_parameters.hh is also converted, following the path of least resistance. Ref #1449. Test: unit (dev) Closes #8323	2021-03-21 09:30:43 +02:00
Tomasz Grabiec	88a019ba21	Merge "raft: respond with snapshot_reply to send_snapshot RPC" from Kostja Currently send_snapshot is the only two-way RPC used by Raft. However, the sender (the leader) does not look at the receiver's reply, other than checks it's not an error. This has the following issues: - if the follower has a newer term and rejects the snapshot for that reason, the leader will not learn about a newer follower term and will not step down - the send_snapshot message doesn't pass through a single-endpoint fsm::step() and thus may not follow the general Raft rules which apply for all messages. - making a general purpose transport that simply calls fsm::step() for every message becomes impossible. Fix it by actually responding with snapshot_reply to send_snapshot RPC, generating this reply in fsm::step() on the follower, and feeding into fsm::step() on the leader. * scylla-dev/raft-send-snapshot-v2: raft: pass snapshot_reply into fsm::step() raft: respond with snapshot_reply to send_snapshot RPC raft: set follower's next_idx when switching to SNAPSHOT mode raft: set the current leader upon getting InstallSnapshot	2021-03-19 18:13:40 +01:00
Raphael S. Carvalho	64d78eae6a	tests: Add unit test for off-strategy sstable compaction Signed-off-by: Raphael S. Carvalho <raphaelsc@scylladb.com>	2021-03-18 16:56:00 -03:00
Nadav Har'El	0b2cf21932	alternator-test: increase read timeout and avoid retries By default the boto3 library waits up to 60 second for a response, and if got no response, it sends the same request again, multiple times. We already noticed in the past that it retries too many times thus slowing down failures, so in our test configuration lowered the number of retries to 3, but the setting of 60-second-timeout plus 3 retries still causes two problems: 1. When the test machine and the build are extremely slow, and the operation is long (usually, CreateTable or DeleteTable involving multiple views), the 60 second timeout might not be enough. 2. If the timeout is reached, boto3 silently retries the same operation. This retry may fail because the previous one really succeeded at least partially! The symptom is tests which report an error when creating a table which already exists, or deleting a table which dooesn't exist. The solution in this patch is first of all to never do retries - if a query fails on internal server error, or times out, just report this failure immediately. We don't expect to see transient errors during local tests, so this is exactly the right behavior. The second thing we do is to increase the default timeout. If 1 minute was not enough, let's raise it to 5 minutes. 5 minutes should be enough for every operation (famous last words...). Even if 5 minutes is not enough for something, at least we'll now see the timeout errors instead of some wierd errors caused by retrying an operation which was already almost done. Fixes #8135 Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20210222125630.1325011-1-nyh@scylladb.com>	2021-03-18 18:58:08 +02:00

1 2 3 4 5 ...

1429 Commits