scylladb

mirror of https://github.com/scylladb/scylladb.git synced 2026-04-27 20:05:10 +00:00

Author	SHA1	Message	Date
Pavel Emelyanov	282a1880a5	forward service: Re-use proxy's helper with duplicated code The get_live_endpoints matches the same method on the proxy side. Since the forward service carries proxy reference, it can use its method (which needs to be made public for that sake). Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-05-03 10:34:51 +03:00
Avi Kivity	36aee57978	storage_proxy: convert rpc handlers from lambdas to member functions Currently, rpc handlers are all lambdas inside storage_proxy::init_messaging_service(). This means any stack trace refers to storage_proxy::init_messaging_service::lambda#n instead of a meaningful function name, and it makes init_messaging_service() very intimidating. Fix that by moving all such lambdas to regular member functions. This is easy now that they don't capture anything except `this`, which we provide during registration via std::bind_front(). A few #includes and forward declarations had to be added to storage_proxy.hh. This is unfortunate, but can only be solved by splitting storage_proxy into a client part and a server part.	2022-04-17 19:03:06 +03:00
Avi Kivity	4cac2eb43e	storage_proxy: don't capture migration_manager in server callbacks We'd like to make the server callbacks member functions, rather than lambdas, so we need to eliminate their captures. This patch eliminates 'mm' by making it a member variable and capturing 'this' instead. In one case 'mm' was used by a handle_write() intermediate lambda so we have to make that non-static and capture it too. uninit_messaging_service() clears the member variable to preserve the same lifetime 'mm' had before, in case that's important.	2022-04-17 17:54:51 +03:00
Nadav Har'El	fa7a302130	cross-tree: split coordinator_result from exceptions.hh Recently, coordinator_result was introduced as an alternative for exceptions. It was placed in the main "exceptions/exceptions.hh" header, which virtually every single source file in Scylla includes. But unfortunately, it brings in some heavy header files and templates, leading to a lot of wasted build time - ClangBuildAnalyzer measured that we include exceptions.hh in 323 source files, taking almost two seconds each on average. In this patch, we split the coordinator_result feature into a separate header file, "exceptions/coordinator_result", and only the few places which need it include the header file. Unfortunately, some of these few places are themselves header, so the new header file ends up being included in 100 source files - but 100 is still much less than 323 and perhaps we can reduce this number 100 later. After this patch, the total Scylla object-file size is reduced by 6.5% (the object size is a proxy for build time, which I didn't directly measure). ClangBuildAnalyzer reports that now each of the 323 includes of exceptions.hh only takes 80ms, coordinator_result.hh is only included 100 times, and virtually all the cost to include it comes from Boost's result.hh (400ms per inclusion). Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20220228204323.1427012-1-nyh@scylladb.com>	2022-03-02 10:12:57 +02:00
Piotr Dulikowski	e5922e650e	storage_proxy: resultify (do_)query Adjusts do_query so that it propagates and returns failed results. The query_result method is added which is result-aware, and the old query method was changed to call query_result.	2022-02-22 16:08:52 +01:00
Piotr Dulikowski	e39c5b6eba	storage_proxy: resultify query_singular Now, query_singular propagates and returns failed results without rethrowing them.	2022-02-22 16:08:52 +01:00
Piotr Dulikowski	2f5f746ae2	storage_proxy: propagate failed results through query_partition_key_range Now, query_partition_key_range propagates the failed result from query_partition_key_range_concurrent.	2022-02-22 16:08:52 +01:00
Piotr Dulikowski	608032b2b5	storage_proxy: resultify query_partition_key_range_concurrent Now, query_partition_key_range_concurrent propagates and returns exceptions as values, if possible.	2022-02-22 16:08:52 +01:00
Piotr Dulikowski	10923d9d58	storage_proxy: modify handle_read_error to also handle exception containers Now, storage_proxy::handle_read_error can work with both exception containers and exception_ptrs.	2022-02-22 16:08:52 +01:00
Piotr Dulikowski	68b5b84fbe	storage_proxy: resultify handling errors from read-repair Now, failed results returned from read-repair are handled without throwing.	2022-02-22 16:08:52 +01:00
Piotr Dulikowski	4c1eae7600	storage_proxy: change mutate_with_triggers to return future<result<>> Changes the interface of `mutate_with_triggers` so that it returns `future<result<>>` instead of `future<>`. No intermediate `mutate_with_triggers_result` method is introduced because all call sites will be changed in this PR so that they properly handle failed `result<>`s with exceptions-as-values.	2022-02-08 11:08:42 +01:00
Piotr Dulikowski	7ed668a177	storage_proxy: add mutate_atomically_result Similarly to `mutate_result` introduced in the previous commit, `mutate_atomically_result` is introduced which returns some exceptions inside `result<>`. The pre-existing `mutate_atomically` keeps the same interface but uses `mutate_atomically_result` internally, converting failed `result<>` to exceptional future if needed.	2022-02-08 11:08:42 +01:00
Piotr Dulikowski	f9ff5e7692	storage_proxy: return result<> from mutate_result In order to be able to propagate exceptions-as-values from storage_proxy but without having to modify all call sites of `mutate`, an in-between method `mutate_result` is introduced which returns some exceptions inside `result<>`. Now, `mutate` just calls the latter and converts those exceptions to exceptional future if needed.	2022-02-08 11:08:42 +01:00
Piotr Dulikowski	f02b8614af	storage_proxy: return result<> from mutate_internal Changes the interface of `mutate_internal` so that it returns a `future<result<>>` instead of `future<>`.	2022-02-08 11:08:42 +01:00
Piotr Dulikowski	e2893368a7	storage_proxy: handle exceptions as values in mutate_end Instead of stupidly rethrowing the exception in failed result<>, the `storage_proxy::mutate_end` function now inspects it with a visitor, which does not involve any rethrows. Moreover, mutate_end now also returns a `future<result<>>` instead of just `future<>`.	2022-02-08 11:08:42 +01:00
Piotr Dulikowski	5c00b27662	storage_proxy: let mutate_end take a future<result<>> Changes the `storage_proxy::mutate_end` method to accept a `future<result<>>` instead of `future<>`. For the time being, all call call sites of that method pass a future which is either exceptional or contains a result<> with a value. Moreover, in case of a failed result<>, mutate_end just rethrows the exception. Both of these will change in the upcoming commits of this PR.	2022-02-08 11:08:42 +01:00
Piotr Dulikowski	59efe085af	storage_proxy: resultify mutate_begin Changes the `storage_proxy::mutate_begin` method to return a future<result<>>.	2022-02-08 11:08:42 +01:00
Piotr Dulikowski	6ac98f26e0	storage_proxy: introduce helpers for dealing with results Adds a number of typedefs in order to make working with coordinator exceptions-as-values easier.	2022-02-08 11:08:42 +01:00
Michał Sala	0fe59082ec	storage_proxy: extract query_ranges_to_vnodes_generator to a separate file Such separation allows using query_ranges_to_vnodes_generator by other services without needing a storage_proxy dependency.	2022-02-01 21:14:41 +01:00
Avi Kivity	fcb8d040e8	treewide: use Software Package Data Exchange (SPDX) license identifiers Instead of lengthy blurbs, switch to single-line, machine-readable standardized (https://spdx.dev) license identifiers. The Linux kernel switched long ago, so there is strong precedent. Three cases are handled: AGPL-only, Apache-only, and dual licensed. For the latter case, I chose (AGPL-3.0-or-later and Apache-2.0), reasoning that our changes are extensive enough to apply our license. The changes we applied mechanically with a script, except to licenses/README.md. Closes #9937	2022-01-18 12:15:18 +01:00
Avi Kivity	bbad8f4677	replica: move ::database, ::keyspace, and ::table to replica namespace Move replica-oriented classes to the replica namespace. The main classes moved are ::database, ::keyspace, and ::table, but a few ancillary classes are also moved. There are certainly classes that should be moved but aren't (like distributed_loader) but we have to start somewhere. References are adjusted treewide. In many cases, it is obvious that a call site should not access the replica (but the data_dictionary instead), but that is left for separate work. scylla-gdb.py is adjusted to look for both the new and old names.	2022-01-07 12:04:38 +02:00
Avi Kivity	ae3a360725	database: Move database, keyspace, table classes to replica/ directory The database, keyspace, and table classes represent the replica-only part of the objects after which they are named. Reading from a table doesn't give you the full data, just the replica's view, and it is not consistent since reconciliation is applied on the coordinator. As a first step in acknowledging this, move the related files to a replica/ subdirectory.	2022-01-06 17:07:30 +02:00
Nadav Har'El	6012f6f2b6	build performance: do not include <seastar/net/ip.hh> In a previous patch, we noticed that the header file <gm/inet_address.hh>, which is included, directly or indirectly, by most source files, includes <seastar/net/ip.hh> which is very slow to compile, and replaced it by the much faster-to-include <seastar/net/ipv[46]_address.hh>. However, we also included <seastar/net/ip.hh> in types.hh - and that too is included by almost every file, so the actual saving from the above patch was minimal. So in this patch we replace this include too. After this patch Scylla does not include <seastar/net/ip.hh> at all. According to ClangBuildAnalyzer, this reduces the average time to include types.hh (multiply this by 312 times!) from 4 seconds to 1.8 seconds, and reduces total build time (dev mode) by about 3%. Some of the source files were now missing some include directives, that were previously included in ip.hh - so we need to add those explicitly. Signed-off-by: Nadav Har'El <nyh@scylladb.com>	2022-01-05 17:29:21 +02:00
Avi Kivity	c2da20484d	storage_proxy: provide access to data_dictionary Probably storage_proxy is not the correct place to supply data_dictionary, but it is available to practically all of the coordinator code, so it is convenient.	2021-12-15 13:54:08 +02:00
Benny Halevy	744275df73	batchlog_manager: get_batch_log_mutation_for: move to storage_proxy And rename to get_batchlog_mutation_for while at it, as it's about the batchlog, not batch_log. This resolves a circular dependency between the batchlog_manager and the storage_proxy that required it in the case. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-11-23 08:27:30 +02:00
Benny Halevy	55967a8597	batchlog_manager: endpoint_filter: move to gossiper There's nothing in this function that actually requries the batchlog manager instance. It uses a random number engine that's moved along with it to class gossiper. This resolves a circular dependency between the batchlog_manager and storage_proxy. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-11-23 08:27:30 +02:00
Benny Halevy	242043368e	main: pass erm_factory to storage_proxy To be used for creating the effective_replication_map per keyspace. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-11-19 10:46:51 +02:00
Pavel Emelyanov	1ea301ad07	storage_proxy: Propagate virtual table exceptions messages The intention is to return some meaningful info to the CQL caller if a virtual table update fails. Unfortunately the "generic" error reporting in CQL is not extremely flexible, so the best option seems to report regular write failre with custom message in it. For now this only works for virtual table errors. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-11-11 16:39:34 +03:00
Pavel Emelyanov	598841a5dd	code: Expell gossiper.hh from other headers This needs to add forward declarations of the gossiper class and re-include some other headers here and there. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-09-22 13:13:06 +03:00
Pavel Emelyanov	dd498273dc	messaging, proxy: Notify connection drops with boost signal The messaging_service keeps track of a list of connection-drop listeners. This list is not auto-removing and is thus not safe on stop (fortunately there's only 1 non-stopping client of it so far). This patch adds a safter notification based on boost/signals. Also storage_proxy is subscribed on it in advance to demonstrate how it looks like altogether and make next patch shorter. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-09-15 17:49:06 +03:00
Avi Kivity	37818170d8	storage_proxy: start_hints_manager(): don't require caller to provide gossiper storage_proxy now maintains a reference to gossiper, so it can simplify its callers.	2021-09-07 20:08:15 +03:00
Avi Kivity	71081be99c	storage_proxy: remove uses of get_local_gossiper() Pass the gossiper as a constructor parameter instead.	2021-09-07 17:14:09 +03:00
Piotr Dulikowski	14b00610b2	storage_proxy: add functions for creating and waiting for hint sync pts Adds functions in storage_proxy which allow to create sync points and wait for them.	2021-08-09 09:24:36 +02:00
Piotr Dulikowski	4a35d138f6	Revert "storage_proxy: add functions for syncing with hints queue" This reverts commit `244738b0d5`. This commit removes create_hint_queue_sync_point and check_hint_queue_sync_point functions from storage_proxy, which were used to wait until local hints are sent out to particular nodes. Similar methods will be reintroduced later in this PR, with a completely different implementation.	2021-08-09 09:24:36 +02:00
Piotr Dulikowski	6c5d2fe0bf	Revert "storage_proxy: coordinate waiting for hints to be sent" This reverts commit `46075af7c4`. This commit removes the logic responsible for waiting for other nodes to replay their hints. The upcoming HTTP API for waiting for hint replay will be restricted to waiting for hints on the node handling the request, so there is no need for coordinating multiple nodes.	2021-08-09 09:24:36 +02:00
Piotr Dulikowski	035da96161	Revert "storage_proxy: add abort_source to wait_for_hints_to_be_replayed" This reverts commit `958a13577c`. The `wait_for_hints_to_be_replayed` function is going to be completely removed in this PR, so this commit needs to be reverted, too.	2021-08-09 09:06:23 +02:00
Avi Kivity	b099e7c254	Merge "Untie hints managers and storage service" from Pavel E " The storage service is carried along storage proxy, hints resource manager and hints managers (two of them) just to subscribe the hints managers on lifecycle events (and stop the subscription on shutdown) emitted from storage service. This dependency chain can be greatly simplified, since the storage proxy is already subscribed on lifecycle events and can kick managers directly from its hooks. tests: unit(dev), dtest.hintedhandoff_additional_test.hintedhandoff_basic_check_test(dev) " * 'br-remove-storage-service-from-hints' of https://github.com/xemul/scylla: hints: Drop storage service from managers hints: Do not subscribe managers on lifecycle events directly	2021-06-17 17:12:31 +03:00
Pavel Emelyanov	92a4278cd1	hints: Drop storage service from managers The storage service pointer is only used so (un)subscribe to (from) lifecycle events. Now the subscription is gone, so can the storage service pointer. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-06-17 15:09:36 +03:00
Pavel Emelyanov	acdc568ecf	hints: Do not subscribe managers on lifecycle events directly Managers sit on storage proxy which is already subscribed on lifecycle events, so it can "notify" hints managers directly. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-06-17 15:06:26 +03:00
Avi Kivity	4d70f3baee	storage_proxy: change unordered_set<inet_address> to small_vector in write path The write paths in storage_proxy pass replica sets as std::unordered_set<gms::inet_address>. This is a complex type, with N+1 allocations for N members, so we change it to a small_vector (via inet_address_vector_replica_set) which requires just one allocation, and even zero when up to three replicas are used. This change is more nuanced than the corresponding change to the read path `abe3d7d7` ("Merge 'storage_proxy: use small_vector for vectors of inet_address' from Avi Kivity"), for two reasons: - there is a quadratic algorithm in abstract_write_response_handler::response(): it searches for a replica and erases it. Since this happens for every replica, it happens N^2/2 times. - replica sets for writes always include all datacenters, while reads usually involve just one datacenter. So, a write to a keyspace that has 5 datacenters will invoke 15*(15-1)/2 =105 compares. We could remove this by sending the index of the replica in the replica set to the replica and ask it to include the index in the response, but I think that this is unnecessary. Those 105 compares need to be only 105/15 = 7 times cheaper than the corresponding unordered_set operation, which they surely will. Handling a response after a cross-datacenter round trip surely involves L3 cache misses, and a small_vector reduces these to a minimum compared to an unordered_set with its bucket table, linked list walking and managent, and table rehashing. Tests using perf_simple_query --write --smp 1 --operations-per-shard 1000000 --task-quota-ms show two allocations removed (as expected) and a nice reduction in instructions executed. before: median 204842.54 tps ( 54.2 allocs/op, 13.2 tasks/op, 49890 insns/op) after: median 206077.65 tps ( 52.2 allocs/op, 13.2 tasks/op, 49138 insns/op) Closes #8847	2021-06-17 13:46:40 +03:00
Avi Kivity	a55b434a2b	treewide: extent copyright statements to present day	2021-06-06 19:18:49 +03:00
Pavel Solodovnikov	2187a59089	treewide: move `service::cas_request` out from `storage_proxy.hh` And remove all remaining inclusions of `storage_proxy.hh` in the headers. Signed-off-by: Pavel Solodovnikov <pa.solodovnikov@scylladb.com>	2021-06-06 19:18:49 +03:00
Pavel Emelyanov	651568318d	api: Get features from proxy The reset_local_schema call needs proxy and feature service to do its job. Right now the features are retrived from global storage service, but they are present on the proxy as well. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2021-05-28 18:15:15 +03:00
Kamil Braun	03ad111beb	tree-wide: comments on deprecated functions to access global variables Closes #8665	2021-05-18 11:31:10 +03:00
Avi Kivity	dd904f7776	storage_proxy: place unique_response_handler:s in small_vector instead of std::vector This cuts an allocation in the write path. Instruction count reduction isn't large, but performance does improve (results are consistent): before: 196369.48 tps ( 55.2 allocs/op, 13.2 tasks/op, 51658 insns/op) after: 199290.32 tps ( 54.2 allocs/op, 13.2 tasks/op, 51600 insns/op) (this is perf_simple_query --write --smp 1 --operations-per-shard 1000000)	2021-05-09 23:53:10 +03:00
Avi Kivity	b015dedd96	storage_proxy: make unique_response_handler friendly to small_vector small_vector wants the move constructor to be noexcept, and move assighment to exist (and be noexcept). These are easy to achieve.	2021-05-09 19:22:20 +03:00
Avi Kivity	8398510382	storage_proxy: give a name to a vector of unique_response_handlers We'd like to change the vector to a more involved type, so to avoid repeating it everywhere, give it a name. The actual type isn't changed in this patch.	2021-05-09 19:17:50 +03:00
Tomasz Grabiec	abe3d7d7d3	Merge 'storage_proxy: use small_vector for vectors of inet_address' from Avi Kivity storage_proxy uses std::vector<inet_address> for small lists of nodes - for replication (often 2-3 replicas per operation) and for pending operations (usually 0-1). These vectors require an allocation, sometimes more than one if reserve() is not used correctly. This series switches storage_proxy to use utils::small_vector instead, removing the allocations in the common case. Test results (perf_simple_query --smp 1 --task-quota-ms 10): ``` before: median 184810.98 tps ( 91.1 allocs/op, 20.1 tasks/op, 54564 insns/op) after: median 192125.99 tps ( 87.1 allocs/op, 20.1 tasks/op, 53673 insns/op) ``` 4 allocations and ~900 instructions are removed (the tps figure is also improved, but it is less reliable due to cpu frequency changes). The type change is unfortunately not contained in storage_proxy - the abstraction leaks to providers of replica sets and topology change vectors. This is sad but IMO the benefits make it worthwhile. I expect more such changes can be applied in storage_proxy, specifically std::unordered_set<gms::inet_address> and vectors of response handles. Closes #8592 * github.com:scylladb/scylla: storage_proxy, treewide: use utils::small_vector inet_address_vector:s storage_proxy, treewide: introduce names for vectors of inet_address utils: small_vector: add print operator for std::ostream hints: messages.hh: add missing #include	2021-05-06 18:00:54 +02:00
Avi Kivity	cea5493cb7	storage_proxy, treewide: introduce names for vectors of inet_address storage_proxy works with vectors of inet_addresses for replica sets and for topology changes (pending endpoints, dead nodes). This patch introduces new names for these (without changing the underlying type - it's still std::vector<gms::inet_address>). This is so that the following patch, that changes those types to utils::small_vector, will be less noisy and highlight the real changes that take place.	2021-05-05 18:36:48 +03:00
Nadav Har'El	58e275e362	cross-tree: reduce dependency on db/config.hh and database.hh Every time db/config.hh is modified (e.g., to add a new configuration option), 110 source files need to be recompiled. Many of those 110 didn't really care about configuration options, and just got the dependency accidentally by including some other header file. In this patch, I remove the include of "db/config.hh" from all header files. It is only needed in source files - and header files only need forward declarations. In some cases, source files were missing certain includes which they got incidentally from db/config.hh, so I had to add these includes explicitly. After this patch, the number of source files that get recompiled after a change to db/config.hh goes down from 110 to 45. It also means that 65 source files now compile faster because they don't include db/config.hh and whatever it included. Additionally, this patch also eliminates a few unnecessary inclusions of database.hh in other header files, which can use a forward declaration or database_fwd.hh. Some of the source files including one of those header files relied on one of the many header files brought in by database.hh, so we need to include those explicitly. In view_update_generator.hh something interesting happened - it needs database.hh because of code in the header file, but only included database_fwd.hh, and the only reason this worked was that the files including view_update_generator.hh already happened to unnecessarily include database.hh. So we fix that too. Refs #1 Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20210505102111.955470-1-nyh@scylladb.com>	2021-05-05 13:23:00 +03:00

1 2 3 4 5 ...

322 Commits