scylladb

mirror of https://github.com/scylladb/scylladb.git synced 2026-04-26 19:35:12 +00:00

Author	SHA1	Message	Date
Tomasz Grabiec	5b046043ea	migration_manager: Make prepare_keyspace_drop_announcement() return a future<> It will be extended with listener notification firing, which is an async operation.	2023-04-24 10:49:37 +02:00
Kefu Chai	ecb5380638	treewide: s/boost::lexical_cast<std::string>/fmt::to_string()/ this change replaces all occurrences of `boost::lexical_cast<std::string>` in the source tree with `fmt::to_string()`. for couple reasons: * `boost::lexical_cast<std::string>` is longer than `fmt::to_string()`, so the latter is easier to parse and read. * `boost::lexical_cast<std::string>` creates a stringstream under the hood, so it can use the `operator<<` to stringify the given object. but stringstream is known to be less performant than fmtlib. * we are migrating to fmtlib based formatting, see #13245. so using `fmt::to_string()` helps us to remove yet another dependency on `operator<<`. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> Closes #13611	2023-04-21 09:43:53 +03:00
Nadav Har'El	d26bb8c12d	Merge 'tree: migrate from std::regex to boost::regex' from Botond Dénes Except for where usage of `std::regex` is required by 3rd party library interfaces. As demonstrated countless times, std::regex's practice of using recursion for pattern matching can result in stack overflow, especially on AARCH64. The most recent incident happened after merging https://github.com/scylladb/scylladb/pull/13075, which (indirectly) uses `sstables::make_entry_descriptor()` to test whether a certain path is a valid scylla table path in a trial-and-error manner. This resulted in stacks blowing up in AARCH64. To prevent this, use the already tried and tested method of switching from `std::regex` to `boost::regex`. Don't wait until each of the `std::regex` sites explode, replace them all preemptively. Refs: https://github.com/scylladb/scylladb/issues/13404 Closes #13452 * github.com:scylladb/scylladb: test: s/std::regex/boost::regex/ utils: s/std::regex/boost::regex/ db/commitlog: s/std::regex/boost::regex/ types: s/std::regex/boost::regex/ index: s/std::regex/boost::regex/ duration.cc: s/std::regex/boost::regex/ cql3: s/std::regex/boost::regex/ thrift: s/std::regex/boost::regex/ sstables: use s/std::regex/boost::regex/	2023-04-09 18:47:41 +03:00
Kefu Chai	7a05cc3a06	thrift: initiaize _config first to avoid dangling reference in `c642ca9e73`, a reference to the a parameter `config` passed to the `thrift_server` 's constructor is passed down to `create_handler_factory()`, which keeps it so it can create connection handler on demand. but unfortunately, - the `config` parameter is a temporary variable - the `config` parameter is moved away in the constructor after `create_handler_factory()` is called hence we have a dangling reference when the factory created by `create_handler_factory()` tries to deference the reference when handling a new incoming connection. in this change, - the definitions of `_config` and `_handler_factory` member variables are transposed, so that the former is initialized first. - `_handler_factory` now keeps a reference to `_config`'s member variable, so that the weak reference it holds is always valid. Fixes #13455 Branches: none Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> Closes #13456	2023-04-09 11:34:34 +03:00
Botond Dénes	c0b72f70d4	thrift: s/std::regex/boost::regex/ The former is prone to producing stack-overflow as it uses recursion in it match implementation. The migration is entirely mechanical.	2023-04-06 09:50:27 -04:00
Kefu Chai	ebf5e138e8	redis,thrift,transport: make timeout_config live-updateable * timeout_config - add `updated_timeout_config` which represents an always-updated options backed by `utils::updateable_value<>`. this class is used by servers which need to access the latest timeout related options. the existing `timeout_config` is more like a snapshot of the `updated_timeout_config`. it is used in the use case where we don't need to most updated options or we update the options manually on demand. * redis, thrift, transport: s/timeout_config/updated_timeout_config/ when appropriate. use the improved version of timeout_config where we need to have the access to the most-updated version of the timeout options. Fixes #10172 Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2023-03-29 20:17:45 +08:00
Kefu Chai	fec35b97ad	thrift: keep a reference of timeout_config in handler_factory this change should keep the timeout settings of handler_factory sync'ed with the ones used by `thrift_server`. so far, the `timeout_config` instance in `thrift_server` is not live-updateable, but in a follow-up change, we will make it so. so, this change prepares the handler_factory for a live-updateable timeout_config. instead keeping a snapshot of the timeout_config, keep a reference of it in handler_factory. the reference points to `thrift_server::_config`. so despite that `thrift_server::_handler_factory` is a shared_ptr, the member variable won't outlive its container, as the only reason to have it as a shared_ptr is to appease the ctor of `CassandraAsyncProcessorFactory`. and the constructed `_processor_factory` is also a member variable of `thrift_server`, so we won't take the risk of a dangling reference held by `handler_factory`. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2023-03-29 20:06:02 +08:00
Kefu Chai	c642ca9e73	redis,thrift,transport: initialize _config with std::move(config) instead of copying the `config` parameter, move away from it. this change also prepares for a non-copyable config. if the class of `config` is not copyable, we will not be able to initialize the member variable by copying from the given `config` parameter. after the live-updateable config change, the `_config` member variable will contain instances of utils::observer<>, which is not copyable, but is move-constructable, hence in this change, we just move away from the give `config`. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2023-03-29 20:06:02 +08:00
Kefu Chai	e0ac2eb770	redis,thrift,transport: pass config via sharded_parameter * pass config via sharded_parameter * initialize config using designated initializer this change paves the road to servers with live-updateable timeout options. before this change, the servers initialize a domain specific combo config, like `redis_server_config`, with the same instance of a timeout_config, and pass the combox config as a ctor parameter to construct each sharded service instance. but this design assumes the value semantic of the config class, say, it should be copyable. but if we want to use utils::updateable_value<> to get updated option values, we would have to postpone the instantiation of the config until the sharded service is about to be initialized. so, in this change, instead of taking a domain specific config created before hand, all services constructed with a `timeout_config` will take a `sharded_parameter()` for creating the config. also, take this opportunity to initialize the config using designated initializer. for two reasons: * less repeatings this way. we don't have to repeat the variable name of the config being initialized for each member variable. * prepare for some member variables which do not have a default constructor. this applies to the timeout_config's updater which will not have a default constructor, as it should be initialized by db::config and a reference to the timeout_config to be updated. we will update the `timeout_config` side in a follow-up commit. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2023-03-29 20:06:00 +08:00
Marcin Maliszkiewicz	339a8fe64d	thrift: return address in listen_addresses() only after server is ready listen_addresses() checks if _server variable is empty and after this patch we assign (move) the value only after server is ready. This is used for readiness API: /storage_service/rpc_server and the fix prevents from returning 'true' prematurely. Some improvement for readiness was added in `a51529dd15` but thrift implementation wasn't fully done. Fixes #12376	2023-03-27 13:20:53 +02:00
Marcin Maliszkiewicz	a38701b9d4	thrift: simplify do_start_server() with seastar:async Code is executed typically on startup only so overhead is very limited. Notably using async avoids managing tserver variable lifetime.	2023-03-27 13:12:10 +02:00
Kefu Chai	3e75df6917	build: cmake: extract thrift out also, move "interface" linkage from scylla to "thrift", because it is "thrift" who is using "interface". Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2023-02-28 21:28:46 +08:00
Kefu Chai	0cb842797a	treewide: do not define/capture unused variables these warnings are found by Clang-17 after removing `-Wno-unused-lambda-capture` and '-Wno-unused-variable' from the list of disabled warnings in `configure.py`. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2023-02-15 22:57:18 +02:00
Avi Kivity	69a385fd9d	Introduce schema/ module Schema related files are moved there. This excludes schema files that also interact with mutations, because the mutation module depends on the schema. Those files will have to go into a separate module. Closes #12858	2023-02-15 11:01:50 +02:00
Avi Kivity	c5e4bf51bd	Introduce mutation/ module Move mutation-related files to a new mutation/ directory. The names are kept in the global namespace to reduce churn; the names are unambiguous in any case. mutation_reader remains in the readers/ module. mutation_partition_v2.cc was missing from CMakeLists.txt; it's added in this patch. This is a step forward towards librarization or modularization of the source base. Closes #12788	2023-02-14 11:19:03 +02:00
Avi Kivity	2739ac66ed	treewide: drop cql_serialization_format Now that we don't accept cql protocol version 1 or 2, we can drop cql_serialization format everywhere, except when in the IDL (since it's part of the inter-node protocol). A few functions had duplicate versions, one with and one without a cql_serialization_format parameter. They are deduplicated. Care is taken that `partition_slice`, which communicates the cql_serialization_format across nodes, still presents a valid cql_serialization_format to other nodes when transmitting itself and rejects protocol 1 and 2 serialization\ format when receiving. The IDL is unchanged. One test checking the 16-bit serialization format is removed.	2023-01-03 19:54:13 +02:00
Botond Dénes	d1d53f1b84	query: add tombstone-limit to read-command Propagate the tombstone-limit from coordinator to replicas, to make sure all is using the same limit.	2022-08-10 06:01:47 +03:00
Benny Halevy	257d74bb34	schema, everywhere: define and use table_id as a strong type Define table_id as a distinct utils::tagged_uuid modeled after raft tagged_id, so it can be differentiated from other uuid-class types, in particular from table_schema_version. Fixes #11207 Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2022-08-08 08:09:41 +03:00
Piotr Dulikowski	a7ad70600d	query-request: add allow_limit flag Adds allow_limit flag to the read_command. The flag decides whether rate limiting of this operation is allowed.	2022-06-22 20:16:49 +02:00
Piotr Dulikowski	e6beab3106	storage_proxy: add allow rate limit flag to mutate/mutate_result Now, mutate/mutate_result accept a flag which decides whether the write should be rate limited or not. The new parameter is mandatory and all call sites were updated.	2022-06-22 20:16:49 +02:00
Avi Kivity	528ab5a502	treewide: change metric calls from make_derive to make_counter make_derive was recently deprecated in favor of make_counter, so make the change throughput the codebase. Closes #10564	2022-05-14 12:53:55 +02:00
Avi Kivity	5937b1fa23	treewide: remove empty comments in top-of-files After `fcb8d040` ("treewide: use Software Package Data Exchange (SPDX) license identifiers"), many dual-licensed files were left with empty comments on top. Remove them to avoid visual noise. Closes #10562	2022-05-13 07:11:58 +02:00
Piotr Dulikowski	e4ff22b4ca	result_message: add result_message::exception In order to propagate exceptions as values through the CQL layer with minimal modifications to the interfaces, a new result_message type is introduced: result_message::exception. Similarly to result_message::bounce_to_shard, this is an internal type which is supposed to be handled before being returned to the client.	2022-02-08 11:08:42 +01:00
Kamil Braun	a664ac7ba5	treewide: require `group0_guard` when performing schema changes `announce` now takes a `group0_guard` by value. `group0_guard` can only be obtained through `migration_manager::start_group0_operation` and moved, it cannot be constructed outside `migration_manager`. The guard will be a method of ensuring linearizability for group 0 operations.	2022-01-24 15:20:35 +01:00
Kamil Braun	86762a1dd9	service: migration_manager: rename `schema_read_barrier` to `start_group0_operation` 1. Generalize the name so it mentions group 0, which schema will be a strict subset of. 2. Remove the fact that it performs a "read barrier" from the name. The function will be used in general to ensure linearizability of group0 operations - both reads and writes. "Read barrier" is Raft-specific terminology, so it can be thought of as an implementation detail.	2022-01-24 15:12:50 +01:00
Kamil Braun	283ac7fefe	treewide: pass mutation timestamp from call sites into `migration_manager::prepare_*` functions The functions which prepare schema change mutations (such as `prepare_new_column_family_announcement`) would use internally generated timestamps for these mutations. When schema changes are managed by group 0 we want to ensure that timestamps of mutations applied through Raft are monotonic. We will generate these timestamps at call sites and pass them into the `prepare_` functions. This commit prepares the APIs.	2022-01-24 15:12:50 +01:00
Avi Kivity	fcb8d040e8	treewide: use Software Package Data Exchange (SPDX) license identifiers Instead of lengthy blurbs, switch to single-line, machine-readable standardized (https://spdx.dev) license identifiers. The Linux kernel switched long ago, so there is strong precedent. Three cases are handled: AGPL-only, Apache-only, and dual licensed. For the latter case, I chose (AGPL-3.0-or-later and Apache-2.0), reasoning that our changes are extensive enough to apply our license. The changes we applied mechanically with a script, except to licenses/README.md. Closes #9937	2022-01-18 12:15:18 +01:00
Gleb Natapov	d65427ad81	thrift: correctly check for keyspace existence `d9c315891a` broke the check for keyspace existence. The condition is opposite. Fix it. Fixes #9927 Message-Id: <YeUhtESDHQeMHiUW@scylladb.com>	2022-01-17 10:20:48 +02:00
Pavel Emelyanov	f22eb22b8b	client_state: Make has_schema_access use data_dictionary::database It's now called with d._d.::database converted to .real_database() right in the argument passing, so this change can be treated as the generalization of that .real_database() call. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-01-14 12:55:53 +03:00
Pavel Emelyanov	b6bc7a9b29	client_state: Make has_column_family_access use data_dictionary::database Straightforward replacement. Internals of the has_column_family_access() temporarily get .real_database(), but it will be changed soon. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-01-14 12:55:15 +03:00
Pavel Emelyanov	1ed237120a	client_state: Make has_keyspace_access use data_dictionary::database Straightforward replacement. Internals of the has_keyspace_access() temporarily get .real_database(), but it will be changed soon. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-01-14 12:54:01 +03:00
Avi Kivity	6205d40d5f	thrift: switch from replica module to data_dictionary module Thrift is a coordinator-side service and should not touch the replica module. Switch it to data_dictionary. The switch is straightforward with two exceptions: - client_state still receives replica::database parameters. After this change it will be easier to adapt client_state too. - calls to replica::database::get_version() remain. They should be rerouted to migration_manager instead, as that deals with schema management.	2022-01-12 19:54:38 +02:00
Avi Kivity	85061b694b	thrift: simplify execute_schema_command() calling convention execute_schema_command is always called with the same first two parameters, which are always defined froom the thrift_handler instance that contains its caller. Simplify it by making it a member function. This simplifies migration to data_dictionary in the next patch.	2022-01-12 18:56:47 +02:00
Gleb Natapov	dd36150a7d	thrift: move system_update_column_family() to raft	2022-01-12 16:33:16 +02:00
Gleb Natapov	bcfdcc51d6	thrift: authenticate a statement before verifying in system_update_column_family() Otherwise it is possible to infer if a table exist without having proper credentials.	2022-01-12 16:33:16 +02:00
Gleb Natapov	aec413d0f7	thrift: co-routinize system_update_column_family()	2022-01-12 16:33:16 +02:00
Gleb Natapov	d9c315891a	thrift: move system_update_keyspace() to raft	2022-01-12 16:33:16 +02:00
Gleb Natapov	7ffbdde554	thrift: authenticate a statement before verifying in system_update_keyspace() Otherwise it is possible to infer if a table exist without having proper credentials.	2022-01-12 16:33:16 +02:00
Gleb Natapov	1b4538f5bd	thrift: co-routinize system_update_keyspace()	2022-01-12 16:33:16 +02:00
Gleb Natapov	64b8f4fe50	thrift: move system_drop_keyspace() to raft	2022-01-12 16:33:16 +02:00
Gleb Natapov	52fc815f24	thrift: authenticate a statement before verifying in system_drop_keyspace() Otherwise it is possible to infer if a table exist without having proper credentials.	2022-01-12 16:33:16 +02:00
Gleb Natapov	45ff7e30a1	thrift: co-routinize system_drop_keyspace()	2022-01-12 16:33:16 +02:00
Gleb Natapov	a17f82c647	thrift: move system_add_keyspace() to raft	2022-01-12 16:33:16 +02:00
Gleb Natapov	3a3a3f693e	thrift: co-routinize system_add_keyspace()	2022-01-12 16:33:16 +02:00
Gleb Natapov	845b617256	thrift: move system_drop_column_family() to raft	2022-01-12 16:33:16 +02:00
Gleb Natapov	9b6a9b104e	thrift: co-routinize system_drop_column_family()	2022-01-12 16:33:16 +02:00
Gleb Natapov	7cfedb50bb	thrift: move system_add_column_family() to raft	2022-01-12 16:33:16 +02:00
Gleb Natapov	e4ac3c2777	thrift: authenticate a statement before verifying in system_add_column_family() Otherwise it is possible to infer if a table exist without having proper credentials.	2022-01-12 16:33:16 +02:00
Gleb Natapov	d5f14306d0	thrift: co-routinize system_add_column_family()	2022-01-12 16:33:16 +02:00
Avi Kivity	bbad8f4677	replica: move ::database, ::keyspace, and ::table to replica namespace Move replica-oriented classes to the replica namespace. The main classes moved are ::database, ::keyspace, and ::table, but a few ancillary classes are also moved. There are certainly classes that should be moved but aren't (like distributed_loader) but we have to start somewhere. References are adjusted treewide. In many cases, it is obvious that a call site should not access the replica (but the data_dictionary instead), but that is left for separate work. scylla-gdb.py is adjusted to look for both the new and old names.	2022-01-07 12:04:38 +02:00

1 2 3 4 5 ...

406 Commits