scylladb

Author	SHA1	Message	Date
Avi Kivity	f3eade2f62	treewide: relicense to ScyllaDB-Source-Available-1.0 Drop the AGPL license in favor of a source-available license. See the blog post [1] for details. [1] https://www.scylladb.com/2024/12/18/why-were-moving-to-a-source-available-license/	2024-12-18 17:45:13 +02:00
Kefu Chai	bab12e3a98	treewide: migrate from boost::adaptors::transformed to std::views::transform now that we are allowed to use C++23. we now have the luxury of using `std::views::transform`. in this change, we: - replace `boost::adaptors::transformed` with `std::views::transform` - use `fmt::join()` when appropriate where `boost::algorithm::join()` is not applicable to a range view returned by `std::view::transform`. - use `std::ranges::fold_left()` to accumulate the range returned by `std::view::transform` - use `std::ranges::fold_left()` to get the maximum element in the range returned by `std::view::transform` - use `std::ranges::min()` to get the minimal element in the range returned by `std::view::transform` - use `std::ranges::equal()` to compare the range views returned by `std::view::transform` - remove unused `#include <boost/range/adaptor/transformed.hpp>` - use `std::ranges::subrange()` instead of `boost::make_iterator_range()`, to feed `std::views::transform()` a view range. to reduce the dependency to boost for better maintainability, and leverage standard library features for better long-term support. this change is part of our ongoing effort to modernize our codebase and reduce external dependencies where possible. limitations: there are still a couple places where we are still using `boost::adaptors::transformed` due to the lack of a C++23 alternative for `boost::join()` and `boost::adaptors::uniqued`. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> Closes scylladb/scylladb#21700	2024-12-03 09:41:32 +02:00
Avi Kivity	907da210b6	compound_compat: replace use of boost ranges with std ranges To reduce the dependency load, replace use of boost ranges with the std equivalent. Files that lost the indirect boost dependency have it added as a direct dependency.	2024-10-30 19:58:07 +02:00
Emil Maskovsky	a840949ea0	treewide: code cleanup and refactoring Fix the clang-tidy warnings, code cleanup and improvements. Applied the clang format to the updated places.	2024-10-08 20:53:54 +02:00
Gleb Natapov	f00ea36f63	gossiper: remove unused REMOVAL_COORDINATOR state This is leftover from `66ff072540`	2024-02-19 15:01:33 +02:00
Gleb Natapov	15a34f650d	gossip: remove unused HIBERNATE gossiper status The status is not used since `2ec1f719de` which is included in scylla-4.6.0. We cannot have mixed cluster with the version so old, so the new version should not carry the compatibility burden.	2023-10-31 14:08:38 +02:00
Gleb Natapov	35a1ac1a9a	gossip: remove unused STATUS_MOVING state Moving operation was removed by `4a0b561376` and since then the state is unused. Even back then it worked only for the case of one token so it is safe to say we never used it. Lets remove the remains of the code instead of carrying it forever.	2023-10-31 13:54:46 +02:00
Gleb Natapov	f80fff3484	gossip: remove unused STATUS_LEAVING gossiper status The status is no longer used. The function that referenced it was removed by `5a96751534` and it was unused back then for awhile already. Message-Id: <ZS92mcGE9Ke5DfXB@scylladb.com>	2023-10-18 11:13:14 +02:00
Gleb Natapov	05aa07835d	storage_service: delete code that handled REMOVING_TOKENS state The state is never advertised so the code is never used.	2023-05-25 14:48:09 +03:00
Kefu Chai	c37f4e5252	treewide: use fmt::join() when appropriate now that fmtlib provides fmt::join(). see https://fmt.dev/latest/api.html#_CPPv4I0EN3fmt4joinE9join_viewIN6detail10iterator_tI5RangeEEN6detail10sentinel_tI5RangeEEERR5Range11string_view there is not need to revent the wheel. so in this change, the homebrew join() is replaced with fmt::join(). as fmt::join() returns an join_view(), this could improve the performance under certain circumstances where the fully materialized string is not needed. please note, the goal of this change is to use fmt::join(), and this change does not intend to improve the performance of existing implementation based on "operator<<" unless the new implementation is much more complicated. we will address the unnecessarily materialized strings in a follow-up commit. some noteworthy things related to this change: * unlike the existing `join()`, `fmt::join()` returns a view. so we have to materialize the view if what we expect is a `sstring` * `fmt::format()` does not accept a view, so we cannot pass the return value of `fmt::join()` to `fmt::format()` * fmtlib does not format a typed pointer, i.e., it does not format, for instance, a `const std::string`. but operator<<() always print a typed pointer. so if we want to format a typed pointer, we either need to cast the pointer to `void` or use `fmt::ptr()`. * fmtlib is not able to pick up the overload of `operator<<(std::ostream& os, const column_definition* cd)`, so we have to use a wrapper class of `maybe_column_definition` for printing a pointer to `column_definition`. since the overload is only used by the two overloads of `statement_restrictions::add_single_column_parition_key_restriction()`, the operator<< for `const column_definition*` is dropped. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2023-03-16 20:34:18 +08:00
Avi Kivity	fcb8d040e8	treewide: use Software Package Data Exchange (SPDX) license identifiers Instead of lengthy blurbs, switch to single-line, machine-readable standardized (https://spdx.dev) license identifiers. The Linux kernel switched long ago, so there is strong precedent. Three cases are handled: AGPL-only, Apache-only, and dual licensed. For the latter case, I chose (AGPL-3.0-or-later and Apache-2.0), reasoning that our changes are extensive enough to apply our license. The changes we applied mechanically with a script, except to licenses/README.md. Closes #9937	2022-01-18 12:15:18 +01:00
Avi Kivity	a55b434a2b	treewide: extent copyright statements to present day	2021-06-06 19:18:49 +03:00
Kamil Braun	4658adbe18	tree-wide: introduce cdc::generation_id_v2 This is a new type of CDC generation identifiers. Compared to old IDs, additionally to the timestamp it contains an UUID. These new identifiers will allow a safer and more efficient algorithm of introducing new generations into a cluster (introduced in a later commit). For now, nodes keep using the old identifier format when creating new generations and whenever they learn about a new CDC generation from gossip they assume that it also is stored in the v1 format. But they do know how to (de)serialize the second format and how to persist new identifiers in local tables.	2021-05-24 17:50:21 +02:00
Pavel Solodovnikov	fff7ef1fc2	treewide: reduce boost headers usage in scylla header files `dev-headers` target is also ensured to build successfully. Signed-off-by: Pavel Solodovnikov <pa.solodovnikov@scylladb.com>	2021-05-20 01:33:18 +03:00
Kamil Braun	99fd2244a3	tree-wide: introduce cdc::generation_id type This is a follow-up to the previous commit. Each CDC generation has a timestamp which denotes a logical point in time when this generation starts operating. That same timestamp is used to identify the CDC generation. We use this identification scheme to exchange CDC generations around the cluster. However, the fact that a generation's timestamp is used as an ID for this generation is an implementation detail of the currently used method of managing CDC generations. Places in the code that deal with the timestamp, e.g. functions which take it as an argument (such as handle_cdc_generation) are often interested in the ID aspect, not the "when does the generation start operating" aspect. They don't care that the ID is a `db_clock::time_point`. They may sometimes want to retrieve the time point given the ID (such as do_handle_cdc_generation when it calls `cdc::metadata::insert`), but they don't care about the fact that the time point actually IS the ID. In the future we may actually change the specific type of the ID if we modify the generation management algorithms. This commit is an intermediate step that will ease the transition in the future. It introduces a new type, `cdc::generation_id`. Inside it contains the timestamp, so: 1. if a piece of code doesn't care about the timestamp, it just passes the ID around 2. if it does care, it can simply access it using the `get_ts` function. The fact that `get_ts` simply accesses the ID's only field is an implementation detail. Using the occasion, we change the `do_handle_cdc_generation_intercept...` function to be a standard function, not a coroutine. It turns out that - depending on the shape of the passed-in argument - the function would sometimes miscompile (the compiled code would not copy the argument to the coroutine frame).	2021-04-07 13:47:13 +02:00
Kamil Braun	e486e0f759	tree-wide: rename "cdc streams timestamp" to "cdc generation id" Each CDC generation always has a timestamp, but the fact that the timestamp identifies the generation is an implementation detail. We abstract away from this detail by using a more generic naming scheme: a generation "identifier" (whatever that is - a timestamp or something else). It's possible that a CDC generation will be identified by more than a timestamp in the (near) future. The actual string gossiped by nodes in their application state is left as "CDC_STREAMS_TIMESTAMP" for backward compatibility. Some stale comments have been updated.	2021-04-06 13:15:31 +02:00
Benny Halevy	68a2920201	gms/versioned_value: mark trivial methods noexcept Also, versioned_value::compare_to() can be marked const. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2020-11-01 16:46:18 +02:00
Kamil Braun	1f7290a0ff	versioned_value: remove versioned_value::factory class If there was a Most Useless Abstraction award, this would be a good candidate.	2020-04-20 12:57:16 +02:00
Kamil Braun	113384b6f8	gms: move TOKENS string deserialization code into versioned_value And do the same with CDC_STREAMS_TIMESTAMP. The code that took a list of tokens represented as a string inside versioned_value (for gossiping) and deserialized it into an `unordered_set<dht::token>` lived in the storage_service module, while the code that did the serializing (set -> string) lived in versioned_value. There was a similar situation with the CDC generation timestamp. To increase maintanability and reusability, the deserialization code is now placed next to the serialization code in versioned_value. Furthermore, the `make_full_token_string`, `make_token_string`, and `make_cdc_streams_timestamp_string` (serialization functions) are moved out of versioned_value::factory and made static methods of versioned_value instead.	2020-04-20 12:57:13 +02:00
Asias He	343986a70b	gossiper: Introduce gossip STATUS_UNKNOWN When a node does not have gossip STATUS application_state, we currently use an empty string to present such state in get_gossip_status. It is better to use an explicit "UNKNOWN" to present it. It makes the log easier to understand when the status is unknown. Before: 'gossip - InetAddress n2 is now UP, status =' After: 'gossip - InetAddress n2 is now UP, status = UNKNOWN' This patch is safe because the STATUS_UNKNOWN is never sent over the cluster. So the presentation is only internal to the node. Fixes #5520	2020-01-20 10:59:14 +02:00
Duarte Nunes	e46ef6723b	Merge seastar upstream * seastar d152f2d...c1e0e5d (6): > scripts: perftune.py: properly merge parameters from the command line and the configuration file > fmt: update to 5.2.1 > io_queue: only increment statistics when request is admitted > Adds `read_first_line.cc` and `read_first_line.hh` to CMake. > fstream: remove default extent allocation hint > core/semaphore: Change the access of semaphore_units main ctor Due to a compile-time fight between fmt and boost::multiprecision, a lexical_cast was added to mediate. sprint("%s", var) no longer accepts numeric values, so some sprint()s were converted to format() calls. Since more may be lurking we'll need to remove all sprint() calls. Signed-off-by: Duarte Nunes <duarte@scylladb.com>	2018-10-25 12:53:30 +03:00
Avi Kivity	ebaeefa02b	Merge seatar upstream (seastar namespace) - introcduced "seastarx.hh" header, which does a "using namespace seastar"; - 'net' namespace conflicts with seastar::net, renamed to 'netw'. - 'transport' namespace conflicts with seastar::transport, renamed to cql_transport. - "logger" global variables now conflict with logger global type, renamed to xlogger. - other minor changes	2017-05-21 12:26:15 +03:00
Pekka Enberg	38a54df863	Fix pre-ScyllaDB copyright statements People keep tripping over the old copyrights and copy-pasting them to new files. Search and replace "Cloudius Systems" with "ScyllaDB". Message-Id: <1460013664-25966-1-git-send-email-penberg@scylladb.com>	2016-04-08 08:12:47 +03:00
Asias He	d7c7994f37	gossip: Drop unused serialization code - versioned_value	2016-01-25 11:28:29 +08:00
Amnon Heiman	8a4d211a99	Changes the versioned_value to make serializeble This patch contains two changes, it make the constructor with parameters public. And it removes the dependency in messaging_service.hh from the header file by moving some of the code to the .cc file. Signed-off-by: Amnon Heiman <amnon@scylladb.com>	2016-01-24 12:13:01 +02:00
Asias He	f62a6f234b	gossip: Add shutdown gossip state Backported: CASSANDRA-8336 and CASSANDRA-9871 84b2846 remove redundant state b2c62bb Add shutdown gossip state to prevent timeouts during rolling restarts 8f9ca07 Cannot replace token does not exist - DN node removed as Fat Client Fixes: When X is shutdown, X sends SHUTDOWN message to both Y and Z, but for some reason, only Y receives the message and Z does not receive the message. If Z has a higher gossip version for X than Y has for X, Z will initiate a gossip with Y and Y will mark X alive again. X ------> Y \ / \ / Z	2015-12-01 17:29:25 +08:00
Asias He	6b7ca7e334	gossip: Cleanup versioned_value class	2015-10-23 10:06:19 +08:00
Avi Kivity	d5cf0fb2b1	Add license notices	2015-09-20 10:43:39 +03:00
Asias He	7ac5277822	gossip: Fix link error with versioned_value With the next patch "gossip: Add storage_service_value_factory helper" in this series. [asias@hjpc urchin]$ ninja-build [8/10] LINK build/release/seastar build/release/gms/gossiper.o: In function `gms::versioned_value::versioned_value_factory::removing_nonlocal(utils::UUID const&)': /home/asias/src/cloudius-systems/urchin/./gms/versioned_value.hh:201: undefined reference to `gms::versioned_value::DELIMITER_STR' build/release/gms/gossiper.o: In function `gms::versioned_value::versioned_value_factory::removal_coordinator(utils::UUID const&)': /home/asias/src/cloudius-systems/urchin/./gms/versioned_value.hh:211: undefined reference to `gms::versioned_value::DELIMITER_STR' build/release/gms/gossiper.o: In function `gms::versioned_value::versioned_value_factory::removed_nonlocal(utils::UUID const&, long)': /home/asias/src/cloudius-systems/urchin/./gms/versioned_value.hh:206: undefined reference to `gms::versioned_value::DELIMITER_STR' /home/asias/src/cloudius-systems/urchin/./gms/versioned_value.hh:206: undefined reference to `gms::versioned_value::DELIMITER_STR' collect2: error: ld returned 1 exit status Fix by defining the symbol in gms/versioned_value.cc.	2015-05-07 15:29:13 +08:00

29 Commits