scylladb

mirror of https://github.com/scylladb/scylladb.git synced 2026-04-23 18:10:39 +00:00

Author	SHA1	Message	Date
Asias He	d3c6e72c69	repair: Allow abort repair jobs in early stage Consider this: - User starts a repair job with http api - User aborts all repair - The repair_info object for the repair job is created - The repair job is not aborted In this patch, the repair uuid is recorded before repair_info object is created, so that repair can now abort repair jobs in the early stage. Fixes #10384 Closes #10428	2022-06-27 16:39:36 +03:00
Avi Kivity	3131cbea62	Merge 'query: allow replica to provide arbitrary continue position' from Botond Dénes Currently, we use the last row in the query result set as the position where the query is continued from on the next page. Since only live rows make it into query result set, this mandates the query to be stopped on a live row on the replica, lest any dead rows or tombstones processed after the live rows, would have to be re-processed on the next page (and the saved reader would have to be thrown away due to position mismatch). This requirement of having to stop on a live row is problematic with datasets which have lots of dead rows or tombstones, especially if these form a prefix. In the extreme case, a query can time out before it can process a single live row and the data-set becomes effectively unreadable until compaction gets rid of the tombstones. This series prepares the way for the solution: it allows the replica to determine what position the query should continue from on the next page. This position can be that of a dead row, if the query stopped on a dead row. For now, the replica supplies the same position that would have been obtained with looking at the last row in the result set, this series merely introduces the infrastructure for transferring a position together with the query result, and it prepares the paging logic to make use of this position. If the coordinator is not prepared for the new field, it will simply fall-back to the old way of looking at the last row in the result set. As I said for now this is still the same as the content of the new field so there is no problem in mixed clusters. Refs: https://github.com/scylladb/scylla/issues/3672 Refs: https://github.com/scylladb/scylla/issues/7689 Refs: https://github.com/scylladb/scylla/issues/7933 Tests: manual upgrade test. I wrote a data set with: ``` ./scylla-bench -mode=write -workload=sequential -replication-factor=3 -nodes 127.0.0.1,127.0.0.2,127.0.0.3 -clustering-row-count=10000 -clustering-row-size=8096 -partition-count=1000 ``` This creates large, 80MB partitions, which should fill many pages if read in full. Then I started a read workload: ``` ./scylla-bench -mode=read -workload=uniform -replication-factor=3 -nodes 127.0.0.1,127.0.0.2,127.0.0.3 -clustering-row-count=10000 -duration=10m -rows-per-request=9000 -page-size=100 ``` I confirmed that paging is happening as expected, then upgraded the nodes one-by-one to this PR (while the read-load was ongoing). I observed no read errors or any other errors in the logs. Closes #10829 * github.com:scylladb/scylla: query: have replica provide the last position idl/query: add last_position to query_result mutlishard_mutation_query: propagate compaction state to result builder multishard_mutation_query: defer creating result builder until needed querier: use full_position instead of ad-hoc struct querier: rely on compactor for position tracking mutation_compactor: add current_full_position() convenience accessor mutation_compactor: s/_last_clustering_pos/_last_pos/ mutation_compactor: add state accessor to compact_mutation introduce full_position idl: move position_in_partition into own header service/paging: use position_in_partition instead of clustering_key for last row alternator/serialization: extract value object parsing logic service/pagers/query_pagers.cc: fix indentation position_in_partition: add to_string(partition_region) and parse_partition_region() mutation_fragment.hh: move operator<<(partition_region) to position_in_partition.hh	2022-06-27 12:23:21 +03:00
Benny Halevy	9c231ad0ce	repair_reader: construct _reader_handle before _reader Currently, the `_reader` member is explicitly initialized with the result of the call to `make_reader`. And `make_reader`, as a side effect, assigns a value to the `_reader_handle` member. Since C++ initializes class members sequentially, in the order they are defined, the assignment to `_reader_handle` in `make_reader()` happens before `_reader_handle` is initialized. This patch fixes that by changing the definition order, and consequently, the member initialization order in the constructor so that `_reader_handle` will be (default-)initialized before the call to `make_reader()`, avoiding the undefined behavior. Fixes #10882 Signed-off-by: Benny Halevy <bhalevy@scylladb.com> Closes #10883	2022-06-26 20:17:47 +03:00
Botond Dénes	119be5d5db	idl: move position_in_partition into own header So it can be used without pulling in all of partition_checksum.idl.hh.	2022-06-23 13:36:24 +03:00
Pavel Emelyanov	b28db0294c	repair: Get rack/datacenter from topology Repair gets token metadata from its local database reference. Not perfect, repair should better have its own private token meta reference, but it's OK for now. The change obsoletes static get_local_dc helper. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-06-22 11:47:26 +03:00
Botond Dénes	74587cb1b5	Merge 'Fix stalls during repair with rbno' from Asias He Found with scylla --blocked-reactor-notify-ms 1 during replace operation with rbno turned on. The stalls showed without this patch were gone after this path set. Closes #10737 * github.com:scylladb/scylla: repair: Avoid stall in working_row_hashes repair: Avoid stall in apply_rows_on_master_in_thread	2022-06-08 06:41:51 +03:00
Asias He	f2c05e21ee	repair: Avoid stall in working_row_hashes Fix the following stall during repair: ``` Reactor stalled for 1 ms on shard 0. Backtrace: [Backtrace #11] {build/release/scylla} 0x4c6deb2: void seastar::backtrace<seastar::backtrace_buffer::append_backtrace_oneline()::{lambda(seastar::frame)#1}>(seastar::backtrace_buffer::append_backtrace_oneline()::{lambda(seastar::frame)#1}&&) at ./ (inlined by) seastar::backtrace_buffer::append_backtrace_oneline() at ./build/release/seastar/./seastar/src/core/reactor.cc:772 (inlined by) seastar::print_with_backtrace(seastar::backtrace_buffer&, bool) at ./build/release/seastar/./seastar/src/core/reactor.cc:791 {build/release/scylla} 0x4c6cb10: seastar::internal::cpu_stall_detector::generate_trace() at ./build/release/seastar/./seastar/src/core/reactor.cc:1366 {build/release/scylla} 0x4c6ddc0: seastar::internal::cpu_stall_detector::maybe_report() at ./build/release/seastar/./seastar/src/core/reactor.cc:1108 (inlined by) seastar::internal::cpu_stall_detector::on_signal() at ./build/release/seastar/./seastar/src/core/reactor.cc:1125 (inlined by) seastar::reactor::block_notifier(int) at ./build/release/seastar/./seastar/src/core/reactor.cc:1349 {build/release/scylla} 0x7f75551bfa1f: ?? ??:0 {build/release/scylla} 0x37abf12: repair_hash::operator<(repair_hash const&) const at ././repair/hash.hh:30 (inlined by) std::less<repair_hash>::operator()(repair_hash const&, repair_hash const&) const at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/bits/stl_function.h:400 (inlined by) bool absl::container_internal::key_compare_adapter<std::less<repair_hash>, repair_hash>::checked_compare::operator()<repair_hash, repair_hash, 0>(repair_hash const&, repair_hash const&) const at ./abseil/absl/containe (inlined by) absl::container_internal::SearchResult<int, false> absl::container_internal::btree_node<absl::container_internal::set_params<repair_hash, std::less<repair_hash>, std::allocator<repair_hash>, 256, false> >::binary_sear (inlined by) _ZNK4absl18container_internal10btree_nodeINS0_10set_paramsI11repair_hashSt4lessIS3_ESaIS3_ELi256ELb0EEEE13binary_searchIS3_NS0_19key_compare_adapterIS5_S3_E15checked_compareEEENS0_12SearchResultIiXsr23btree_is_key_com (inlined by) absl::container_internal::SearchResult<int, false> absl::container_internal::btree_node<absl::container_internal::set_params<repair_hash, std::less<repair_hash>, std::allocator<repair_hash>, 256, false> >::lower_bound (inlined by) absl::container_internal::SearchResult<absl::container_internal::btree_iterator<absl::container_internal::btree_node<absl::container_internal::set_params<repair_hash, std::less<repair_hash>, std::allocator<repair_hash (inlined by) std::pair<absl::container_internal::btree_iterator<absl::container_internal::btree_node<absl::container_internal::set_params<repair_hash, std::less<repair_hash>, std::allocator<repair_hash>, 256, false> >, repair_hash (inlined by) std::pair<absl::container_internal::btree_iterator<absl::container_internal::btree_node<absl::container_internal::set_params<repair_hash, std::less<repair_hash>, std::allocator<repair_hash>, 256, false> >, repair_hash (inlined by) operator() at ./repair/row_level.cc:896 (inlined by) seastar::future<void> seastar::futurize<void>::invoke<repair_meta::working_row_hashes()::{lambda(absl::btree_set<repair_hash, std::less<repair_hash>, std::allocator<repair_hash> >&)#1}::operator()(absl::btree_set<repa (inlined by) auto seastar::futurize_invoke<repair_meta::working_row_hashes()::{lambda(absl::btree_set<repair_hash, std::less<repair_hash>, std::allocator<repair_hash> >&)#1}::operator()(absl::btree_set<repair_hash, std::less<repai{build/release/scylla} 0x37ac70f: seastar::internal::do_for_each_state<std::_List_iterator<repair_row>, repair_meta::working_row_hashes()::{lambda(absl::btree_set<repair_hash, std::less<repair_hash>, std::allocator<repair_hash> >&) {build/release/scylla} 0x4c7ee64: seastar::reactor::run_tasks(seastar::reactor::task_queue&) at ./build/release/seastar/./seastar/src/core/reactor.cc:2356 (inlined by) seastar::reactor::run_some_tasks() at ./build/release/seastar/./seastar/src/core/reactor.cc:2769 {build/release/scylla} 0x4c80247: seastar::reactor::do_run() at ./build/release/seastar/./seastar/src/core/reactor.cc:2938 {build/release/scylla} 0x4c7f49c: seastar::reactor::run() at ./build/release/seastar/./seastar/src/core/reactor.cc:2821 {build/release/scylla} 0x4c264d8: seastar::app_template::run_deprecated(int, char, std::function<void ()>&&) at ./build/release/seastar/./seastar/src/core/app-template.cc:265 {build/release/scylla} 0x4c259b1: seastar::app_template::run(int, char, std::function<seastar::future<int> ()>&&) at ./build/release/seastar/./seastar/src/core/app-template.cc:156 {build/release/scylla} 0xf5c16f: scylla_main(int, char) at ./main.cc:535 {build/release/scylla} 0xf5999a: std::function<int (int, char)>::operator()(int, char**) const at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/bits/std_function.h:590 (inlined by) main at ./main.cc:1575 {build/release/scylla} 0x27b74: ?? ??:0 {build/release/scylla} 0xf5892d: _start at ??:? ``` Found with scylla --blocked-reactor-notify-ms 1 Refs #10665	2022-06-07 16:04:50 +08:00
Asias He	45bcacf672	repair: Avoid stall in apply_rows_on_master_in_thread Fix the following stall during repair: ``` Reactor stalled for 3 ms on shard 0. Backtrace: [Backtrace #20] {build/release/scylla} 0x4c6deb2: void seastar::backtrace<seastar::backtrace_buffer::append_backtrace_oneline()::{lambda(seastar::frame)#1}>(seastar::backtrace_buffer::append_backtrace_oneline()::{lambda(seastar::frame)#1}&&) at ./build/release/seastar/./seastar/include/seastar/util/backtrace.hh:59 (inlined by) seastar::backtrace_buffer::append_backtrace_oneline() at ./build/release/seastar/./seastar/src/core/reactor.cc:772 (inlined by) seastar::print_with_backtrace(seastar::backtrace_buffer&, bool) at ./build/release/seastar/./seastar/src/core/reactor.cc:791 {build/release/scylla} 0x4c6cb10: seastar::internal::cpu_stall_detector::generate_trace() at ./build/release/seastar/./seastar/src/core/reactor.cc:1366 {build/release/scylla} 0x4c6ddc0: seastar::internal::cpu_stall_detector::maybe_report() at ./build/release/seastar/./seastar/src/core/reactor.cc:1108 (inlined by) seastar::internal::cpu_stall_detector::on_signal() at ./build/release/seastar/./seastar/src/core/reactor.cc:1125 (inlined by) seastar::reactor::block_notifier(int) at ./build/release/seastar/./seastar/src/core/reactor.cc:1349 {build/release/scylla} 0x7f75551bfa1f: ?? ??:0 {build/release/scylla} 0x11293e9: std::default_delete<bytes_ostream::chunk>::operator()(bytes_ostream::chunk) const at database.cc:? (inlined by) std::default_delete<bytes_ostream::chunk>::operator()(bytes_ostream::chunk) const at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/bits/unique_ptr.h:85 {build/release/scylla} 0x37b18e6: ~unique_ptr at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/bits/unique_ptr.h:361 (inlined by) ~bytes_ostream at ././bytes_ostream.hh:26 (inlined by) ~frozen_mutation_fragment at ././frozen_mutation.hh:265 (inlined by) std::_Optional_payload_base<frozen_mutation_fragment>::_M_destroy() at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/optional:260 (inlined by) std::_Optional_payload_base<frozen_mutation_fragment>::_M_reset() at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/optional:280 (inlined by) ~_Optional_payload at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/optional:401 (inlined by) ~_Optional_base at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/optional:472 (inlined by) ~repair_row at ././repair/row.hh:24 (inlined by) void std::destroy_at<repair_row>(repair_row) at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/bits/stl_construct.h:88 (inlined by) void std::allocator_traits<std::allocator<std::_List_node<repair_row> > >::destroy<repair_row>(std::allocator<std::_List_node<repair_row> >&, repair_row) at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/bits/alloc_traits.h:537 (inlined by) std::__cxx11::_List_base<repair_row, std::allocator<repair_row> >::_M_clear() at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/bits/list.tcc:77 (inlined by) ~_List_base at /usr/lib/gcc/x86_64-redhat-linux/11/../../../../include/c++/11/bits/stl_list.h:499 (inlined by) repair_meta::apply_rows_on_master_in_thread(std::__cxx11::list<partition_key_and_mutation_fragments, std::allocator<partition_key_and_mutation_fragments> >, gms::inet_address, seastar::bool_class<update_working_row_buf_tag>, seastar::bool_class<update_peer_row_hash_sets_tag>, unsigned int) at ./repair/row_level.cc:1273 {build/release/scylla} 0x37ad9dc: repair_meta::get_row_diff_source_op(seastar::bool_class<update_peer_row_hash_sets_tag>, gms::inet_address, unsigned int, seastar::rpc::sink<repair_hash_with_cmd>&, seastar::rpc::source<repair_row_on_wire_with_cmd>&) at ./repair/row_level.cc:1617 {build/release/scylla} 0x37a2982: repair_meta::get_row_diff_with_rpc_stream(absl::btree_set<repair_hash, std::less<repair_hash>, std::allocator<repair_hash> >, seastar::bool_class<needs_all_rows_tag>, seastar::bool_class<update_peer_row_hash_sets_tag>, gms::inet_address, unsigned int) at ./repair/row_level.cc:1683 ``` Found with scylla --blocked-reactor-notify-ms 1 Refs #10665	2022-06-07 16:04:50 +08:00
Michael Livshin	029508b77c	flat_mutation_reader ist tot Signed-off-by: Michael Livshin <michael.livshin@scylladb.com>	2022-05-31 23:42:34 +03:00
Michael Livshin	7a11a22cd6	repair/row_level: mutation_fragment_v1_stream() instead of downgrade_to_v1() Signed-off-by: Michael Livshin <michael.livshin@scylladb.com>	2022-05-31 23:42:34 +03:00
Avi Kivity	4b53af0bd5	treewide: replace parallel_for_each with coroutine::parallel_for_each in coroutines coroutine::parallel_for_each avoids an allocation and is therefore preferred. The lifetime of the function object is less ambiguous, and so it is safer. Replace all eligible occurences (i.e. caller is a coroutine). One case (storage_service::node_ops_cmd_heartbeat_updater()) needed a little extra attention since there was a handle_exception() continuation attached. It is converted to a try/catch. Closes #10699	2022-05-31 09:06:24 +03:00
Avi Kivity	e65b3ed50a	Merge 'Allow trigger off strategy compaction early for node operations' from Asias He This patch set adds two commits to allow trigger off strategy early for node operations. ) repair: Repair table by table internally This patch changes the way a repair job walks through tables and ranges if multiple tables and ranges are requested by users. Before: ``` for range in ranges for table in tables repair(range, table) ``` After: ``` for table in tables for range in ranges repair(range, table) ``` The motivation for this change is to allow off-strategy compaction to trigger early, as soon as a table is finished. This allows to reduce the number of temporary sstables on disk. For example, if there are 50 tables and 256 ranges to repair, each range will generate one sstable. Before this change, there will be 50 256 sstables on disk before off-strategy compaction triggers. After this change, once a table is finished, off-strategy compaction can compact the 256 sstables. As a result, this would reduce the number of sstables by 50X. This is very useful for repair based node operations since multiple ranges and tables can be requested in a single repair job. Refs: #10462 ) repair: Trigger off strategy compaction after all ranges of a table is repaired When the repair reason is not repair, which means the repair reason is node operations (bootstrap, replace and so on), a single repair job contains all the ranges of a table that need to be repaired. To trigger off strategy compaction early and reduce the number of temporary sstable files on disk, we can trigger the compaction as soon as a table is finished. Refs: #10462 Closes #10551 github.com:scylladb/scylla: repair: Trigger off strategy compaction after all ranges of a table is repaired repair: Repair table by table internally	2022-05-23 18:58:21 +03:00
Avi Kivity	528ab5a502	treewide: change metric calls from make_derive to make_counter make_derive was recently deprecated in favor of make_counter, so make the change throughput the codebase. Closes #10564	2022-05-14 12:53:55 +02:00
Asias He	7a38b806be	repair: Trigger off strategy compaction after all ranges of a table is repaired When the repair reason is not repair, which means the repair reason is node operations (bootstrap, replace and so on), a single repair job contains all the ranges of a table that need to be repaired. To trigger off strategy compaction early and reduce the number of temporary sstable files on disk, we can trigger the compaction as soon as a table is finished. Refs: #10462	2022-05-12 10:46:11 +08:00
Asias He	3dc9a81d02	repair: Repair table by table internally This patch changes the way a repair job walks through tables and ranges if multiple tables and ranges are requested by users. Before: ``` for range in ranges for table in tables repair(range, table) ``` After: ``` for table in tables for range in ranges repair(range, table) ``` The motivation for this change is to allow off-strategy compaction to trigger early, as soon as a table is finished. This allows to reduce the number of temporary sstables on disk. For example, if there are 50 tables and 256 ranges to repair, each range will generate one sstable. Before this change, there will be 50 * 256 sstables on disk before off-strategy compaction triggers. After this change, once a table is finished, off-strategy compaction can compact the 256 sstables. As a result, this would reduce the number of sstables by 50X. This is very useful for repair based node operations since multiple ranges and tables can be requested in a single repair job. Refs: #10462	2022-05-12 10:46:11 +08:00
Pavel Emelyanov	598ce8111d	repair: Handle discarded stopping future When repair_meta stops it does so in the background and reports back a shared future into whose shared promise peer it resolves that background activity. There's a shorter way to forward a future result into another, even shared, promise. And this method doesn't need to discard a future. tests: https://jenkins.scylladb.com/job/releng/job/Scylla-CI/253 Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-05-09 17:23:12 +03:00
Avi Kivity	19ab3edd77	gms: feature_service: remove variable/helper function duplication Each feature has a private variable and a public accessor. Since the accessor effectively makes the variable public, avoid the intermediary and make the variable public directly. To ease mechanical translation, the variable name is chosen as the function name (without the cluster_supports_ prefix). References throughout the codebase are adjusted.	2022-05-04 18:59:56 +03:00
Pavel Solodovnikov	47834313d8	repair: avoid infinite recursion on stringifying unknown node_ops_cmd Cast the cmd representation to underlying type and avoid infinite recursion in the `operator <<(node_ops_cmd)`. Tests: unit(dev) Signed-off-by: Pavel Solodovnikov <pa.solodovnikov@scylladb.com> Message-Id: <20220430180701.1012190-1-pa.solodovnikov@scylladb.com>	2022-05-02 10:08:34 +03:00
Botond Dénes	f527956cdb	readers: remove v1 empty_reader The only user is row level repair: it is replaced with downgrade_to_v1(make_empty_flat_reader_v2()). The row level reader has lots of downgrade_to_v1() calls, we will deal with these later all at once. Another use is the empty mutation source, this is trivially converted to use the v2 variant.	2022-04-28 14:12:24 +03:00
Botond Dénes	b061acb668	Merge 'Remove queue reader v1' from Mikołaj Sielużycki The patchset embeds the mutation_fragment upgrading logic from v1 to v2 into the mutation_fragment_queue. This way the mutation fragments coming to the mutation_fragment_queue can be v1, but the underlying query_reader receives mutation_fragment_v2, eliminating the last usage of query_reader (v1). The last commit removes query_reader, query_reader_handle and associated factory functions. tests: unit(dev), dtest(incremental_repair_test, read_repair_test, repair_additional_test, repair_test) Closes #10371 * github.com:scylladb/scylla: readers: Remove queue_reader v1 and associated code. repair: Make mutation_fragment_queue internally upgrade fragments to v2 repair: Make mutation_fragment_queue::impl a seastar::shared_ptr	2022-04-21 12:34:48 +03:00
Mikołaj Sielużycki	339b60e5b0	repair: Make mutation_fragment_queue internally upgrade fragments to v2	2022-04-20 17:55:58 +02:00
Mikołaj Sielużycki	eeb2b458de	repair: Make mutation_fragment_queue::impl a seastar::shared_ptr It makes mutation_fragment_queue copyable and makes the pointer to pending mutation fragments in next commit stable. This allows moving the mutation_fragment_queue without breaking the underlying upgrading_consumer.	2022-04-20 17:51:58 +02:00
Avi Kivity	5da586271f	repair: explicityl ignore tombstone gc update response The response struct is empty and we have nothing to do with it. Cast it to void to avoid a gcc warning.	2022-04-18 12:27:18 +03:00
Botond Dénes	75786c42cb	Merge 'Add repair unit tests/v1' from Mikołaj Sielużycki This patch series splits up parts of repair pipeline to allow unit testing various bits of code without having to run full dtest suite. The reason why repair pipeline has no unit tests is that by definition repair requires multiple nodes, while unit test environment works only for a single node. However, it is possible to explicitly define interfaces between various parts of the pipeline, inject dependencies and test them individually. This patch series is focused on taking repair_rows_on_wire (frozen mutation representation of changes coming from another node) and flushing them to an sstable. The commits are split into the following parts: - pulling out classes to separate headers so that they can be included (potentially indirectly) from the test, - pulling out repair_meta::to_repair_rows_list and part of repair_meta::flush_rows_in_working_row_buf so that they can be tested, - refactoring repair_writer so that the actual writing logic can be injected as dependency, - creating the unit test. tests: unit(dev), dtest(incremental_repair_test, read_repair_test, repair_additional_test, repair_test) Closes #10345 * github.com:scylladb/scylla: repair: Add unit test for flushing repair_rows_on_wire to disk. repair: Extract mutation_fragment_queue and repair_writer::impl interfaces. repair: Make parts of repair_writer interface private. repair: Rename inputs to flush_rows. repair: Make repair_meta::flush_rows a free function. repair: Split flush_rows_in_working_row_buf to two functions and make one static. repair: Rename inputs to to_repair_rows_list. repair: Make to_repair_rows_list a free function. repair: Make repair_meta::to_repair_rows_list a static function repair: Fix indentation in repair_writer. repair: Move repair_writer to separate header. repair: Move repair_row to a separate header. repair: Move repair_sync_boundary to a separate header. repair: Move decorated_key_with_hash to separate header. repair: Move row_repair hashing logic to separate class and file.	2022-04-14 18:17:03 +03:00
Pavel Emelyanov	05eb9c9416	repair, system_keyspace: Query repair_history with a helper Querying the table is now done with the help of qctx directly. This patch replaces it with a querying helper that calls the consumer function with the entry struct as the argument. After this change repair code can stop including query_context and mess with untyped_result_set. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-04-12 14:04:21 +03:00
Pavel Emelyanov	59f4aa0934	repair: Update loader code to use system_keyspace entry Patch the history entry loader to use the recently introduced history entry. This is just to reduce the churn in the next patch Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-04-12 13:59:55 +03:00
Pavel Emelyanov	9940016e05	repair, system_keyspace: Update repair_history with a helper Current code works directly on the qctx which is not nice. Instead, make it use the system keyspace reference. To make it work, the patch adds a helper method and introduces a helper struct for the table entry. This struct will also be used to query the table (next patch). Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-04-12 13:57:57 +03:00
Pavel Emelyanov	e501ebd6c2	repair: Keep system keyspace reference Repair updates (and queries on start) the system.repair_history table and thus depends on the system_keyspace object Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2022-04-12 13:57:08 +03:00
Mikołaj Sielużycki	39205917a8	repair: Extract mutation_fragment_queue and repair_writer::impl interfaces.	2022-04-12 09:22:03 +02:00
Mikołaj Sielużycki	a52126d861	repair: Make parts of repair_writer interface private.	2022-04-12 09:20:14 +02:00
Mikołaj Sielużycki	826e0e9d8a	repair: Rename inputs to flush_rows.	2022-04-12 09:20:14 +02:00
Mikołaj Sielużycki	4dd32064a3	repair: Make repair_meta::flush_rows a free function.	2022-04-12 09:20:14 +02:00
Mikołaj Sielużycki	046e8c31db	repair: Split flush_rows_in_working_row_buf to two functions and make one static. It allows pulling out the logic of writing internal representation of repair mutations to disk. This in turn is needed to unit test this functionality without spinning up clusters, which significantly improves developer iteration time.	2022-04-12 09:20:14 +02:00
Mikołaj Sielużycki	ca53a7fcc9	repair: Rename inputs to to_repair_rows_list.	2022-04-12 09:20:14 +02:00
Mikołaj Sielużycki	c7a7680c7d	repair: Make to_repair_rows_list a free function.	2022-04-12 09:20:14 +02:00
Mikołaj Sielużycki	69fc74ffbe	repair: Make repair_meta::to_repair_rows_list a static function It allows pulling out the logic of convering on-the-wire representation of repair mutations to an internal representation used later for flushing repair mutations to disk. This in turn is needed to unit test the functionality without spinning up clusters, which significantly improves developer iteration time.	2022-04-12 09:20:14 +02:00
Mikołaj Sielużycki	4ba48e5739	repair: Fix indentation in repair_writer.	2022-04-12 09:20:14 +02:00
Mikołaj Sielużycki	3ff738db6b	repair: Move repair_writer to separate header.	2022-04-12 09:20:03 +02:00
Mikołaj Sielużycki	04986e8c8e	repair: Move repair_row to a separate header.	2022-04-12 08:50:34 +02:00
Mikołaj Sielużycki	7b0cbdeac5	repair: Move repair_sync_boundary to a separate header.	2022-04-12 08:50:34 +02:00
Mikołaj Sielużycki	f9c75952ea	repair: Move decorated_key_with_hash to separate header.	2022-04-12 08:50:34 +02:00
Mikołaj Sielużycki	0fa703de3e	repair: Move row_repair hashing logic to separate class and file.	2022-04-12 08:50:34 +02:00
Botond Dénes	11c378a175	mutation_reader: move queue reader to readers/	2022-03-30 15:42:51 +03:00
Botond Dénes	d0ea895671	readers: move multishard reader & friends to reader/multishard.cc Since the multishard reader family weighs more than 1K SLOC, it gets its own .cc file.	2022-03-30 15:42:51 +03:00
Mikołaj Sielużycki	1d84a254c0	flat_mutation_reader: Split readers by file and remove unnecessary includes. The flat_mutation_reader files were conflated and contained multiple readers, which were not strictly necessary. Splitting optimizes both iterative compilation times, as touching rarely used readers doesn't recompile large chunks of codebase. Total compilation times are also improved, as the size of flat_mutation_reader.hh and flat_mutation_reader_v2.hh have been reduced and those files are included by many file in the codebase. With changes real 29m14.051s user 168m39.071s sys 5m13.443s Without changes real 30m36.203s user 175m43.354s sys 5m26.376s Closes #10194	2022-03-14 13:20:25 +02:00
Botond Dénes	ad1b157452	streaming/consumer: convert to v2 At least on the API level, internally there are still conversions, but these are going to be sorted out in the next patches too.	2022-03-02 09:55:09 +02:00
Asias He	ec59f7a079	repair: Do not flush hints and batchlog if tombstone_gc_mode is not repair The flush of hints and batchlog are needed only for the table with tombstone_gc_mode set to repair mode. We should skip the flush if the tombstone_gc_mode is not repair mode. Fixes #10004 Closes #10124	2022-02-25 07:26:11 +02:00
Asias He	680195564d	repair: Unify repair uuid report in the log More and more places are using the repair[uuid]: format for logging repair jobs with the uuid. Convert more places to use the new format to unify the log format. This makes it easier to grep a specific repair job in the log. Closes #10125	2022-02-23 09:13:12 +02:00
Botond Dénes	4aa9b90ba9	repair/row_level: use evictable reader v2	2022-02-21 12:29:24 +02:00
Benny Halevy	f8db9e1bd8	repair_service: deglobalize get_next_repair_meta_id Rather than using a static unit32_t next_id, move the next_id variable into repair_service shard 0 and manage it there. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2022-01-27 11:34:21 +02:00

1 2 3 4 5 ...

643 Commits