scylladb

mirror of https://github.com/scylladb/scylladb.git synced 2026-05-29 11:10:40 +00:00

Author	SHA1	Message	Date
Kefu Chai	afeff0a792	docs: explain task status retention and one-time query behavior Task status information from nodetool commands is not retained permanently: - Status of completed tasks is only kept for `task_ttl_in_seconds` - Status is removed after being queried, making it a one-time operation This behavior is important for users to understand since subsequent queries for the same completed task will not return any information. Add documentation to make this clear to users. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> Closes scylladb/scylladb#21386	2024-11-29 16:36:27 +01:00
Pavel Emelyanov	f2509d90a5	Merge 'mutation: remove unused "#include"s' from Kefu Chai these unused includes are identified by clang-include-cleaner. after auditing the source files, all of the reports have been confirmed. please note, because `mutation/mutation.hh` does not include `seastar/coroutine/maybe_yield.hh` anymore, and quite a few source files were relying on this header to bring in the declaration of `maybe_yield()`, we have to include this header in the places where this symbol is used. the same applies to `seastar/core/when_all.hh`. --- it's a cleanup, hence no need to backport. Closes scylladb/scylladb#21727 * github.com:scylladb/scylladb: .github: add "mutation" to CLEANER_DIR mutation: remove unused "#include"s	2024-11-29 13:01:53 +03:00
Kefu Chai	efbf6e5526	.github: add "mutation" to CLEANER_DIR in order to prevent future inclusion of unused headers, let's include "mutation" subdirectory to CLEANER_DIR, so that this workflow can identify the regressions in future. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2024-11-29 14:01:44 +08:00
Kefu Chai	f436edfa22	mutation: remove unused "#include"s these unused includes are identified by clang-include-cleaner. after auditing the source files, all of the reports have been confirmed. please note, because `mutation/mutation.hh` does not include `seastar/coroutine/maybe_yield.hh` anymore, and quite a few source files were relying on this header to bring in the declaration of `maybe_yield()`, we have to include this header in the places where this symbol is used. the same applies to `seastar/core/when_all.hh`. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2024-11-29 14:01:44 +08:00
Botond Dénes	055a36ae55	main: dump diagnostics on SIGQUIT Dump a diagnostics report on each shard when receiving a SIGQUIT. The report is logged with a dedicated logger, called diagnostics. The report has multiple parts: * seastar memory diagnostics, similar to that printed by the scylla memory command (from scylla-gdb.py). * reader concurrency semaphore diagnostics for each semaphore. Example report: INFO 2024-11-27 01:31:55,882 [shard 0:main] diagnostics - Diagnostics dump requested via SIGQUIT: Dumping seastar memory diagnostics Used memory: 3988M Free memory: 58M Total memory: 4G Hard failures: 0 LSA allocated: 4M used: 16 free: 4G Cache: total: 1M used: 642K free: 398K Memtables: total: 3M Regular: real dirty: 0B virt dirty: 0B System: real dirty: 3M virt dirty: 3M Replica: Read Concurrency Semaphores: user: 0/100, 0B/81M, queued: 0 streaming: 0/10, 0B/81M, queued: 0 system: 0/10, 0B/81M, queued: 0 compaction: 0/unlimited, 0B/unlimited view update: 0/50, 0B/40M, queued: 0 Execution Stages: apply stage: Total: 0 Tables - Ongoing Operations: Pending writes (top 10): 0 Total (all) Pending reads (top 10): 0 Total (all) Pending streams (top 10): 0 Total (all) Small pools: objsz spansz usedobj memory unused wst% 8 4K 858 16K 9K 58 10 4K 5 8K 8K 99 12 4K 5 8K 8K 99 14 4K 0 0B 0B 0 16 4K 2k 44K 15K 35 32 4K 4k 136K 16K 11 32 4K 8k 280K 24K 8 32 4K 3k 92K 6K 6 32 4K 4k 140K 21K 14 48 4K 3k 180K 25K 14 48 4K 2k 120K 27K 22 64 4K 2k 156K 18K 11 64 4K 19k 1M 11K 0 80 4K 3k 236K 16K 6 96 4K 6k 572K 49K 8 112 4K 2k 276K 72K 25 128 4K 477 80K 20K 25 160 4K 194 60K 30K 49 192 4K 1k 232K 39K 16 224 4K 2k 468K 15K 3 256 4K 182 100K 55K 54 320 8K 349 152K 43K 28 384 8K 332 288K 164K 56 448 4K 243 180K 74K 40 512 4K 256 244K 116K 47 640 16K 185 192K 76K 39 768 16K 394 432K 137K 31 896 8K 54 192K 144K 75 1024 4K 288 432K 144K 33 1280 32K 92 256K 140K 54 1536 32K 11 128K 111K 86 1792 16K 10 144K 126K 87 2048 8K 487 1M 90K 8 2560 64K 113 384K 100K 26 3072 64K 9 256K 228K 89 3584 32K 3 288K 277K 96 4096 16K 129 912K 396K 43 5120 128K 21 384K 275K 71 6144 128K 4 512K 486K 94 7168 64K 3 576K 553K 96 8192 32K 373 3M 56K 1 10240 64K 6 832K 770K 92 12288 64K 17 960K 756K 78 14336 128K 2 1M 1M 97 16384 64K 14 1M 992K 81 Page spans: index size free used spans 0 4K 4K 5M 1k 1 8K 8K 2M 213 2 16K 16K 2M 106 3 32K 64K 6M 200 4 64K 64K 4M 71 5 128K 384K 3934M 31k 6 256K 1M 256K 5 7 512K 512K 512K 2 8 1M 2M 0B 2 9 2M 2M 2M 2 10 4M 4M 0B 1 11 8M 16M 0B 2 12 16M 32M 0B 2 13 32M 0B 32M 1 14 64M 0B 0B 0 15 128M 0B 0B 0 16 256M 0B 0B 0 17 512M 0B 0B 0 18 1G 0B 0B 0 19 2G 0B 0B 0 20 4G 0B 0B 0 21 8G 0B 0B 0 22 16G 0B 0B 0 23 32G 0B 0B 0 24 64G 0B 0B 0 25 128G 0B 0B 0 26 256G 0B 0B 0 27 512G 0B 0B 0 28 1T 0B 0B 0 29 2T 0B 0B 0 30 4T 0B 0B 0 31 8T 0B 0B 0 INFO 2024-11-27 01:31:55,882 [shard 0:main] diagnostics - Diagnostics dump requested via SIGQUIT: Semaphore user with 0/100 count and 0/84850769 memory resources: user request, dumping permit diagnostics: permits count memory table/operation/state 0 0 0B total Stats: permit_based_evictions: 0 time_based_evictions: 0 inactive_reads: 0 total_successful_reads: 0 total_failed_reads: 0 total_reads_shed_due_to_overload: 0 total_reads_killed_due_to_kill_limit: 0 reads_admitted: 0 reads_enqueued_for_admission: 0 reads_enqueued_for_memory: 0 reads_admitted_immediately: 0 reads_queued_because_ready_list: 0 reads_queued_because_need_cpu_permits: 0 reads_queued_because_memory_resources: 0 reads_queued_because_count_resources: 0 reads_queued_with_eviction: 0 total_permits: 0 current_permits: 0 need_cpu_permits: 0 awaits_permits: 0 disk_reads: 0 sstables_read: 0 INFO 2024-11-27 01:31:55,882 [shard 0:main] diagnostics - Diagnostics dump requested via SIGQUIT: Semaphore streaming with 0/10 count and 0/84850769 memory resources: user request, dumping permit diagnostics: permits count memory table/operation/state 0 0 0B total Stats: permit_based_evictions: 0 time_based_evictions: 0 inactive_reads: 0 total_successful_reads: 6 total_failed_reads: 0 total_reads_shed_due_to_overload: 0 total_reads_killed_due_to_kill_limit: 0 reads_admitted: 6 reads_enqueued_for_admission: 0 reads_enqueued_for_memory: 0 reads_admitted_immediately: 6 reads_queued_because_ready_list: 0 reads_queued_because_need_cpu_permits: 0 reads_queued_because_memory_resources: 0 reads_queued_because_count_resources: 0 reads_queued_with_eviction: 0 total_permits: 6 current_permits: 0 need_cpu_permits: 0 awaits_permits: 0 disk_reads: 0 sstables_read: 0 INFO 2024-11-27 01:31:55,882 [shard 0:main] diagnostics - Diagnostics dump requested via SIGQUIT: Semaphore compaction with 0/2147483647 count and 0/9223372036854775807 memory resources: user request, dumping permit diagnostics: permits count memory table/operation/state 0 0 0B total Stats: permit_based_evictions: 0 time_based_evictions: 0 inactive_reads: 0 total_successful_reads: 0 total_failed_reads: 0 total_reads_shed_due_to_overload: 0 total_reads_killed_due_to_kill_limit: 0 reads_admitted: 0 reads_enqueued_for_admission: 0 reads_enqueued_for_memory: 0 reads_admitted_immediately: 0 reads_queued_because_ready_list: 0 reads_queued_because_need_cpu_permits: 0 reads_queued_because_memory_resources: 0 reads_queued_because_count_resources: 0 reads_queued_with_eviction: 0 total_permits: 27 current_permits: 0 need_cpu_permits: 0 awaits_permits: 0 disk_reads: 0 sstables_read: 0 INFO 2024-11-27 01:31:55,882 [shard 0:main] diagnostics - Diagnostics dump requested via SIGQUIT: Semaphore system with 0/10 count and 0/84850769 memory resources: user request, dumping permit diagnostics: permits count memory table/operation/state 1 0 0B ./view_builder/active 1 0 0B total Stats: permit_based_evictions: 0 time_based_evictions: 0 inactive_reads: 0 total_successful_reads: 234 total_failed_reads: 0 total_reads_shed_due_to_overload: 0 total_reads_killed_due_to_kill_limit: 0 reads_admitted: 234 reads_enqueued_for_admission: 154 reads_enqueued_for_memory: 0 reads_admitted_immediately: 80 reads_queued_because_ready_list: 154 reads_queued_because_need_cpu_permits: 0 reads_queued_because_memory_resources: 0 reads_queued_because_count_resources: 0 reads_queued_with_eviction: 0 total_permits: 235 current_permits: 1 need_cpu_permits: 0 awaits_permits: 0 disk_reads: 0 sstables_read: 0 INFO 2024-11-27 01:31:55,882 [shard 0:main] diagnostics - Diagnostics dump requested via SIGQUIT: Semaphore view_update with 0/50 count and 0/42425384 memory resources: user request, dumping permit diagnostics: permits count memory table/operation/state 0 0 0B total Stats: permit_based_evictions: 0 time_based_evictions: 0 inactive_reads: 0 total_successful_reads: 0 total_failed_reads: 0 total_reads_shed_due_to_overload: 0 total_reads_killed_due_to_kill_limit: 0 reads_admitted: 0 reads_enqueued_for_admission: 0 reads_enqueued_for_memory: 0 reads_admitted_immediately: 0 reads_queued_because_ready_list: 0 reads_queued_because_need_cpu_permits: 0 reads_queued_because_memory_resources: 0 reads_queued_because_count_resources: 0 reads_queued_with_eviction: 0 total_permits: 0 current_permits: 0 need_cpu_permits: 0 awaits_permits: 0 disk_reads: 0 sstables_read: 0 Fixes: scylladb/scylladb#7400 Closes scylladb/scylladb#21692	2024-11-28 18:52:29 +02:00
Botond Dénes	ff90a77f5b	scylla-sstable: revamp schema sources Demote --scylla-data-dir and --scylla-yaml-file to schema source helpers, rather than schema source in themselves. This practically means that when these options are used, they won't define where the tool will attempt to load the schema from, they will just be helpers to help locate the schema, for whichever schema source the tool was instructed to use (or left to choose). --scylla-data-dir and --scylla-yaml-file being schema sources were problematic with encryption at rest and for S3 support (not yet implemented). With encryption, the tool needs access to the configuration, so --scylla-yaml-file is often used to provide the path to the configuration file, which contains encryption configuration, needed for the tool to decrypt the sstable. Currently, using this option implies forcing the tool to read the schema from the schema tables, which is a problematic option for tests -- Scylla might be compacting a schema sstable and this will make the tool fail to load the schema. Demoting these options the schema helpers, allows providing them, while at the same time having the option to use a different schema-source. To allow the user to force the tool to load the schema from the schema tables, a new --schema-tables option is added. Similarly, a --sstable-schema option is introduced to force the tool to load the schema from the sstable itself. With this, each 4 schema source now has an option to force the use of said schema source. There are various helper options to be used along with these. The documentation as well as the tests are updated with the changes. The schema related documentation gets an rather extensive facelift because it was a bit out-of-date and incomplete. Fixes: scylladb/scylladb#20534 Closes scylladb/scylladb#21678	2024-11-28 18:36:09 +02:00
Kefu Chai	2c9c654798	build: cmake: Enforce explicit library linkage visibility This change improves dependency management by explicitly specifying library linkage visibility in CMake targets. Previously, some ScyllaDB targets used `target_link_libraries()` without `PUBLIC` or `PRIVATE` keywords, which resulted in transitive library dependencies by default. This unintentionally exposed non-public dependencies to downstream targets. Changes: - Always use explicit `PRIVATE` or `PUBLIC` keywords with `target_link_libraries()` - Tighten build dependency tree - Enforce a more modular linkage model See: [CMake documentation on library dependencies](https://cmake.org/cmake/help/latest/command/target_link_libraries.html) Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> Closes scylladb/scylladb#21686	2024-11-28 18:15:23 +02:00
Piotr Smaron	a49ed7074d	Update in-memory ks.metadata.init_tablets after ALTER KS Once e.g. `ALTER KEYSPACE` is performed, all in-memory objects should be updated accordingly, but this is not entirely true for keyspace metadata object. The reason for that is that keyspace metadata are stored in 2 system tables: `system_schema.keyspaces` and `system_schema.scylla_keyspaces`. Up until now the in-memory keyspace metadata object has been updated only with entries from the first table, and missed updates when entries from the 2nd table changed. These entries were e.g. initial tablets or storage options. This change fixes this oversight by considering both tables when checking if keyspace metadata need to be updated. From the implementation point of view, the change is simple: we're considering `system_schema.scylla_keyspaces` also in `merge_keyspaces()` and if old and new schemas have any differences, we include that when altering ks. Fixes #20768 Backport: no need, I don't think the issue is severe, atm it seems like it can only influence the tablets number, which should not bring the cluster down nor result in returning bad data, it can mostly influence the speed of the db. Closes scylladb/scylladb#20852	2024-11-28 13:46:32 +01:00
Nikos Dragazis	6091d5d789	sstables: Fix range of input stream in checksummed file data source The checksummed file data source uses the chunk size to enforce that the reads from the underlying file input stream will be aligned at the chunk boundary. This is necessary so that we can validate the checksum of each chunk. However, a mismatch in the numeric types caused a bug where the underlying file input stream would read a smaller portion of the data file than expected. The bug is located in the following lines: ``` auto start = _beg_pos & ~(chunk_size - 1); auto end = (_end_pos & ~(chunk_size - 1)) + chunk_size; ``` `_beg_pos` and `_end_pos` are `uint64_t`, whereas `chunk_size` is `uint32_t`. When executing the AND operation, the compiler converts the right operand from `uint32_t` to `uint64_t`. Since the integer is unsigned, the four most-significant bytes are filled with zeros, thus erroneously truncating the corresponding bytes of the position. Fix the bug by explicitly converting the chunk size to `uint64_t` before any arithmetic operations. Also, replace the handwritten alignment implementations with the `align_up()` and `align_down()` helpers. Finally, restrict the file end position to not exceed the file length. Since the last chunk can be smaller than the chunk size, it could happen that the end position exceeds the file length after the round-up. This is not a bug on its own since `make_file_input_stream()` can accept lengths that go beyond end-of-file, but still it makes the code more error prone and should be avoided. Signed-off-by: Nikos Dragazis <nikolaos.dragazis@scylladb.com> Closes scylladb/scylladb#21665	2024-11-28 12:53:05 +02:00
Dawid Mędrek	7cce9a8f64	db/hints: Prevent dereferencing a null pointer Before these changes, we dereferenced `app_state` in `manager::endpoint_downtime_not_bigger_than()` before checking that it's not a null pointer. We fix that. Fixes scylladb/scylladb#21699 Closes scylladb/scylladb#21676	2024-11-28 11:31:57 +01:00
Ernest Zaslavsky	4035e0877d	s3_tests: Add s3 test to check object re-uploading Add s3 test to check existing object re-uploading succeeds Closes scylladb/scylladb#21544	2024-11-28 12:46:59 +03:00
Pavel Emelyanov	58a2c6a7c3	Merge 'replica,sstables: track download progress of download_task_impl' from Kefu Chai Previously, the progress of download_task_impl launched by the "restore" API was not tracked. Since restore operations can involve large data transfers, this makes it difficult for users to monitor progress. The restore process happens in two sequential steps: 1. Open specified SSTables from object storage 2. Download and stream mutation fragments from the opened SSTables to mapped destinations While both steps contribute to overall progress, they use different units of measurement, making a unified progress metric challenging. Because the load-and-stream step (step 2) is the largest time-consuming part of the restore. This change implements progress tracking for this step as an initial improvement to provide users with partial visibility into the restore operation. Fixes scylladb/scylladb#21427 Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> --- this a part of experimental feature, hence no need to backport. Closes scylladb/scylladb#21562 * github.com:scylladb/scylladb: test/object_store: Enable tablets to match production settings sstables_loader: Track download progress of download_task_impl sstables_loader: improve batch tracking using ranges library sstables_loader: print streaming progress with moving range sstables_loader: mark sstable_streamer::stream_sstable_mutations() private sstables_loader: fix indentation in stream_sstable_mutations()	2024-11-28 12:46:31 +03:00
Laszlo Ersek	5f8549a9a0	configure.py: honor "--build-dir" when using CMake The "--use-cmake" option currently hardwires the build directory as "$source_dir/build". Adhere to the "--build-dir" option's argument instead: - If the option is not specified, its argument defaults to "build"; thus, there is no change in behavior. - If the option specifies a relative pathname, append it to $source_dir. - If the option specifies an absolute pathname, use it as-is. This is especially useful for keeping the build directory on a filesystem separate from the source directory (without resorting to creating "build" as a symlink, before running "configure.py"). For example, the source tree can be accessed remotely over sshfs, from a build host, while keeping the build artifacts (and hence the link stage) local to the build host. Signed-off-by: Laszlo Ersek <laszlo.ersek@scylladb.com> Closes scylladb/scylladb#21694	2024-11-28 11:26:26 +03:00
Avi Kivity	7e02f9bbaa	tombstone_gc.hh: remove include of boost/icl/interval_map.hh tombstone_gc.hh is relatively lightweight and is used in many places, but it includes the heavyweight boost/icl/interval_map.hh. Lighten the load for its users by wrapping lw_shared_ptr<some icl map type> in a forward-declared class. Define the class in a new header tombstone_gc-internals.hh, to be used by the two translation units that need it. Ref #1. Closes scylladb/scylladb#21706	2024-11-28 11:24:51 +03:00
Kefu Chai	23a7e9a6d0	docs: align tablestats documentation with actual output Update the tablestats documentation to correctly describe the "Number of partitions" metric. The previous documentation incorrectly referred to "estimated row count" when the command actually shows estimated partition count. Before: ``` Number of keys (estimate) \| The estimated row count ``` After: ``` Number of partitions (estimate) \| The estimated partition count ``` This distinction is important since a partition (identified by its partition key) can contain multiple rows in ScyllaDB. The updated format also matches Cassandra's nodetool output for better compatibility. Fixes scylladb/scylladb#21586 Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> Closes scylladb/scylladb#21598	2024-11-28 09:36:21 +02:00
Kefu Chai	79cc90141b	test/object_store: Enable tablets to match production settings Enable the `enable_tablets` configuration flag in object store tests to better align with production environments, where it is enabled by default via the `scylla.yaml` in Scylla's relocatable tarball. This change will improve test coverage of tablet-related features. Previously, `enable_tablets` defaulted to false in tests, creating a mismatch with typical production deployments. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2024-11-28 10:00:45 +08:00
Kefu Chai	5ab4932f34	sstables_loader: Track download progress of download_task_impl Previously, the progress of download_task_impl launched by the "restore" API was not tracked. Since restore operations can involve large data transfers, this makes it difficult for users to monitor progress. The restore process happens in two sequential steps: 1. Open specified SSTables from object storage 2. Download and stream mutation fragments from the opened SSTables to mapped destinations While both steps contribute to overall progress, they use different units of measurement, making a unified progress metric challenging. Because the load-and-stream step (step 2) is the largest time-consuming part of the restore. This change implements progress tracking for this step as an initial improvement to provide users with partial visibility into the restore operation. Fixes scylladb/scylladb#21427 Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2024-11-28 10:00:45 +08:00
Kefu Chai	e57f674066	sstables_loader: improve batch tracking using ranges library Replace manual vector iteration with ranges library to preserve batch size information. When streaming SSTable mutations, we need to track progress across batches. The previous implementation used a loop to move elements from the vector's end, but this approach lost the batch size information since the SSTable set was moved away during streaming. Now use std::ranges to take elements from the vector's end instead of manual iteration. This preserves the original batch size, enabling accurate progress tracking which will be implemented in a follow-up commit. Technical changes: - Replace manual vector iteration with ranges::take_view - Preserve batch size information for progress tracking - Maintain existing batch processing behavior Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2024-11-28 10:00:45 +08:00
Kefu Chai	38af123e05	sstables_loader: print streaming progress with moving range After `d1db17d490`, the processed SSTable counter remained at 0 while streaming progress was still being displayed. This fix properly tracks and displays streaming progress by: - Moving SSTable counter (`nr_sst_current`) to `sstable_streamer::stream_sstables()` - Generating UUID at the streaming initialization - Relocating progress reporting to `stream_sstables()` for accurate tracking This ensures the progress indicator correctly reflects the actual number of processed SSTables during streaming operations. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2024-11-28 10:00:45 +08:00
Kefu Chai	83f55fdb84	sstables_loader: mark sstable_streamer::stream_sstable_mutations() private the only user of `sstable_streamer::stream_sstable_mutations()` is `sstable_st6reamer::stream_sstables()`, so mark this member function as private. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2024-11-28 10:00:45 +08:00
Kefu Chai	5a39be9b8c	sstables_loader: fix indentation in stream_sstable_mutations() Fix indentation regression from `d1db17d490` where the function body of `sstable_streamer::stream_sstable_mutations()` was left incorrectly indented after the function was extracted to decouple streaming from sstable selection. Pure style fix, no functional changes. in this change, we correct the indent. Refs `d1db17d490` Signed-off-by: Kefu Chai <kefu.chai@scylladb.com>	2024-11-28 10:00:45 +08:00
Kefu Chai	5e391eee25	treewide: use coroutine::parallel_for_each(range) when appropriate `coroutine::parallel_for_each` accepts both a range and a pair of iterators. let's use the former when appropriate. it is simpler this way. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> Closes scylladb/scylladb#21684	2024-11-27 21:00:47 +02:00
Botond Dénes	20bbb1113e	test/cqlpy: test_tools.py: use xfail more selectively ScyllaDB doesn't support counters with tablets yet. So scylla-sstable tests which use counter schema are marked with xfail, but this is done too aggressively, disabling too many tests that are otherwise fine. There are two tests affected: * test_scylla_sstable_script - this test uses early return when the schema parameter is the one with counters and tablets are enabled. This is still too eager because tablets are now always enabled. Also, the early return make the fact that this test is disabled hidden. So change the check to check whether tablets are used on the test keyspace and use xfail instead of sneaky early return. * test_scylla_sstable_dump_data - this test is blanket-disabled when run with the tablets parameter. Even though only 1 out of 5 schemas tested use counters. Remove the blanket xfail and only add it when test keyspace uses tablets and the schema parameter is the one with counters. This makes dozens of test run again, restoring the test coverage lost with the too eager use of xfail (and sneaky return). Refs: #18180 Closes scylladb/scylladb#21685	2024-11-27 12:17:56 +03:00
Kefu Chai	8ca1c57de0	test: s3_proxy: bring back InjectingHandler.log_message in `0dff187b7a`, we dropped `InjectingHandler.log_message()`, but this method was defined to override the default implementation provided by `BaseHTTPRequestHandler.log_message()`. this change flooded the standard output when testing `aws_error_injection_test` with `test.py` with logging messages like: ``` 127.0.0.1 - - [26/Nov/2024 17:27:34] "PUT /?Policy=0&Key=%2Ftest%2Ftestobject-large-817295 HTTP/1.1" 200 127.0.0.1 - - [26/Nov/2024 17:27:34] "PUT /?Policy=1&Key=%2Ftest%2Ftestobject-large-817306 HTTP/1.1" 200 ``` this is unexpected. in this change, we bring this method back, and additionally, we format the logging message lazily. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> Closes scylladb/scylladb#21689	2024-11-27 12:16:36 +03:00
Botond Dénes	87bdfb80aa	docs/dev/reader-concurrency-semaphore.md: fix formatting of diagnostics dump Indent the whole thing so it is formatted as code, not as text. Closes scylladb/scylladb#21693	2024-11-27 12:13:16 +03:00
Botond Dénes	ccb433d767	Merge 'tasks: add api_task_ttl for tasks started with API' from Aleksandra Martyniuk When users start an operation asynchronously with API, they are expected to check the operation's status. Hence, the status should be kept in task manager for reasonable time after the operation is done. The operations that are started internally usually don't need to stay in task manager for that long. Add api_task_ttl that will be used for tasks started with API. By default it's 1 hour. The time for which non-API tasks stay in task manager isn't changed. Fixes: #21499. Refs: #21425. No backport needed - previous versions may use task_ttl Closes scylladb/scylladb#21505 * github.com:scylladb/scylladb: test: add test to check user_task_ttl tasks: api: move make_task method docs: nodetool: update backup and restore commands docs docs: update task manager docs nodetool: add nodetool tasks user-ttl command node_ops: use user task ttl for node ops virtual task tasks: use user_task_ttl for tasks started by user api: task_manager: add /task_manager/user_ttl to get and set user task ttl tasks: add task_manager::task::is_user_task method tasks: keep updateable_value of task_ttl in task manager db: config: add user_task_ttl_seconds named value	2024-11-27 09:57:57 +02:00
Nikita Kurashkin	4ba8a6b1b4	Fix test for DESC TABLE on materialised view to be compatible with Scylla AND Cassandra Fixes #21026 Refs #21500 Closes scylladb/scylladb#21526	2024-11-27 09:49:23 +02:00
Pavel Emelyanov	4d10cd40f0	s3: Remove unused boost/algorithm/string/classification.hpp inclusion Signed-off-by: Pavel Emelyanov <xemul@scylladb.com> Closes scylladb/scylladb#21690	2024-11-26 19:51:21 +02:00
Botond Dénes	98fa00499c	Update seastar submodule * seastar a5432364...72c7ac57 (13): > json2code: convert boost integer_range to std iota_view > utils/program-options: selection_value: make get_candidate_names() public > treewide: Trim trailing spaces > build: Clean up `compile_option` after FindSanitizers module > file: Convert file operations to use coroutines > websocket: Remove unnecessary condition in frame parsing > websocket: Fix logic when parsing header > websocket: Avoid memory copy when full websocket frames are received > websocket: Fix websocket frame parsing on partial packets > sharded.hh: remove inline from templates > reactor: fix reserve_io_control_blocks config name in error message > reactor: next_waitpid_timeout: contants can be defined as constexpr > reactor: coroutinize waitpid Closes scylladb/scylladb#21688	2024-11-26 19:50:45 +02:00
Kamil Braun	1f5b83dc56	Merge 'docs: update admin-tools docs with deprecation and removal notice for java tools' from Botond Dénes Java tools are deprecated and slated for removal in the next ScyllaDB release. Update the admin-tools docs and make sure all java tool documentation pages have a notice reflecting this fact. Fixes: https://github.com/scylladb/scylladb/issues/21149 Should be backported to 6.2, so users of the latest stable version can see the notice. Closes scylladb/scylladb#21522 * github.com:scylladb/scylladb: docs: sstableloader.rst: add deprecation notice docs: admin-tools: update deprecation notice for sstable{dump,metadata} docs: tools_index.rst: remove deprecated sstablereset and sstablerepairedset tools	2024-11-26 17:03:56 +01:00
Ernest Zaslavsky	793f2c95d1	snapshots: Stop taking snapshots of MVs Stop taking snapshots of MVs and allow taking snapshot of individual tables, now one can take a snapshot of any base table, any view or index. Also add tests to cover new cases both boost test (using cc code) and pytest (using the API) Also, update documentation to reflect the change fixes: #21339 fixes: #20760 Closes scylladb/scylladb#21433	2024-11-26 15:27:30 +02:00
Kamil Braun	9dc8926252	Merge 'a bunch of cleanups and enhancements to various services' from Gleb Mostly no functional changes here except in patch 3. * 'gleb/cleanups' of github.com:scylladb/scylla-dev: migration_manager: move migration manager verbs to the IDL storage_proxy: remove unused function storage_proxy: co-routinize handle_paxos_prepare storage_proxy: co-routinise handle_paxos_prune service: raft: no need to sync schema if the cluster is in raft topology mode messaging_service: co-routinize messaging_service::stop_client gossiper: rename apply_state_locally_without_listener_notification to apply_state_locally_in_shadow_round	2024-11-26 14:18:00 +01:00
Kefu Chai	a5ee0c896b	treewide: migrate from boost::adaptors::filtered to std::views::filter Modernize the codebase by replacing Boost range adaptors with C++23 standard library views, reducing external dependencies and leveraging modern C++ language features. Key Changes: - Replace `boost::adaptors::filtered` with `std::views::filter` - Remove `#include <boost/range/adaptor/filtered.hpp>` - Utilize standard library range views Motivation: - Reduce project's external dependency footprint - Leverage standard library's range and view capabilities - Improve long-term code maintainability - Align with modern C++ best practices Implementation Challenges and Considerations: 1. Range Conversion and Move Semantics - `std::ranges::to` adaptor requires rvalue references - Necessitated updates to variable and parameter constness - Example: `cql3/restrictions/statement_restrictions.cc` modified to remove `const` from `common` to enable efficient range conversion 2. Range Iteration and Mutation - Range views may mutate internal state during iteration - Cannot pass ranges by const reference in some scenarios - Solution: Pass ranges by rvalue reference to explicitly indicate state invalidation Limitations: - One instance of `boost::adaptors::filtered` temporarily preserved due to lack of a C++23 alternative for `boost::join()` - A comprehensive replacement will be addressed in a follow-up change This change is part of our ongoing effort to modernize the codebase, reducing external dependencies and adopting modern C++ practices. Signed-off-by: Kefu Chai <kefu.chai@scylladb.com> Closes scylladb/scylladb#21648	2024-11-26 14:26:50 +02:00
Aleksandra Martyniuk	ac6a07117a	test: add test to check user_task_ttl	2024-11-26 09:57:42 +01:00
Aleksandra Martyniuk	1712c93261	tasks: api: move make_task method task_manager::module::make_task method template is used only for test_task_impl. Move it to api/task_manager_test.cc and modify it to be test_task_impl-specific.	2024-11-26 09:57:42 +01:00
Aleksandra Martyniuk	1244982071	docs: nodetool: update backup and restore commands docs	2024-11-26 09:57:41 +01:00
Aleksandra Martyniuk	3b86150e88	docs: update task manager docs	2024-11-26 09:57:41 +01:00
Aleksandra Martyniuk	1ade668d79	nodetool: add nodetool tasks user-ttl command	2024-11-26 09:57:23 +01:00
Evgeniy Naydanov	1e9d780e89	test.py: deselect random failures which can cause #21534 Following combinations of error injections and cluster events can cause #21534. Disable them for now because they break CI. Closes scylladb/scylladb#21658	2024-11-26 10:38:15 +02:00
Aleksandra Martyniuk	e703ba08f8	node_ops: use user task ttl for node ops virtual task Use user task ttl for node ops virtual task. Modify the test accordingly.	2024-11-25 14:21:53 +01:00
Aleksandra Martyniuk	6241d49b64	tasks: use user_task_ttl for tasks started by user	2024-11-25 14:21:53 +01:00
Aleksandra Martyniuk	19a90e3697	api: task_manager: add /task_manager/user_ttl to get and set user task ttl	2024-11-25 14:21:53 +01:00
Aleksandra Martyniuk	292d00463a	tasks: add task_manager::task::is_user_task method	2024-11-25 14:21:53 +01:00
Aleksandra Martyniuk	16e204dfdb	tasks: keep updateable_value of task_ttl in task manager Drop task_ttl observer from task manager and use updateable_value.	2024-11-25 14:20:43 +01:00
Aleksandra Martyniuk	1bf073704c	db: config: add user_task_ttl_seconds named value Add user_task_ttl_seconds config option and keep the value in task manager. In the following patches tasks started by user will be kept in task manager for user_task_ttl_seconds after they are finished.	2024-11-25 14:16:06 +01:00
Nadav Har'El	cb6c55209a	Merge 'locator: token_metadata: replace boost range with std range' from Avi Kivity Reduce dependency load by standardizing on std::ranges. This is a little involved since a we use a custom iterator. Code cleanup; no backport. Closes scylladb/scylladb#21421 * github.com:scylladb/scylladb: locator: token_metadata: switch from boost ranges to std ranges locator: token_metadata: make iterator support std::input_iterator concept locator: tokens_metadata: move tokens_iterator to namespace scope	2024-11-25 14:58:45 +02:00
Botond Dénes	510e09c648	docs/ddl: document memtable_flush_period_in_ms This option was implemented by scylladb/scylladb#20999 but it wasn't documented. Add a description of this option to the create table page. Note that the option was accepted already before scylladb/scylladb#20999, but it's value was ignored. Fixes: scylladb/scylladb#21671 Closes scylladb/scylladb#21673	2024-11-25 13:53:21 +02:00
Nadav Har'El	61e8975930	Merge 'test/boost/view_schema_test: Improve test_view_update_generating_writetime' from Dawid Mędrek In this PR, we improve various aspects of the test: * increase obtained information whenever any test case fails, * split test cases, * elaborate on the semantics of generating view updates and what exactly we check and why. Backport: not needed, this is an enhancement. Closes scylladb/scylladb#21579 * github.com:scylladb/scylladb: test/boost/view_schema_test: Improve comments in test_view_update_generating_writetime test/boost/view_schema_test.cc: Improve checks in test_view_update_generating_writetime test/boost/view_schema_test.cc: Split test cases in test_view_update_generating_writetime	2024-11-25 13:46:56 +02:00
Botond Dénes	090ab796dd	Merge 'repair: Enable small table optimization for RBNO bootstrap and decommission' from Asias He The non local strategy system keyspaces usually contain very litte data. All the tables within them have to be repaired for all the token ranges, which could be large in clusters with a large number of nodes. In multiple DC setup, the repair in RBNO is dominated by the network latency. As a result, it takes a long time to repair those tables even if they are almost empty. To speed up the RBNO bootstrap, especially for starting empty clusters, this patch enables small table optimization for RBNO for system tables. We could enable it for small user tables as a follow up. Tests: 1) A 5ms latency is added to simulate cross dc network delay, 256 tokens per node, 10 nodes: - Before topology_custom dev topology_custom.test_boot_time.1 1287.06s - After topology_custom dev topology_custom.test_boot_time.1 12.48s The test shows 100X boot time improvement 2) A SCT test to bootstrap 3 DCs, 3 nodes in each DC. - Before Time to bootstrap = 1h23m - After Time to bootstrap = 13m The test shows 6X bootstrap time improvement Fixes #19131 New feature. No backport is needed. Closes scylladb/scylladb#21207 * github.com:scylladb/scylladb: repair: Enable small table optimization for RBNO bootstrap and decommission repair: Move flush_rows after repair_meta class	2024-11-25 11:50:55 +02:00
Nadav Har'El	71c671eeaa	docs: copy-edit docs/alternator/compatibility.md I reread the "ScyllaDB Alternator for DynamoDB users" document (alternator/compatibility.md) and improved various places that I thought needed improvement. Two of the more significant changes is moving the not-really-important "Scan ordering" section much lower in the document and explaining it better, and improving the "provisioning" section to focus on the available and missing functionality, and not on minor API details. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Closes scylladb/scylladb#21605	2024-11-25 10:02:36 +03:00

1 2 3 4 5 ...

45576 Commits