scylladb

mirror of https://github.com/scylladb/scylladb.git synced 2026-05-01 05:35:48 +00:00

Author	SHA1	Message	Date
Piotr Jastrzebski	9348006092	Implement reading rows and columns in data_consume_rows_context_m Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 19:54:16 +02:00
Piotr Jastrzebski	f6e1c38486	Introduce column_flags_m This will be used for reading columns from data file. Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 19:54:16 +02:00
Piotr Jastrzebski	609854e21a	Add column_translation to data_consume_rows_context_m Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 19:54:16 +02:00
Piotr Jastrzebski	7fd222e639	Pass schema to data_consume_context It will be needed to obtain column_translation that will be added to data_consume_context in the next patch. Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 19:54:16 +02:00
Piotr Jastrzebski	d3f3cd36dd	Add column_translation.hh It contains a class that manages mapping between sstable columns and schema column definitions. Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 19:54:16 +02:00
Piotr Jastrzebski	25b8cf9e4c	consumer_m: Add consume methods for consuming rows and columns Also implement them in mp_row_consumer_m. Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 19:53:29 +02:00
Piotr Jastrzebski	94e3138dc5	Extract make_atomic_cell from mp_row_consumer_k_l It will be used in both mp_row_consumer_k_l and mp_row_consumer_m. Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:39:52 +02:00
Piotr Jastrzebski	c6d5ebc274	Rename NON_STATIC_ROW_* states to ROW_BODY_* New name describes the states in a better way as those states will be used both for static and non-static rows. Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:39:52 +02:00
Piotr Jastrzebski	10c669d2b5	Add liveness_info and use it in reading sstables Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:39:52 +02:00
Piotr Jastrzebski	b2f9841dd4	Add helper methods for parsing simple types. Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:39:52 +02:00
Piotr Jastrzebski	d8cd8e04ed	Add unfiltered_flags_m::has_all_columns Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:39:52 +02:00
Piotr Jastrzebski	51d079e17c	data_consume_context: use make_unique instead of new Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:39:52 +02:00
Piotr Jastrzebski	54ef775501	Pass serialization_header to data_consume_rows_context* This header is needed to parse data for SSTable 3.0 format Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:39:52 +02:00
Piotr Jastrzebski	b849eefc8c	Use disk_string_vint_size for bytes_array_vint_size Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:39:52 +02:00
Piotr Jastrzebski	76f0f2693d	Introduce disk_string_vint_size type Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:30:03 +02:00
Piotr Jastrzebski	5ca4bfd69a	disk_array_vint_size: Remove unused Size template parameter Signed-off-by: Piotr Jastrzebski <piotr@scylladb.com>	2018-05-23 16:15:44 +02:00
Avi Kivity	20271b3890	Update scylla-ami submodule * dist/ami/files/scylla-ami e0b35dc...025644d (1): > Merge "AMI build fix" from Takuya	2018-05-16 12:33:45 +03:00
Avi Kivity	05cec4a265	Merge "Reduce LSA memory reclamation overhead" from Tomasz " Main optimization is in the patch titled "lsa: Reduce amount of segment compactions". I measured 50% reduction of cache update run time in a steady state for an append-only workload with large partition, in perf_row_cache_update version from: `c3f9e6ce1f/tests/perf_row_cache_update.cc` Other workloads, and other allocation sites probably also could see the improvement. " * tag 'tgrabiec/reduce-lsa-segment-compactions-v1' of github.com:tgrabiec/scylla: lsa: Expose counters for allocation and compaction throughput lsa: Reduce amount of segment compactions lsa: Avoid the call to segment_pool::descriptor() in compact() lsa: Make reclamation on reserve refill more efficient	2018-05-16 10:24:20 +03:00
Tomasz Grabiec	534068a0f7	Update seastar submodule Fixes #3339 * seastar 840002c...0a1a327 (7): > Merge "fix perftune.py issues with cpu-masks on big machines" from Vlad > Merge 'Handle Intel's NICs in a special way' from Vlad > reactor: fix calculation of idle ticks > log: streamline logging internals a little > Merge "CMake imrovements and compatibility" from Jesse > iotune: fix typo in property name > cmake: do not find_package(Boost ...) if Boost is a target	2018-05-16 09:11:22 +02:00
Avi Kivity	832e8fb1e0	Merge "Support writing counters in SSTables 3.x format." from Vladimir " This patchset adds support for writing counter cells in SSTables 3.x format ('m'). The logic of writing counters is almost identical to that used for the old 2.x format ('k'/'l') with the only difference that the data length preceding serialised shards is written as a vint. Tests: unit {release}. Generated SSTables are verified to be processed fine by sstabledump (note that sstabledump only outputs the binary data for counters, not their actual values, same as sstable2json). Verified with Cassandra 3.11 to get the expected values from the counters table: cqlsh> SELECT * from sst3.counter_table; pk \| ck \| rc1 \| rc2 -----+-----+-----+----- key \| ck1 \| 10 \| 1 (1 rows) Verified that the deleted counter can no longer be updated: cqlsh> use sst3 ; cqlsh:sst3> UPDATE counter_table SET rc1 = rc1 + 2 WHERE pk = 'key' AND ck = 'ck2'; cqlsh:sst3> SELECT * from sst3.counter_table; pk \| ck \| rc1 \| rc2 -----+-----+-----+----- key \| ck1 \| 10 \| 1 (1 rows) " * 'projects/sstables-30/write_counters/v1' of https://github.com/argenet/scylla: tests: Unit tests to cover writing counters in SSTables 3.x format. sstables: Support writing counters for SSTables 3.x. sstables: Move code writing counter value into a separate helper.	2018-05-16 08:46:15 +03:00
Raphael S. Carvalho	59c57861ae	tests/sstable_test: switch to dynamic temporary dir creation sstable test fails when running concurrently (for example, release and debug mode) because it uses a static temporary dir in lots of tests. Let's fix it by switching to dynamic temporary dir, which is created using mkdtemp(). Also the sstable tests will now run in /tmp, and so it's made much faster. Signed-off-by: Raphael S. Carvalho <raphaelsc@scylladb.com> Message-Id: <20180516042044.15336-1-raphaelsc@scylladb.com>	2018-05-16 08:00:29 +03:00
Tomasz Grabiec	4fdd61f1b0	lsa: Expose counters for allocation and compaction throughput Allow observing amplification induced by segment compaction.	2018-05-15 21:49:01 +02:00
Tomasz Grabiec	3775a9ecec	lsa: Reduce amount of segment compactions Reclaiming memory through segment compaction is expensive. For occupancy of 85%, in order to reclaim one free segment, we need to compact 7 segments, by migrating 6 segments worth of data. This results in significant amplification. Compaction involves moving objects, which in some cases is expensive in itself as well (See https://github.com/scylladb/scylla/issues/3247). This patch reduces amount of segment compactions in favor of doing more eviction. It especially helps workloads in which LRU order matches allocation order, in which case there will be no segment compaction, and just eviction. In perf_row_cache_update test case for large partition with lots of rows, which simulates appending workload, I measured that for each new object allocated, 2 need to be migrated, before the patch. After the patch, only 0.003 objects are migrated. This reduces run time of cache update part by 50%.	2018-05-15 21:49:01 +02:00
Vladimir Krivopalov	a16b8d5d77	tests: Unit tests to cover writing counters in SSTables 3.x format. Signed-off-by: Vladimir Krivopalov <vladimir@scylladb.com>	2018-05-15 11:44:44 -07:00
Vladimir Krivopalov	ffd8886da9	sstables: Support writing counters for SSTables 3.x. Signed-off-by: Vladimir Krivopalov <vladimir@scylladb.com>	2018-05-15 11:44:44 -07:00
Vladimir Krivopalov	28c3c21c73	sstables: Move code writing counter value into a separate helper. Signed-off-by: Vladimir Krivopalov <vladimir@scylladb.com>	2018-05-15 11:44:44 -07:00
Avi Kivity	5f3a5c436e	Merge "chunked vector memory estimation" from Glauber " The memory estimations we have when using the chunked vector are usually slightly wrong. We can make them more accurate by exporting the memory usage directly as a chunked_vector API. " * 'chunked_memory-v2' of github.com:glommer/scylla: large_bitset: be more accurate with memory usage chunked_vector: exports its current memory usage	2018-05-15 19:00:36 +03:00
Glauber Costa	2ba08178ca	large_bitset: be more accurate with memory usage We are slightly underestimating the amount of memory we use. Now that the chunked vector can exports its internal memory usage we can use that directly. Signed-off-by: Glauber Costa <glauber@scylladb.com>	2018-05-15 11:22:21 -04:00
Glauber Costa	7190bb4f95	chunked_vector: exports its current memory usage There are times in which we would like to estimate how much memory a chunked_vector is using. We have two strategies to do it: 1) multiply the size by the size of the elements. That is wrong, because the chunked_vector can allocate larger chunks in anticipation of more elements to come. 2) multiply the number of chunks by 128kB. That is also wrong, because the chunk_vector will not always allocate the entire chunk if there are only a few elements in it. The best way to deal with it is to allow the chunked_vector to exports its current memory usage. Signed-off-by: Glauber Costa <glauber@scylladb.com>	2018-05-15 11:22:21 -04:00
Raphael S. Carvalho	83e64192d3	tests/perf: fix compaction and write mode of perf_sstable storage_service_for_tests must be instantiated only once at a global scope. Fixes #3369. Signed-off-by: Raphael S. Carvalho <raphaelsc@scylladb.com> Message-Id: <20180510042200.2548-1-raphaelsc@scylladb.com>	2018-05-15 18:00:18 +03:00
Avi Kivity	e0ef39705f	dist: redhat: properly package scylla_blocktune.py Commit `9eb8ea8b11` installed scylla_blocktune.py as part of preparing the rpm, but forgot to add it to the installed file list, breaking the rpm build. Fix by listing the file in the %files section. Message-Id: <20180506202807.5719-1-avi@scylladb.com>	2018-05-15 18:00:05 +03:00
Piotr Sarna	40bf5d671b	cql: add secondary index metrics This commit adds basic secondary index metrics to cql_stats: * total number of indexes creates * total number of indexes dropped * total number of reads from a secondary index * total number of rows read from a secondary index References #3384 Reviewed-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <d5eda7a343cee547c921dd4d289ecb1ac1c2bf24.1526374243.git.sarna@scylladb.com>	2018-05-15 17:59:53 +03:00
Avi Kivity	4f81e1f55a	Merge "Use CRC32 to calculate checksums for SSTables 3.0." from Vladimir " SSTables 3.x (format 'm') use CRC32 instead of Adler32 for calculating checksums. This patchset introduces support for CRC32 along with Adler32 in checksummed_file_writer to be used for SSTables written in 'mc' format. Structures and helpers introduced for CRC32 will be later used for calculating checksums for compressed files as well (not a part of this patchset). Tests: unit {release} " * 'projects/sstables-30/write-digest-crc/v3' of https://github.com/argenet/scylla: tests: Add test covering checksumming SSTables 3.0 with CRC32. sstables: Support CRC32 checksum for SSTables 3.x. sstables: Move adler32 routines under the scope of a class. sstables: Move checksum utils into separate header. sstables: Remove unused 'checksum_file' flag from checksummed_file_writer.	2018-05-15 10:18:14 +03:00
Duarte Nunes	3a7d655d01	Merge 'transport: reduce unneeded continuations' from Avi " The native protocol server generates mant reactor tasks that can be easily eliminated. I measured a read workload with 100% cache hit rate, seeing the number of tasks per request drop from ~31 to ~27, and an increase of 3% in throughput. " * tag 'transport-optimize-1/v1' of https://github.com/avikivity/scylla: transport: remove unused capture of flags variable transport: merge response write and error handling continuations transport: make write_repsonse() return void transport: de-template a lambda transport: merge memory-management and logging continuations transport: remove gate continuation transport: merge two response processing continuations transport: simplify response processing continuation transport: remove gratuitous continuation from process_request_one()	2018-05-14 10:12:07 +01:00
Avi Kivity	4500baaaf4	transport: remove unused capture of flags variable	2018-05-14 09:41:06 +03:00
Avi Kivity	88f8fe3168	transport: merge response write and error handling continuations The response write continuation does not defer, so traditional try/catch works well and saves a continuation.	2018-05-14 09:41:06 +03:00
Avi Kivity	3e8d1c8fd7	transport: make write_repsonse() return void It just schedules the response, and returns immediately. (I thought about calling it schedule_response(), but usually it will write the response immediately, since waiting for network writes is rare in a local network).	2018-05-14 09:41:06 +03:00
Avi Kivity	b26f36c2ec	transport: de-template a lambda Generic templates = annoying.	2018-05-14 09:41:06 +03:00
Avi Kivity	7a9b73f166	transport: merge memory-management and logging continuations Merge a continuation that just keeps things alive with another that just logs things.	2018-05-14 09:41:06 +03:00
Avi Kivity	f0887a55e4	transport: remove gate continuation with_gate() generates a continuation if the protected function defers. Avoid that by merging a gate::leave() call with another, preexisting, continuation.	2018-05-14 09:41:06 +03:00
Avi Kivity	876837a5da	transport: merge two response processing continuations We have one coninuation transforming the result, and another shutting down tracing. Since the first cannot defer, we can merge the two, reducing the number of tasks processed by the reactor.	2018-05-14 09:41:06 +03:00
Avi Kivity	38619138be	transport: simplify response processing continuation A continuation in the response processing path is only doing transformation on the output. Make that clear by returning a value, not a future.	2018-05-14 09:41:06 +03:00
Avi Kivity	f0a1478b6c	transport: remove gratuitous continuation from process_request_one() No need to call then() just to convert exceptions to futures, futurize_apply() does this with less ado.	2018-05-14 09:41:06 +03:00
Vladimir Krivopalov	1da6144f90	tests: Add test covering checksumming SSTables 3.0 with CRC32. Signed-off-by: Vladimir Krivopalov <vladimir@scylladb.com>	2018-05-13 12:38:25 -07:00
Vladimir Krivopalov	e6dfa008d8	sstables: Support CRC32 checksum for SSTables 3.x. Signed-off-by: Vladimir Krivopalov <vladimir@scylladb.com>	2018-05-13 12:38:25 -07:00
Vladimir Krivopalov	adb43959d1	sstables: Move adler32 routines under the scope of a class. This is a step towards making digest algorithm customizable at compile time. Signed-off-by: Vladimir Krivopalov <vladimir@scylladb.com>	2018-05-13 12:38:25 -07:00
Vladimir Krivopalov	4e4030676f	sstables: Move checksum utils into separate header. Checksummed writer doesn't need to include all compression stuff. Signed-off-by: Vladimir Krivopalov <vladimir@scylladb.com>	2018-05-13 12:38:25 -07:00
Nadav Har'El	f5536d607e	secondary index: fix multiple appearance of rows This patch fixes a bug where queries using a secondary index would, in some cases, produce the same rows multiple times. The problem was that the code begins by finding a list of primary keys that match the search, and then work on the partitions containing them. If multiple rows matched in the same partition, the partition was considered multiple times, and the same rows were output multiple times. Signed-off-by: Nadav Har'El <nyh@scylladb.com> Message-Id: <20180510203141.17157-1-nyh@scylladb.com>	2018-05-13 20:08:14 +02:00
Avi Kivity	7d29addb1f	mutation_reader: optimize make_combined_reader for the single-reader case If we're given a single reader (can be common in a low-write-rate table, where most of the data will be in a single large sstable, or in leveled tables) then we can avoid the overhead of the combining reader by returning the single input. Tests: unit (release) Message-Id: <20180513130333.15424-1-avi@scylladb.com>	2018-05-13 20:07:10 +02:00
Duarte Nunes	a23bda3393	Merge 'Implement separate timeout for range queries' from Avi " This patchset implements separate timeouts for range queries, and lays the foundations for separate timeouts for other query types. While the feature in itself is worthy, the real motivation is to have the timeouts decided by the caller, instead of storage_proxy. This in turn is required to disentangle each layer behaving differently depending on whether the query is internal or not; instead, the goal is to have each caller declare its needs in terms of consistency level and timeouts, and have the lower layers implement its requirements instead of making their own decisions. Fixes #3013. Tests: unit (release) " * tag '3013/v1.1' of https://github.com/avikivity/scylla: storage_proxy: remove default_query_timeout() storage_proxy: don't use default timeouts query_options: augment with timeout_config thrift: configure thrift transport and handler with a timeout_config transport: configure native transport with a timeout_config cql3: define and populate timeout_config_selector timeout_config: introduce timeout configuration	2018-05-13 20:05:50 +02:00

1 2 3 4 5 ...

15397 Commits