scylladb

mirror of https://github.com/scylladb/scylladb.git synced 2026-04-22 09:30:45 +00:00

Author	SHA1	Message	Date
Juliusz Stasiewicz	00a6fda7b9	tracing: Trace slow queries on replicas wrt. parent's clock Secondary tracing sessions used to compute the execution time from the point of their `begin()`-ning, not the parent session's `begin()`. As a result, replica reported a slow query if it exceeded the entire threshold on that replica too. This change augments `trace_info` with the TS of parent's session starting point, to be used as a reference on replicas. Fixes #9403 Closes #10005	2022-02-10 12:03:53 +01:00
Avi Kivity	fcb8d040e8	treewide: use Software Package Data Exchange (SPDX) license identifiers Instead of lengthy blurbs, switch to single-line, machine-readable standardized (https://spdx.dev) license identifiers. The Linux kernel switched long ago, so there is strong precedent. Three cases are handled: AGPL-only, Apache-only, and dual licensed. For the latter case, I chose (AGPL-3.0-or-later and Apache-2.0), reasoning that our changes are extensive enough to apply our license. The changes we applied mechanically with a script, except to licenses/README.md. Closes #9937	2022-01-18 12:15:18 +01:00
Benny Halevy	d96a67eb57	abstract_replication_strategy: use shared_ptr in registry Enable creating shared_ptr<BaseClass> in nonstatic_class_registry using BaseClass::ptr_type and use that for abstract_replication_strategy. While at it, also clean up compressor with that respect to define compressor::ptr_type as shared_ptr<compressor> thus simplifying compressor_registry. Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2021-10-13 12:39:36 +03:00
Avi Kivity	a55b434a2b	treewide: extent copyright statements to present day	2021-06-06 19:18:49 +03:00
Pavel Emelyanov	887a1b0d3d	tracing: Stop tracing in main's deferred action Tracing is created in two steps and is destroyed in two too. The 2nd step doesn't have the corresponding stop part, so here it is -- defer tracing stop after it was started. But need to keep in mind, that tracing is also shut down on drain, so the stopping should handle this. Fixes #8382 tests: unit(dev), manual(start-stop, aborted-start) Signed-off-by: Pavel Emelyanov <xemul@scylladb.com> Message-Id: <20210331092221.1602-1-xemul@scylladb.com>	2021-03-31 12:28:37 +03:00
Ivan Prisyazhnyy	85fbca2049	tracing: omit tracing session events and subsessions in fast mode If tracing::tracing::_ignore_trace_events is enabled then the tracing system must ignore all sessions events for non full_tracing sessions (probability tracing and user requested) and creating subsessions with the make_trace_info. Patch introduces the slow query tracing fast mode that omits all events during tracing. Signed-off-by: Ivan Prisyazhnyy <ivan@scylladb.com>	2021-03-18 15:04:47 +02:00
Pavel Emelyanov	87f1223965	tracing: Push query processor through init methods The goal is to make tracing keyspace helper reference query processor, so this patch adds the needed arguments through the initialization stack. Signed-off-by: Pavel Emelyanov <xemul@scylladb.com>	2020-10-06 15:45:12 +03:00
Benny Halevy	7aef39e400	tracing: one_session_records: keep local tracing ptr Similar to trace_state keep shared_ptr<tracing> _local_tracing_ptr in one_session_records when constructed so it can be used during shutdown. Fixes #5243 Signed-off-by: Benny Halevy <bhalevy@scylladb.com>	2019-11-28 15:24:10 +01:00
Duarte Nunes	fa2b0384d2	Replace std::experimental types with C++17 std version. Replace stdx::optional and stdx::string_view with the C++ std counterparts. Some instances of boost::variant were also replaced with std::variant, namely those that called seastar::visit. Scylla now requires GCC 8 to compile. Signed-off-by: Duarte Nunes <duarte@scylladb.com> Message-Id: <20190108111141.5369-1-duarte@scylladb.com>	2019-01-08 13:16:36 +02:00
Avi Kivity	6641353854	tracing: remove static class_registry Static class_registries hinder librarification by requiring linking with all object files (instead of a library from which objects are linked on demand) and reduce readability by hiding dependencies and by their horrible syntax. Hide them behind a non-static, non-template tracing backend registry. Message-Id: <20181229121000.7885-1-avi@scylladb.com>	2018-12-31 13:24:54 +00:00
Vlad Zolotarov	896c1822b5	tracing: move all tracing related API functions to a cold path This patch completes what was started in `a4282c2c6e` Make trace_state_ptr to be a wrapper class around lw_shared_ptr<trace_state> that hints that bool(trace_state_ptr) is likely to return FALSE. Signed-off-by: Vlad Zolotarov <vladz@scylladb.com>	2018-08-03 12:32:54 -04:00
Vlad Zolotarov	6db90a2e63	tracing: store a query response size Add a new "response_size" column to system_traces.sessions and store a size of an uncompressed response for a traced query. Signed-off-by: Vlad Zolotarov <vladz@scylladb.com>	2018-08-03 12:29:36 -04:00
Vlad Zolotarov	05020921bb	tracing: store request size Add a new column "request_size" to system_traces.sessions and store the uncompressed request frame data size. Signed-off-by: Vlad Zolotarov <vladz@scylladb.com>	2018-08-03 12:29:36 -04:00
Vlad Zolotarov	fcff872089	tracing: make the session state modifying methods and tracing::trace(...) noexcept Make state session creation, stop_forground() and tracing::trace(...) methods noexcept. Most of them have already been implemented the way that they won't throw but this patch makes it official... Signed-off-by: Vlad Zolotarov <vladz@scylladb.com>	2017-12-14 15:05:48 -05:00
Avi Kivity	c6fa727af0	tracing: add missing include The IDE doesn't understand what lw_shared_ptr<> means without it, though it does compile.	2017-11-21 13:24:07 +02:00
Avi Kivity	a9f19e37b5	tracing: add missing include "log.hh" It's currently made available via another include, which is going away.	2017-08-27 15:18:41 +03:00
Vlad Zolotarov	b0f660331a	tracing: introduce a span ID and parent span ID This patch makes the tracing framework follow the general idea of Google's Dapper paper: traces generated in a context of the same query are forming a single-rooted acyclic tree where in a ScyllaDB case vertexes are spans running on each involved replica Node and edges are RPCs sent from one Node to another. - Each vertex in the tree above has an ID - "span ID". - In order to be able to build the tree from the sessions traces we need to know the parent "span ID" - the ID of a span that sent an RPC that created the current span. - Each span of a tracing session is given a 64-bit random span ID. - The root span has a span_id::illegal_id value. This patch adds: - The described above parent span ID and a span ID to the one_session_records object. - The current span ID is passed in the trace_info struct to the remote replica. - Add parent_id and span_id columns to system_traces.events table for the parent ID and span ID. Signed-off-by: Vlad Zolotarov <vladz@scylladb.com>	2017-04-25 21:52:23 -04:00
Vlad Zolotarov	aa1f8ccea4	tracing::tracing: allow slow query TTL only in the signed 32-bit integer range Any TTL is eventually converted into the gc_clock::duration value, which is based on int32_t type. Limit the node_slow_log TTL user configurable value to the same values range for consistency. Signed-off-by: Vlad Zolotarov <vladz@scylladb.com>	2017-04-12 12:24:08 -04:00
Amnon Heiman	e19fa02a17	remove scollectd from headers As the metrics migration progressed, some include to scollectd.hh left behind. Because of the nature of the scollecd implementation those include brings alot of code with them to the header files and eventually to many source file. This patch remove those include and add a missing include to storage_proxy.cc. The reason the compiler didn't complain is an indication to the problematic nature of those include in the first place. Before this patch, change in metrics.hh would cause 169 files to compile, after this change 17. Signed-off-by: Amnon Heiman <amnon@scylladb.com> Message-Id: <1484667536-2185-1-git-send-email-amnon@scylladb.com>	2017-01-17 17:39:47 +02:00
Vlad Zolotarov	6267bb63f4	tracing::tracing: move collectd metrics registration to metrics registration layer Signed-off-by: Vlad Zolotarov <vladz@scylladb.com>	2017-01-10 16:24:54 -05:00
Pekka Enberg	a443dfa95e	tracing: Add seastar/core/scollectd.hh include Fix the following build breakage: FAILED: build/release/gen/cql3/CqlParser.o g++ -MMD -MT build/release/gen/cql3/CqlParser.o -MF build/release/gen/cql3/CqlParser.o.d -std=gnu++1y -g -Wall -Werror -fvisibility=hidden -pthread -I/home/penberg/scylla/seastar -I/home/penberg/scylla/seastar/fmt -I/home/penberg/scylla/seastar/build/release/gen -march=nehalem -Ifmt -DBOOST_TEST_DYN_LINK -Wno-overloaded-virtual -DFMT_HEADER_ONLY -DHAVE_HWLOC -DHAVE_NUMA -DHAVE_LZ4_COMPRESS_DEFAULT -O2 -DBOOST_TEST_DYN_LINK -Wno-maybe-uninitialized -DHAVE_LIBSYSTEMD=1 -I. -I build/release/gen -I seastar -I seastar/build/release/gen -c -o build/release/gen/cql3/CqlParser.o build/release/gen/cql3/CqlParser.cpp In file included from ./query-request.hh:31:0, from ./locator/token_metadata.hh:51, from ./locator/abstract_replication_strategy.hh:29, from ./database.hh:26, from ./service/storage_proxy.hh:44, from ./db/schema_tables.hh:43, from ./db/system_keyspace.hh:46, from ./cql3/functions/function_name.hh:45, from ./cql3/selection/selectable.hh:48, from ./cql3/selection/writetime_or_ttl.hh:45, from build/release/gen/cql3/CqlParser.hpp:63, from build/release/gen/cql3/CqlParser.cpp:44: ./tracing/tracing.hh:357:5: error: ‘scollectd’ does not name a type scollectd::registrations _registrations; ^~~~~~~~~ Message-Id: <1482939751-8756-1-git-send-email-penberg@scylladb.com>	2016-12-28 18:40:18 +02:00
Vlad Zolotarov	62cad0f5f5	tracing: don't start tracing until a Tracing service is fully initialized RPC messaging service is initialized before the Tracing service, so we should prevent creation of tracing spans before the service is fully initialized. We will use an already existing "_down" state and extend it in a way that !_down equals "started", where "started" is TRUE when the local service is fully initialized. We will also split the Tracing service initialization into two parts: 1) Initialize the sharded object. 2) Start the tracing service: - Create the I/O backend service. - Enable tracing. Fixes issue #1939 Signed-off-by: Vlad Zolotarov <vladz@scylladb.com> Message-Id: <1481836429-28478-1-git-send-email-vladz@scylladb.com>	2016-12-21 12:40:14 +02:00
Vlad Zolotarov	a491ac0f18	tracing: introduce a log_slow_query logic The main idea is to log queries that take "too long" to complete. The "too long" is above the given threshold. To achieve the above this patch does the following: - Introduce two new properties to the tracing::trace_state: - "Full tracing": when the tracing of this query was explicitly requested. In this state we will record all possible traces related to this query: both on the coordinator and on any replica involved. - "Log slow query": when slow query logging is enabled. If slow query logging is enabled and a session's "duration" is above the specified threshold we will create a record in the "slow queries log" and write all trace records created on the coordinator and on a replica if a replica's session lasts longer than that threshold. (We will propagate the Coordinator's slow query logging threshold to replicas in the context of a specific tracing/logging session). The properties above are independent, namely they may be enabled and/or disabled independently and any combination of them is legal (naturally, creating a tracing session when both states above are disabled makes no sense). - Instrument the tracing::tracing service to allow the following: - Enable/disable slow query logging. - Set/get the slow query duration threshold (in microseconds). - Set/get the slow query log record TTL value (in seconds). - Instrument the trace_keyspace_helper to write a slow query log entry when requested. - The slow query logging is disabled by default and the threshold is set to half a second. - The TTL of a slow log record is set to 86400 seconds by default. - It makes sense to use the same "slow query logging threshold" and a "slow query record TTL" both on a coordinator and on a replica Nodes in a context of the same tracing session: - Pass both TTL and a threshold to the replica in a trace_info. This patch also implements the new slow query logging specific logic: - Don't write the pending tracing records before the end of a tracing session until "duration" reaches the logging threshold. - Don't build the parameters<sstring, sstring> map unless we know we will write it to I/O. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-28 18:28:44 +03:00
Vlad Zolotarov	8609900621	tracing: introduce trace_state capabilities bit field - Instead of keeping separate booleans introduce a trace_state_props_set enum_set and pass it around instead of separate booleans. - Change the trace_info to hold this value in addition to write_on_close. Initialize a corresponding bit in an enum_set based on a write_on_close value in a trace_info constructor for a backward compatibility. - Separate a trace_state constructor into two: - For a primary session object. - For a secondary session object. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-23 18:34:36 +03:00
Vlad Zolotarov	ed21398ce9	trace_keyspace_helper: create a system_traces.node_slow_log table This table is going to be used to store information about queries which are slower than a specified threshold. Also added a column caching and mutation creation functions Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-23 17:58:42 +03:00
Vlad Zolotarov	6c3e1935b0	tracing::session_record: change a type of a "ttl" field to be std::chrono::seconds TTL is always defined in seconds - make its type explicitly reflect that. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-23 17:58:42 +03:00
Vlad Zolotarov	372da7e71b	tracing: add support for setting a username and a table name parameters - "username" is a name used in the authentication process. - "table name" is a <keyspace>.<cf name> string representing a name of a table used for a query in question. Note that there may be more than one table name in a batch query. Therefore we store an unordered set of tables names. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-23 17:58:42 +03:00
Vlad Zolotarov	eaf5db66a8	tracing::session_record: store "parameters" data in an std::map instead of in an unordered_map Avoid sorting (and creating a new one) container at a backend code when a sorted container is needed. The overhead for the backends where it's not needed is minimal since the size of the map is very small. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-23 17:58:10 +03:00
Vlad Zolotarov	37da6f53f8	tracing: fix a session "duration" semantics A session's "duration" should be a time it took to handle a request, which is a time till response to a user. In other words - till a consistency level is reached. Before this patch is was a time that takes a complete handling of a request, which is the time it takes to handle all replicas and not only those required to reach a CL. This patch fixes this situation by extending the trace_state's state values to 3 states: inactive, foreground and background. A primary session may be in 3 states: - "inactive": between the creation and a begin() call. - "foreground": after a begin() call and before a stop_foreground_and_write() call. - "background": after a stop_foreground_and_write() call and till the state object is destroyed. - Traces are not allowed while state is in an "inactive" state. - The time the primary session was in a "foreground" state is the time reported as a session's "duration". - Traces that have arrived during the "background" state will be recorded as usual but their "elapsed" time will be greater or equal to the session's "duration". Secondary sessions may only be in an "inactive" or in a "foreground" states. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-16 12:32:34 +03:00
Vlad Zolotarov	f83e33fc13	tracing: make "elapsed" be std::chrono::duration - Define an tracing::elapsed_clock type (std::chrono::steady_clock). Use it instead of trace_state::clock_type. - Store the "elapsed" information in a form of elapsed_clock::duration. - Make all keyspace_backend specific conversions inside the trace_keyspace_helper class, where they belong. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-16 12:32:34 +03:00
Vlad Zolotarov	ebf13da9c9	tracing::session_record: make start_at to be a time_point Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-16 12:32:34 +03:00
Vlad Zolotarov	5deec0e327	tracing::write_complete(): improve a message in case of a logic error Improve a message if there is a logic error and add logging of such errors. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-09 19:00:43 +03:00
Vlad Zolotarov	67d537ecb5	tracing: issue a write event if a single session creates a lot of events Currently write events are issued every time a trace session is closed. However if a single session creates a lot of events we will start dropping them after the total amount of pending records bypasses the limit. This patch will issue a write event before the session end in that case. Since now new events may be added to the active tracing session while it's scheduled for write we have to ensure the following: - Not to add the already pending for write session to the pending bulk. - Grab all pending data in a specific session in a synchronous way during the write event. - Serialize creation of events mutations - otherwise the "monotonic nanos" logic won't work. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-09 19:00:43 +03:00
Vlad Zolotarov	5391bcc5a9	tracing: improve a back pressure policy Use a per-shard tracing records budget instead of maintaining a fixed-size per-session records budget and a per-shard sessions budget. The original policy could lead to some irrational situations, when we have a single tracing session that creates a substantial amount of records that we can handle but we would start dropping new records after it surpasses the per-session limit. The new policy handles a per-shard trace records budget that is being consumed by each trace() call and by a primary session destructor when a session record is created. Each active record may only be in one of the following states: - cached: stored in its session's object. When record is in this state it's not going to be written to I/O during the next write event. - pending for write: when record is in this state it's going to be written to I/O during the next write event. - flushing: the record is being currently written to the I/O. There are counters of the total amount of records in each state above. Each record may only be in a specific state at every point of time and thereby it must be accounted only in one and only one of the three counters. The sum of all three counters should not be greater than (max_pending_trace_records + write_event_records_threshold) at any time (actually it can get as high as a value above plus (max_pending_sessions) if all sessions are primary but we won't take this into an account for simplicity). The same is about the number of outstanding sessions: it may not be greater than (max_pending_sessions + write_event_sessions_threshold) at any time. If total number of tracing records is greater or equal to the limit above, the new trace point is going to be dropped. If current number or records plus the expected number of trace records per session (exp_trace_events_per_session) is greater than the limit above new sessions will be dropped. A new session will also be dropped if there are too many active sessions. When the record or a session is dropped the appropriate statistics counters are updated and there is a rate-limited warning message printed to the log. Every time a number of records pending for write is greater or equal to (write_event_records_threshold) or a number of sessions pending for write is greater or equal to (write_event_sessions_threshold) a write event is issued. Every 2 seconds a timer would write all pending for write records available so far. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-09 19:00:43 +03:00
Vlad Zolotarov	d8fe5317d1	tracing::trace_keyspace_helper: make events' mutations applying loop interruptible When building events' mutation don't apply them in a tight loop but rather apply each of them in a separate continuation to allow reactor to interrupt this loop if it takes too long for it to complete (e.g. where there are a lot of mutations to apply). Since building all events' mutations is asynchronous now we can no longer keep the "nanos" state in a global trace_keyspace_helper object but rather have to move it into the per-session backend_session_state class. backend_session_state class is a backend-specific implementation of a tracing::backend_session_state_base class. An instance of the above object is created by a tracing::i_tracing_backend_helper::allocate_session_state() virtual method and is stored in a tracing::one_session_records object. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-09 19:00:39 +03:00
Vlad Zolotarov	63a0502ed1	tracing: rework the interface between the tracing/trace_state and the backend Before this patch the interaction between the layers above was as follows: - trace_state was passing the trace event data to a backend object every time trace() method was called. - trace_state was passing the session data to a backend object in a destructor. - A backend object was storing this data in a form of lambda where all data above was caught in a capture list. This was primarily done in order to delay the call for make_xxx_mutation(). Lambdas were stored in a map by a session ID and they were executed when a kick() method was called. - A tracing::tracing object was periodically calling a kick() method of a backend that was initiating a write of all pending data to the storage. All backend methods used in the described above interactions were virtual. Thereby, for instance, for each and every trace record we were calling a virtual method that was receiving a significant amount of parameters, store a lambda in a map and return. This is clearly a suboptimal way of using virtual functions since we prevent a compiler from inlining an obviously inlinable operations. This patch changes the interaction scheme to be as follows: - Trace events and session data are stored and passed around in a form of structs that hold all relevant information (no more lambdas). - As long as a trace session is active its data is aggregated inside the corresponding trace_state object. - The object containing all records is passed and stored as a lw_shared_ptr to save extra copies and to shorten capture lists. - All aggregated data is passed to a tracing::tracing object in a trace_state destructor. The data is stored in a std::deque in a tracing::tracing object (instead of a map by a session ID). - A single backend's virtual method call writes all data aggregated so far (kick() method is not needed any more), every time a write event occurs. - Backend has only one virtual method now: - Write a bulk of sessions' data aggregated so far. - Backend's virtual method receives a records bulk object by reference. As a result: - A latency of a single trace event that has no formatting improved from 0.2us to 0.1us. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-08-09 15:25:52 +03:00
Vlad Zolotarov	9c0a725c56	tracing: add a _local_tracing to a i_tracing_backend_helper A backend helper has to constantly communicate with the corresponding tracing::tracing instance. By saving a reference to the tracing::tracing instance will save us a lot of tracing::get_local_tracing_instance() calls and thus a lot of dereferencing. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-07-19 18:21:58 +03:00
Vlad Zolotarov	2bb054748e	tracing: record events' time stamps - Extend the i_tracing_backend_helper interface to accept the event record timestamp. - Grab the current timestamp when the event record is taken. - Add the instrumentation to the trace_keyspace_helper to create a unique time-UUID from a given std::chrono::duration object. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-07-19 18:21:58 +03:00
Vlad Zolotarov	da4836becc	tracing::trace_state: add support for a formatted message in trace() Add an support for passing a format string plus positional parameters for creation of a trace point message. Format string should be given in a fmt library native format described here: http://fmtlib.net/latest/syntax.html#syntax . Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-07-19 18:21:57 +03:00
Vlad Zolotarov	ee0e986e96	tracing: make a service shutdown stages more strict kick() backend during shutdown and restrict accessing a backend after that. Flush pending records when service is being shut down. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-07-19 18:21:57 +03:00
Vlad Zolotarov	6e38133f82	tracing: prevent a destruction of a tracing::tracing while it's used Prevent the destruction of tracing::tracing instances while there are still tracing::trace_state objects that are using it: - Make tracing::tracing inherit from seastar::async_sharded_service<tracing::tracing>. - Grab a tracing::tracing.shared_from_this() in each tracing::trace_state object using it. - Use a saved pointer to the local tracing::tracing instance in a destructor instead of accessing it via tracing::get_local_tracing_instance() to avoid "local is not initialized" assert when sessions are being destroyed after the service was stopped. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-07-19 18:21:57 +03:00
Vlad Zolotarov	a5022a09a4	tracing: use 'write' instead of 'flush' and 'store' for consistency with seastar's API In names of functions and variables: s/flush_/write_/ s/store_/write_/ In a i_tracing_backend_helper: s/flush()/kick()/ Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-07-19 18:21:57 +03:00
Vlad Zolotarov	d3960f0bbb	tracing: rearrange shut down tracing::tracing local instance is dereferenced from a cql_server::connection::process_request(), therefore tracing::tracing service may be stop()ed only after a CQL server service is down. On the other hand it may not be stopped before RPC service is down because a remote side may request a tracing for a specific command too. This patch splits the tracing::tracing stop() into two phases: 1) Flush all pending tracing records and stop the backend. 2) Stop the service. The first phase is called after CQL server is down and before RPC is down. The second phase is called after RPC is down. Fixes #1339 Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com> Message-Id: <1465840496-19990-1-git-send-email-vladz@cloudius-systems.com>	2016-06-14 07:58:04 +03:00
Vlad Zolotarov	ce08bc611c	tracing: fix debug compilation Define flush_period as a const and not as constexpr. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com> Message-Id: <1465240516-20128-1-git-send-email-vladz@cloudius-systems.com>	2016-06-06 22:15:27 -04:00
Vlad Zolotarov	905190ac06	tracing: add support for probabilistic tracing Add a support for defining a probability (a value in a [0,1] range) for tracing the next CQL request. Traces for requests that are chosen to be traced due to this feature are not going to flushed immediately. Use std::subtract_with_carry_engine (implements the "lagged Fibonacci" algorithm) random number engine for fastest generation of random integer values. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-06-06 15:41:01 +03:00
Vlad Zolotarov	779ff88c76	tracing: add flush timer Flush pending sessions to the storage every 2 seconds. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-06-06 14:34:08 +03:00
Vlad Zolotarov	4b008ac5ea	tracing: rework maximum sessions amount back pressure strategy A tracing session life cycle includes 3 stages: 1) Active: when new trace records are being added to this session. 2) Pending for flushing to a storage: when session is over but not yet flushed to the storage ("backend"). 3) Flushing: when session's records are being flushed to the storage and this process is not yet completed. Sessions may accumulate in each of the stages above and we should limit the maximum amount of sessions being accumulated in each of them in order to avoid OOM situation. Current in-tree implementation only limits the number of tracing sessions accumulated in the first ("Active") stage. Since currently every closing session is being immediately flushed (as long as "settraceprobability" is not implemented) the second stage never accumulates tracing sessions. The third stage is currently not controlled at all and if, for instance, we succeed to push enough tracing session towards a slow storage backend, they may accumulate there consuming an uncontrolled amount of memory and may eventually consume all of it. This patch fixes this unpleasant situation by implying the following strategy: - Limit the total amount of accumulated tracing sessions in all stages above together by a static value - 2 times "flush threshold". "2 times" is needed to allow new tracing sessions to accumulate in the stage 2 while sessions in the stage 3 are still being processed. - Forcefully flush sessions in the stage 2 to the storage when their count reaches a "flush threshold". This would ensure that there will not more than totally (2 * "flush threshold") sessions (in any stage) on each shard. An advantage of this strategy is its simplicity - we only need a single threshold to control all stages. If we feel that we needed a finer graining for each stage we may add separate limits for each of them in the future. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-06-06 13:50:41 +03:00
Vlad Zolotarov	a53d329b25	tracing: add a serializable trace_info object tracing::trace_info is used to pass the tracing information between nodes. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-06-01 20:16:53 +03:00
Vlad Zolotarov	c965528a03	tracing: add a trace_state and tracing classes trace_state: Is a single tracing session. tracing: A sharded service that contains an i_trace_backend_helper instance and is a "factory" of trace_state objects. trace_state main interface functions are: - begin(): Start time counting (should be used via tracing::begin() wrapper). - trace(): Create a tracing event - it's coupled with a time passed since begin() (should be used via tracing::trace() wrapper). - ~trace_state(): Destructor will close the tracing session. "tracing" service main interface function is: - start(): Initialize a backend. - stop(): Shut down a backend. - create_session(): Creates a new tracing session. (tracing::end_session(): Is called by a trace_state destructor). When trace_state needs to store a tracing event it uses a backend helper from a "tracing" service. A "tracing" service limits a number of opened tracing session by a static number. If this number is reached - next sessions will be dropped. trace_state implements a similar strategy in regard to tracing events per singe session. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-06-01 20:13:42 +03:00
Vlad Zolotarov	a2994ffd7f	tracing: add i_tracing_backend_helper interface This class represents an interface for a specific backend that is going to store tracing information. The specific implementation may and expected to implement caching of pending tracing records. Interface functions are: - start(): Initialize a backend (e.g. create keyspace and tables). - stop(): Flush all pending work and shut down the backend. - store_session_record()/store_event_record(): Cache/store the corresponding tracing records. - flush(): Flush pending tracing records. Signed-off-by: Vlad Zolotarov <vladz@cloudius-systems.com>	2016-06-01 20:12:13 +03:00

50 Commits