mariadb-columnstore-engine

mirror of https://github.com/mariadb-corporation/mariadb-columnstore-engine.git synced 2025-07-29 08:21:15 +03:00

Author	SHA1	Message	Date
Gagan Goel	7f456e58cc	MCOL-4868 UPDATE on a ColumnStore table containing an IN-subquery on a non-ColumnStore table does not work. As part of MCOL-4617, we moved the in-to-exists predicate creation and injection from the server into the engine. However, when query with an IN Subquery contains a non-ColumnStore table, the server still performs the in-to-exists predicate transformation for the foreign engine table. This caused ColumnStore's execution plan to contain incorrect WHERE predicates. As a fix, we call mutate_optimizer_flags() for the WRITE lock, in addition to the READ table lock. And in mutate_optimizer_flags(), we change the optimizer flag from OPTIMIZER_SWITCH_IN_TO_EXISTS to OPTIMIZER_SWITCH_MATERIALIZATION.	2021-12-16 23:11:26 +00:00
Roman Nozdrin	af36f9940f	This patch introduces support for scanning/filtering vectorized execution for numeric-based data types TEXT, CHAR, VARCHAR, FLOAT and DOUBLE are not yet supported by vectorized path This patch introduces an example for Google benchmarking suite to measure a perf diff b/w legacy scan/filtering code and the templated version	2021-12-10 10:30:00 +00:00
Gagan Goel	340a90fc8d	MCOL-4874 Crossengine JOIN involving a ColumnStore table and a wide decimal column in a non-ColumnStore table throws an exception. ROW::getSignedNullValue() method does not support wide decimal fields yet. To fix this exception, we remove the call to this method from CrossEngineStep::setField().	2021-12-08 22:26:52 +00:00
Gagan Goel	d91cab2ff5	MCOL-4925 Suppress the warning message when a non-cached table is (#2164 ) dropped with the insert cache enabled.	2021-12-06 11:27:37 -06:00
Roman Nozdrin	6f1c0d6c5a	Merge pull request #2157 from denis0x0D/warnings [MCOL-4849] Fix build warnings.	2021-11-25 12:14:07 -06:00
Gagan Goel	affb2ae770	MCOL-4769 Fix cache bugs. (#2151 ) * MCOL-4769 Do not replay INSERTs and LDIs on the replica nodes when the write cache is enabled. * MCOL-4769 If a table is created with the write cache disabled (i.e. when columnstore_cache_inserts=OFF), make it accessible when the cache feature is enabled (columnstore_cache_inserts=ON).	2021-11-22 14:20:50 -06:00
Denis Khalikov	f8bd566b0f	[MCOL-4849] Fix build warnings.	2021-11-17 17:33:22 +03:00
Denis Khalikov	b382f681a1	[MCOL-4849] Parallelize the processing of the bytestream vector. This patch changes the logic of the `receiveMultiPrimitiveMessages` function in the following way: 1. We have only one aggregation thread which reads the data from Queue (which is populated by messages from BPPs). 2. Processing of the received `bytestream vector` could be in parallel depends on the type of `TupleBPS` operation (join, fe2, ...) and actual thread pool workload. The motivation is to eliminate some amount of context switches.	2021-11-04 13:28:22 +03:00
Leonid Fedorov	6a00fa9839	byson unused function fix	2021-10-29 14:57:11 +00:00
Leonid Fedorov	1973168e03	c++17 fix	2021-10-29 14:57:11 +00:00
Roman Nozdrin	3de038c1da	MCOL-4876 This patch enables continues buffer to be used by ColumnCommand and aligns BPP::blockData that in most cases was unaligned	2021-10-06 09:23:40 +00:00
Alexander Barkov	fa9f18553a	MCOL-4728 Query with unusual use of aggregate functions on ColumnStore table crashes MariaDB Server After an AggreateColumn corresponding to SUM(1+1) is created, it is pushed to the list: gwi.count_asterisk_list.push_back(ac) Later, in getSelectPlan(), the expression SUM(1+1) was erroneously treated as a constant: if (!hasNonSupportItem && !nonConstFunc(ifp) && !(parseInfo & AF_BIT) && tmpVec.size() == 0) { srcp.reset(buildReturnedColumn(item, gwi, gwi.fatalParseError)); This code freed the original AggregateColumn and replaced to a ConstantColumn. But gwi.count_asterisk_list still pointer to the freed AggregateColumn(). The expression SUM(1+1) was treated as a constant because tmpVec was empty due to a bug in this code: // special handling for count(). This should not be treated as constant. if (isp->argument_count() == 1 && ( sfitempp[0]->type() == Item::CONST_ITEM && (sfitempp[0]->cmp_type() == INT_RESULT \|\| sfitempp[0]->cmp_type() == STRING_RESULT \|\| sfitempp[0]->cmp_type() == REAL_RESULT \|\| sfitempp[0]->cmp_type() == DECIMAL_RESULT) ) ) { field_vec.push_back((Item_field)item); //dummy Notice, it handles only aggregate functions with explicit literals passed as an argument, while it does not handle constant expressions such as 1+1. Fix: - Adding new classes ConstantColumnNull, ConstantColumnString, ConstantColumnNum, ConstantColumnUInt, ConstantColumnSInt, ConstantColumnReal, ValStrStdString, to reuse the code easier. - Moving a part of the code from the case branch handling CONST_ITEM in buildReturnedColumn() into a new function newConstantColumnNotNullUsingValNativeNoTz(). This makes the code easier to read and to reuse in the future. - Adding a new function newConstantColumnMaybeNullFromValStrNoTz(). Removing dulplicate code from !!!four!!! places, using the new function instead. - Adding a function isSupportedAggregateWithOneConstArg() to properly catch all constant expressions. Using the new function parse_item() in the code commented as "special handling for count(*)". Now it pushes all constant expressions to field_vec, not only explicit literals. - Moving a part of the code from buildAggregateColumn() to a helper function processAggregateColumnConstArg(). Using processAggregateColumnConstArg() in the CONST_ITEM and NULL_ITEM branches. - Adding a new branch in buildReturnedColumn() handling FUNC_ITEM. If a function has constant arguments, a ConstantColumn() is immediately created, without going to buildArithmeticColumn()/buildFunctionColumn(). - Reusing isSupportedAggregateWithOneConstArg() and processAggregateColumnConstArg() in buildAggregateColumn(). A new branch catches aggregate function has only one constant argument and immediately creates a single ConstantColumn without traversing to the argument sub-components.	2021-09-21 14:00:56 +04:00
Roman Nozdrin	67c85dae15	MCOL-4809 The patch replaces legacy scanning/filtering code with a number of templates that simplifies control flow removing needless expressions	2021-09-06 17:04:52 +00:00
Leonid Fedorov	5c5f103f98	MCOL-4839: Fix clang build (#2100 ) * Fix clang build * Extern C returned to plugin_instance Co-authored-by: Leonid Fedorov <l.fedorov@mail.corp.ru>	2021-08-23 10:45:10 -05:00
benthompson15	923bbf4033	MCOL-1356: Add convert_tz (#2099 )	2021-08-19 17:47:10 -05:00
Gagan Goel	98473a45cc	Merge pull request #2079 from dhall-MariaDB/MCOL-3741 Mcol 3741 Change IDB-xxxx error codes to MCS-xxxx	2021-08-18 14:01:04 -04:00
Leonid Fedorov	dcacbbd3a9	Wrong power of 2 in esimator` (#2088 )	2021-08-18 11:48:48 -05:00
Leonid Fedorov	469e5c7881	WriteBatchFieldMariaDB m_type was wrong (#2090 )	2021-08-18 11:36:53 -05:00
Leonid Fedorov	145a7f7217	Wrong concatenation in between predicate (#2092 )	2021-08-18 11:32:01 -05:00
David Hall	ecde2719b1	MCOL-3741 Change IDB-xxxx error codes to MCS-xxxx	2021-08-09 11:33:09 -05:00
Gagan Goel	649ca10429	MCOL-4805 Follow up.	2021-08-04 23:54:02 +00:00
Gagan Goel	afb638b9bd	MCOL-4805 For functions in the plugin code that disable replication on the slave threads, we now check for this condition early on in the function block.	2021-08-03 22:49:22 +00:00
David Hall	a202bda485	MCOL-4719 iterate into subquery looking for windowfunctions When an outer query filter accesses an subquery column that contains an aggregate or a window function, certain optimizations can't be performed. We had been looking at the surface of the returned column. We now iterate into any functions or operations looking for aggregates and window functions.	2021-07-22 13:56:21 -05:00
Roman Nozdrin	4cdef40a55	Merge pull request #2052 from drrtuy/MCOL-4815 MCOL-4815 ColumnCommand was replaced with a set of derived classes sp…	2021-07-21 17:36:35 +03:00
Roman Nozdrin	a292585b8c	MCOL-4815 ColumnCommand was replaced with a set of derived classes specified by column width RTSCommand was modified to use a fabric that produces CC class based on column width NB this patch doesn't affect PseudoCC that also leverages ColumnCommand	2021-07-21 12:54:14 +00:00
Gagan Goel	c5502c02fa	Rename columnstore_use_cpimport_for_cache_inserts system variable to (#2053 ) columnstore_cache_use_import.	2021-07-19 12:47:15 -05:00
Denis Khalikov	fa8dc815a7	MCOL-4814 Add a cmake build option to enable LZ4 compression. This patch adds an option for cmake flags to enable lz4 compression.	2021-07-16 17:57:11 +03:00
benthompson15	91945fe271	Fix warnings for vla, unused variables.	2021-07-14 20:08:46 -05:00
benthompson15	2ae3da45eb	MCOL-1175: add ability to encrypt CEJ password and use in Columnstore.xml (#2045 )	2021-07-13 11:42:36 -05:00
Gagan Goel	b3a560300c	Revert "Merge pull request #2022 from mariadb-corporation/bar-develop-MCOL-4791" This reverts commit `4016e25e5b`, reversing changes made to `85435f6b1e`.	2021-07-13 11:06:56 +00:00
David.Hall	2b37c2c7bc	Merge pull request #2046 from denis0x0D/MCOL-4786_fix_regression [MCOL-4786] Fix filter comparison.	2021-07-12 13:51:42 -05:00
Denis Khalikov	dc51dbf6cf	[MCOL-4786] Fix filter comparison. Compare ParseTree by dereferencing pointers.	2021-07-12 19:18:02 +03:00
Gagan Goel	3d557a2f1e	Merge pull request #2044 from dhall-MariaDB/MCOL-3738 MCOL-3738 COUNT(DISTINCT) with multiple parms	2021-07-12 07:34:56 -04:00
David Hall	76607be63a	MCOL-3738 COUNT(DISTINCT) with multiple parms Fixed regression Added a few more mtr tests	2021-07-09 09:07:03 -05:00
Roman Nozdrin	4d265472ee	Merge pull request #2036 from mariadb-SergeyZefirov/MCOL-4766-UPDATE-INSERT-in-a-transaction-does-not-revert-back-extent-ranges-on-a-rollback MCOL-4766 ROLLBACK kept ranges changed inside rolled back transaction	2021-07-09 16:33:06 +03:00
Denis Khalikov	adace6e0c7	MCOL-4786 Fix wrong comparison for the filters. Fix wrong comparison for the filters while creating case function.	2021-07-09 12:18:26 +03:00
Roman Nozdrin	3391eda89d	Merge pull request #2038 from mariadb-corporation/MCOL-4603-replace-long-double Replace LONG DOUBLE with wide decimal for aggregates	2021-07-08 22:14:17 +03:00
Leonid Fedorov	f81f743282	Replace underlying type for avg and sum for int types from long double to wide decimal	2021-07-08 17:04:43 +00:00
Gagan Goel	a0bd790005	ColumnStore Cache changes. 1. Add a new system variable, columnstore_use_cpimport_for_cache_inserts, that when set to ON, uses cpimport for the cache flush into ColumnStore. This variable is set to OFF by default. By default, we perform batch inserts for the cache flush. 2. Disable DMLProc logging of the SQL statement text for the cache flush operation in case of batch inserts. Under certain heavy loads involving INSERT statements, this logging becomes a bottleneck for the cache flush, causing subsequent inserts into the cache table to hang.	2021-07-07 19:02:28 +00:00
Sergey Zefirov	9e0851e4cf	MCOL-4766 ROLLBACK kept ranges changed inside rolled back transaction Now ROLLBACK drops ranges to INVALID state which makes engine to rescan blocks and discover correct ranges.	2021-07-07 18:16:56 +03:00
Roman Nozdrin	866dc25729	Merge pull request #1842 from denis0x0D/MCOL-987_LZ MCOL-987 LZ4 compression support.	2021-07-07 13:13:18 +03:00
Roman Nozdrin	7b4f759592	Merge pull request #2032 from drrtuy/MCOL-4802 MCOL-4802 Removed ByteStream methods for bool and add some logging in…	2021-07-07 13:03:54 +03:00
Roman Nozdrin	fb5ba84212	MCOL-4802 Removed ByteStream methods for bool manipulations and add some logging into I_S.columnstore_files	2021-07-07 07:16:30 +00:00
Alexander Barkov	9794f24369	MCOL-4801 Replace Row methods getStringLength() and getStringPointer() to getConstString()	2021-07-06 21:15:32 +04:00
Denis Khalikov	cc1c3629c5	MCOL-987 Add LZ4 compression. * Adds CompressInterfaceLZ4 which uses LZ4 API for compress/uncompress. * Adds CMake machinery to search LZ4 on running host. * All methods which use static data and do not modify any internal data - become `static`, so we can use them without creation of the specific object. This is possible, because the header specification has not been modified. We still use 2 sections in header, first one with file meta data, the second one with pointers for compressed chunks. * Methods `compress`, `uncompress`, `maxCompressedSize`, `getUncompressedSize` - become pure virtual, so we can override them for the other compression algos. * Adds method `getChunkMagicNumber`, so we can verify chunk magic number for each compression algo. * Renames "s/IDBCompressInterface/CompressInterface/g" according to requirement.	2021-07-06 18:04:37 +03:00
Roman Nozdrin	b9bd207d3b	Merge pull request #2029 from tntnatbry/MCOL-641-cleanup MCOL-641 Cleanup.	2021-07-06 14:35:21 +03:00
Gagan Goel	8520f87237	MCOL-641 Cleanup.	2021-07-06 09:01:49 +00:00
Leonid Fedorov	3fb5579708	Delete duplicate precision initilization in distinct aggregate prepare	2021-07-05 15:54:58 +03:00
David.Hall	237cad347f	MCOL-4758 Limit LONGTEXT and LONGBLOB to 16MB (#1995 ) MCOL-4758 Limit LONGTEXT and LONGBLOB to 16MB Also add the original test case from MCOL-3879.	2021-07-05 02:09:41 -04:00
Roman Nozdrin	6b823db28b	Merge pull request #1913 from denis0x0D/MCOL-1205 MCOL-1205 Support queries with circular joins	2021-07-02 20:47:14 +03:00

... 3 4 5 6 7 ...

1621 Commits