mariadb-columnstore-engine

mirror of https://github.com/mariadb-corporation/mariadb-columnstore-engine.git synced 2025-08-01 06:46:55 +03:00

Author	SHA1	Message	Date
Leonid Fedorov	04752ec546	clang format apply	2022-01-21 16:43:49 +00:00
Leonid Fedorov	01f3ceb437	replace header guards with #pragma once	2022-01-21 15:24:58 +00:00
Gagan Goel	195425924d	MCOL-4936 Disable binlog for DML statements. DML statements executed on the primary node in a ColumnStore cluster do not need to be written to the primary's binlog. This is due to ColumnStore's distributed storage architecture. With this patch, we disable writing to binlog when a DML statement (INSERT/DELETE/UPDATE/LDI/INSERT..SELECT) is performed on a ColumnStore table. HANDLER::external_lock() calls are used to 1. Turn OFF the OPTION_BIN_LOG flag 2. Turn ON the OPTION_BIN_TMP_LOG_OFF flag in THD::variables.option_bits during a WRITE lock call. THD::variables.option_bits is restored back to the original state during the UNLOCK call in HANDLER::external_lock(). Further, isDMLStatement() function is added to reduce code verbosity to check if a given statement is a DML statement. Note that with this patch, not writing to primary's binlog means DML replication from a ColumnStore cluster to another ColumnStore cluster or to another foreign engine will not work.	2022-01-04 17:31:59 +00:00
Roman Nozdrin	b3ab3fb514	Merge pull request #2203 from mariadb-AlexeyAntipovsky/auto-query-stats [MCOL-4944] Automatically enable stats collection	2021-12-23 15:13:43 +03:00
Roman Nozdrin	94806e7ee0	Merge pull request #2200 from drrtuy/MCOL-4943-dev MCOL-4943 Moved SQL script call into columnstore-post-install	2021-12-21 11:18:07 +03:00
Alexey Antipovsky	683a6b3d19	[MCOL-4944] Automatically enable stats collection if it is enabled in the config	2021-12-21 11:15:25 +03:00
Roman Nozdrin	22b0e4addc	MCOL-4943 Moved SQL script call into columnstore-post-install	2021-12-18 03:59:58 +00:00
Roman Nozdrin	7b5845a4aa	MCOL-4871 Bar's patch to do proper extent elimination for short CHAR	2021-12-17 17:41:03 +00:00
Gagan Goel	7f456e58cc	MCOL-4868 UPDATE on a ColumnStore table containing an IN-subquery on a non-ColumnStore table does not work. As part of MCOL-4617, we moved the in-to-exists predicate creation and injection from the server into the engine. However, when query with an IN Subquery contains a non-ColumnStore table, the server still performs the in-to-exists predicate transformation for the foreign engine table. This caused ColumnStore's execution plan to contain incorrect WHERE predicates. As a fix, we call mutate_optimizer_flags() for the WRITE lock, in addition to the READ table lock. And in mutate_optimizer_flags(), we change the optimizer flag from OPTIMIZER_SWITCH_IN_TO_EXISTS to OPTIMIZER_SWITCH_MATERIALIZATION.	2021-12-16 23:11:26 +00:00
Gagan Goel	d91cab2ff5	MCOL-4925 Suppress the warning message when a non-cached table is (#2164 ) dropped with the insert cache enabled.	2021-12-06 11:27:37 -06:00
Gagan Goel	affb2ae770	MCOL-4769 Fix cache bugs. (#2151 ) * MCOL-4769 Do not replay INSERTs and LDIs on the replica nodes when the write cache is enabled. * MCOL-4769 If a table is created with the write cache disabled (i.e. when columnstore_cache_inserts=OFF), make it accessible when the cache feature is enabled (columnstore_cache_inserts=ON).	2021-11-22 14:20:50 -06:00
Alexander Barkov	fa9f18553a	MCOL-4728 Query with unusual use of aggregate functions on ColumnStore table crashes MariaDB Server After an AggreateColumn corresponding to SUM(1+1) is created, it is pushed to the list: gwi.count_asterisk_list.push_back(ac) Later, in getSelectPlan(), the expression SUM(1+1) was erroneously treated as a constant: if (!hasNonSupportItem && !nonConstFunc(ifp) && !(parseInfo & AF_BIT) && tmpVec.size() == 0) { srcp.reset(buildReturnedColumn(item, gwi, gwi.fatalParseError)); This code freed the original AggregateColumn and replaced to a ConstantColumn. But gwi.count_asterisk_list still pointer to the freed AggregateColumn(). The expression SUM(1+1) was treated as a constant because tmpVec was empty due to a bug in this code: // special handling for count(). This should not be treated as constant. if (isp->argument_count() == 1 && ( sfitempp[0]->type() == Item::CONST_ITEM && (sfitempp[0]->cmp_type() == INT_RESULT \|\| sfitempp[0]->cmp_type() == STRING_RESULT \|\| sfitempp[0]->cmp_type() == REAL_RESULT \|\| sfitempp[0]->cmp_type() == DECIMAL_RESULT) ) ) { field_vec.push_back((Item_field)item); //dummy Notice, it handles only aggregate functions with explicit literals passed as an argument, while it does not handle constant expressions such as 1+1. Fix: - Adding new classes ConstantColumnNull, ConstantColumnString, ConstantColumnNum, ConstantColumnUInt, ConstantColumnSInt, ConstantColumnReal, ValStrStdString, to reuse the code easier. - Moving a part of the code from the case branch handling CONST_ITEM in buildReturnedColumn() into a new function newConstantColumnNotNullUsingValNativeNoTz(). This makes the code easier to read and to reuse in the future. - Adding a new function newConstantColumnMaybeNullFromValStrNoTz(). Removing dulplicate code from !!!four!!! places, using the new function instead. - Adding a function isSupportedAggregateWithOneConstArg() to properly catch all constant expressions. Using the new function parse_item() in the code commented as "special handling for count(*)". Now it pushes all constant expressions to field_vec, not only explicit literals. - Moving a part of the code from buildAggregateColumn() to a helper function processAggregateColumnConstArg(). Using processAggregateColumnConstArg() in the CONST_ITEM and NULL_ITEM branches. - Adding a new branch in buildReturnedColumn() handling FUNC_ITEM. If a function has constant arguments, a ConstantColumn() is immediately created, without going to buildArithmeticColumn()/buildFunctionColumn(). - Reusing isSupportedAggregateWithOneConstArg() and processAggregateColumnConstArg() in buildAggregateColumn(). A new branch catches aggregate function has only one constant argument and immediately creates a single ConstantColumn without traversing to the argument sub-components.	2021-09-21 14:00:56 +04:00
Leonid Fedorov	5c5f103f98	MCOL-4839: Fix clang build (#2100 ) * Fix clang build * Extern C returned to plugin_instance Co-authored-by: Leonid Fedorov <l.fedorov@mail.corp.ru>	2021-08-23 10:45:10 -05:00
benthompson15	923bbf4033	MCOL-1356: Add convert_tz (#2099 )	2021-08-19 17:47:10 -05:00
Gagan Goel	98473a45cc	Merge pull request #2079 from dhall-MariaDB/MCOL-3741 Mcol 3741 Change IDB-xxxx error codes to MCS-xxxx	2021-08-18 14:01:04 -04:00
Leonid Fedorov	469e5c7881	WriteBatchFieldMariaDB m_type was wrong (#2090 )	2021-08-18 11:36:53 -05:00
David Hall	ecde2719b1	MCOL-3741 Change IDB-xxxx error codes to MCS-xxxx	2021-08-09 11:33:09 -05:00
Gagan Goel	649ca10429	MCOL-4805 Follow up.	2021-08-04 23:54:02 +00:00
Gagan Goel	afb638b9bd	MCOL-4805 For functions in the plugin code that disable replication on the slave threads, we now check for this condition early on in the function block.	2021-08-03 22:49:22 +00:00
Gagan Goel	c5502c02fa	Rename columnstore_use_cpimport_for_cache_inserts system variable to (#2053 ) columnstore_cache_use_import.	2021-07-19 12:47:15 -05:00
Denis Khalikov	fa8dc815a7	MCOL-4814 Add a cmake build option to enable LZ4 compression. This patch adds an option for cmake flags to enable lz4 compression.	2021-07-16 17:57:11 +03:00
benthompson15	91945fe271	Fix warnings for vla, unused variables.	2021-07-14 20:08:46 -05:00
Denis Khalikov	dc51dbf6cf	[MCOL-4786] Fix filter comparison. Compare ParseTree by dereferencing pointers.	2021-07-12 19:18:02 +03:00
Denis Khalikov	adace6e0c7	MCOL-4786 Fix wrong comparison for the filters. Fix wrong comparison for the filters while creating case function.	2021-07-09 12:18:26 +03:00
Roman Nozdrin	3391eda89d	Merge pull request #2038 from mariadb-corporation/MCOL-4603-replace-long-double Replace LONG DOUBLE with wide decimal for aggregates	2021-07-08 22:14:17 +03:00
Leonid Fedorov	f81f743282	Replace underlying type for avg and sum for int types from long double to wide decimal	2021-07-08 17:04:43 +00:00
Gagan Goel	a0bd790005	ColumnStore Cache changes. 1. Add a new system variable, columnstore_use_cpimport_for_cache_inserts, that when set to ON, uses cpimport for the cache flush into ColumnStore. This variable is set to OFF by default. By default, we perform batch inserts for the cache flush. 2. Disable DMLProc logging of the SQL statement text for the cache flush operation in case of batch inserts. Under certain heavy loads involving INSERT statements, this logging becomes a bottleneck for the cache flush, causing subsequent inserts into the cache table to hang.	2021-07-07 19:02:28 +00:00
Roman Nozdrin	866dc25729	Merge pull request #1842 from denis0x0D/MCOL-987_LZ MCOL-987 LZ4 compression support.	2021-07-07 13:13:18 +03:00
Roman Nozdrin	fb5ba84212	MCOL-4802 Removed ByteStream methods for bool manipulations and add some logging into I_S.columnstore_files	2021-07-07 07:16:30 +00:00
Denis Khalikov	cc1c3629c5	MCOL-987 Add LZ4 compression. * Adds CompressInterfaceLZ4 which uses LZ4 API for compress/uncompress. * Adds CMake machinery to search LZ4 on running host. * All methods which use static data and do not modify any internal data - become `static`, so we can use them without creation of the specific object. This is possible, because the header specification has not been modified. We still use 2 sections in header, first one with file meta data, the second one with pointers for compressed chunks. * Methods `compress`, `uncompress`, `maxCompressedSize`, `getUncompressedSize` - become pure virtual, so we can override them for the other compression algos. * Adds method `getChunkMagicNumber`, so we can verify chunk magic number for each compression algo. * Renames "s/IDBCompressInterface/CompressInterface/g" according to requirement.	2021-07-06 18:04:37 +03:00
Gagan Goel	8520f87237	MCOL-641 Cleanup.	2021-07-06 09:01:49 +00:00
David.Hall	237cad347f	MCOL-4758 Limit LONGTEXT and LONGBLOB to 16MB (#1995 ) MCOL-4758 Limit LONGTEXT and LONGBLOB to 16MB Also add the original test case from MCOL-3879.	2021-07-05 02:09:41 -04:00
Roman Nozdrin	a4aecc120e	Merge pull request #2006 from tntnatbry/fix-const-scalar-subselect Fixes for queries containing constant scalar subselects in the WHERE clause.	2021-07-02 20:15:09 +03:00
Gagan Goel	8d0ca55495	Fixes for queries containing constant scalar subselects in the WHERE clause. For queries of the form: SELECT col1 FROM t1 WHERE col2 = (SELECT 2); We fix the execution plan which earlier had an empty filters expression. For this query, we now build a SimpleFilter with a SimpleColumn and a ConstantColumn as the LHS and the RHS operands respectively. For queries of the form: SELECT ... WHERE col1 NOT IN (SELECT <const_item>); The execution plan earlier built a SimpleFilter with an "=" as the predicate operator of the filter. We fix this by assigning the correct "<>" operator instead.	2021-07-02 16:40:30 +00:00
Roman Nozdrin	6dc356ed60	Merge pull request #1989 from denis0x0D/MCOL-4713 MCOL-4713 Analyze table implementation.	2021-07-02 16:17:07 +03:00
Denis Khalikov	c20015a7b2	MCOL-4713 Analyze table implementation.	2021-07-02 12:37:12 +03:00
Roman Nozdrin	a465b60bdd	MCOL-1482 Future repetition reduction	2021-07-01 12:27:03 +00:00
Roman Nozdrin	325bb6c9e0	Merge pull request #1986 from tntnatbry/MCOL-1482 MCOL-1482 An UPDATE operation on a non-ColumnStore table involving a cross-engine join	2021-07-01 14:25:32 +03:00
David.Hall	132146b9c8	Mcol 3738 Allow COUNT(DISTINCT to have multiple parms) (#2002 ) * MCOL-3738 allow COUNT(DISTINCT) multiple parameters Changes in the way tupleaggregatestep sets up the aggregate arrays. * MCOL-3738 mtr test	2021-06-28 20:14:44 +03:00
Gagan Goel	49255f5cbd	MCOL-1482 An UPDATE operation on a non-ColumnStore table involving a cross-engine join with a ColumnStore table errors out. ColumnStore cannot directly update a foreign table. We detect whether a multi-table UPDATE operation is performed on a foreign table, if so, do not create the select_handler and let the server execute the UPDATE operation instead.	2021-06-25 15:27:54 +00:00
David.Hall	28fd12a008	Merge pull request #1874 from mariadb-corporation/bar-develop-MCOL-4681 MCOL-4681 Fix install_mcs_mysql.sh.in to do CREATE FUNCTION instead o…	2021-06-24 09:11:06 -05:00
Gagan Goel	7c8b502dc2	Fix regression in a query involving an aggregate function on a non-wide decimal column in the HAVING clause. In buildAggregateColumn(), if an aggregate function (such as avg) is applied on a non-wide decimal column, we were setting the precision of the resulting column as -1. This later down in the execution got converted to 255 as in some cases, precision is stored as uint8_t. The predicate operations on a DECIMAL column has logic that uses the wide Decimal::s128value field if precision > 18. This logic incorrectly used the Decimal::s128value instead of the correct value stored in the narrow Decimal::value field, since precision of the Decimal column was 255. The fix is to set the aggregate column precision to datatypes::INT64MAXPRECISION (18) in buildAggregateColumn() when the aggregate is applied on a non-wide decimal column. This commit also partially fixes -Wstrict-aliasing GCC warnings.	2021-06-22 11:11:34 +00:00
Roman Nozdrin	e153486361	Merge pull request #1994 from denis0x0D/MCOL-4685_rename MCOL-4685 Remname UNUSED -> SNAPPY	2021-06-16 11:24:17 +03:00
Denis Khalikov	e2a5956ef8	MCOL-4685 Remname UNUSED -> SNAPPY	2021-06-15 21:19:09 +03:00
Roman Nozdrin	96f2a55eea	Merge pull request #1970 from tntnatbry/MCOL-4525 MCOL-4525 Implement columnstore_select_handler=AUTO.	2021-06-14 10:43:34 +03:00
Gagan Goel	e3d8100150	MCOL-4525 Implement columnstore_select_handler=AUTO. This feature allows a query execution to fallback to the server, in case query execution using the select_handler (SH) fails. In case of fallback, a warning message containing the original reason for query failure using SH is generated. To accomplish this task, SH execution is moved to an earlier step when we create the SH in create_columnstore_select_handler(), instead of the previous call to SH execution in ha_columnstore_select_handler::init_scan(). This requires some pre-requisite steps that occur in the server in JOIN::optimize() and JOIN::exec() to be performed before starting SH execution. In addition, missing test cases from MCOL-424 are also added to the MTR suite, and the corresponding fix using disable_indices_for_CEJ() is reverted back since the original fix now appears to be redundant.	2021-06-11 11:35:34 +00:00
Alexander Barkov	d00ace2398	MCOL-4757 Empty set in SELECT * INFORMATION_SCHEMA.COLUMNSTORE_TABLES WHERE TABLE_NAME='t1'	2021-06-11 12:00:23 +04:00
Roman Nozdrin	47e9fc0312	Merge pull request #1922 from denis0x0D/MCOL-4685 MCOL-4685: Eliminate some irrelevant settings (uncompressed data and extents per file)	2021-06-06 16:00:29 +03:00
Gagan Goel	3537c0d635	Merge pull request #1962 from tntnatbry/MCOL-4642 MCOL-4642 NOT IN subquery containing an isnull in the OR predicate crashes server.	2021-06-04 07:18:46 -04:00
Denis Khalikov	606194e6e4	MCOL-4685: Eliminate some irrelevant settings (uncompressed data and extents per file). This patch: 1. Removes the option to declare uncompressed columns (set columnstore_compression_type = 0). 2. Ignores [COMMENT '[compression=0] option at table or column level (no error messages, just disregard). 3. Removes the option to set more than 2 extents per file (ExtentsPreSegmentFile). 4. Updates rebuildEM tool to support up to 10 dictionary extent per dictionary segment file. 5. Adds check for `DBRootStorageType` for rebuildEM tool. 6. Renamed rebuildEM to mcsRebuildEM.	2021-06-03 14:44:33 +03:00

1 2 3 4 5 ...

1028 Commits